Laya for benchmarks & evals

Speed, accuracy and cost measurements of Laya.

227 projects

laya-mlx

mizorewww

Runs Laya typed-decision models locally on Apple Silicon using MLX

6.3k·GitHub repo

laya-coreml

mizorewww

Runs Laya typed-decision models locally on Apple Silicon using Core ML and the Neural Engine

1.4k·GitHub repo

openJev-verdict-2.0

Heman10x-NGU

Develops and benchmarks a non-autoregressive typed-decision model against Jev and Laya

285·GitHub repo

laya-playground

wdobry

A local website combines Laya model demos, games, a benchmark, and an agent skill

157·GitHub repo

laya-vs-jev

virajbhartiya

A side-by-side T-Rex game compares local Laya and hosted Jev decision models

99·GitHub repo

laya-mps

afshinm

Runs Laya typed-decision inference locally on Apple Silicon using Metal Performance Shaders

21·GitHub repo

lev

jlt-commons

Implements a typed decision engine using Laya checkpoints and GGUF chat models

17·GitHub repo

edgejev

yzfly

Exports, quantizes, and serves Laya and other Jev-style models for offline CPU inference

11·GitHub repo

laya-jev-lab

yibie

Compares Jev and Laya decision models and evaluates a local-first inference cascade

9·GitHub repo

laya_router

glukicov

A Python model router compares local Laya decisions with a GPT-5 nano routing model

6·GitHub repo

jevbench

dhruvmehra

A reproducible benchmark compares JEV, Laya, and other classifiers across datasets and metrics

6·GitHub repo

laya-onnx

MstyAI

Runs Laya decision models locally with ONNX Runtime through a Go library and CLI

5·GitHub repo

laya-plays-smb3

cv

Demonstrates Laya controlling Super Mario Bros. 3 with recorded, replay-verified decisions

5·GitHub repo

reflexbench

brida-ai

Benchmarks Laya and other typed-decision engines across quality, calibration, robustness, and latency

5·GitHub repo
L

laya-jev-benchmark

Luni

Benchmarks Laya against Jev and other models on phishing detection and calibration

4·HF dataset
C

MacJev-322M-4K-Laya

chaoliangUNSW

Fine-tunes Laya for long-context local Mac agents and compares it with the base model

4·HF model

jev-laya-benchmark

harrymunro

Benchmarks local MLX Laya against TypeSafe's hosted Jev on synthetic decision tasks

4·GitHub repo

sysone-bench

instax-dutta

Compares Laya and Jev on identical inputs across multiple decision benchmarks

4·GitHub repo

jev-tests

schacon

Compares Laya, Jev, Kev, and Claude in three macOS typed-decision demos

4·GitHub repo

chunklaya

myxamediyar

Chunks long documents and uses Laya for typed-decision lookup across inputs up to one million tokens

3·GitHub repo

zero-shot-ie-bench

umstek

Compares Laya and other zero-shot systems across information-extraction and classification tasks

3·GitHub repo

structured-decision-bench

zhengbangbo

Benchmarks structured decisions from Qwen3, Jev, and Laya CoreML

3·GitHub repo

runtime-tutorials

Runtime-weekly

Provides runnable Python guides for Laya classification and a comparison of Laya with other systems

3·GitHub repo

doomLaya

azalio

Trains and compares Laya and Jev agents playing FreeDoom with reproducible results

3·GitHub repo

More ways to use Laya