Laya for runtimes & ports
Ports and inference engines that run the Laya weights.
245 projects
chunklaya
myxamediyar
Chunks long documents and uses Laya for typed-decision lookup across inputs up to one million tokens
laya-web
r4ai
Runs Laya typed-decision models in browsers and Node.js using ONNX Runtime Web
laya-typed-decisions-GGUF
mys
Provides GGUF builds of Laya's typed-decisions model compiled with ggmlc
laya-onnx
inferenceprince
Exports Laya to ONNX for inference with ONNX Runtime on CPUs, GPUs, and browsers
laya-onnx
receptron
Exports Laya to ONNX for use with Node.js or ONNX Runtime
laya-onnx
gqgs
Exports the Laya model to quantized ONNX for browser inference
docker-laya
chneau
Serves Laya typed-decision predictions through an authenticated, multi-checkpoint FastAPI service
laya-portable
MatteoGauthier
Exports Laya to ONNX and provides JavaScript runtimes for Node.js and browser inference
JevCoreML
GodModeAI2025
Provides Core ML decision models, a Swift package, an HTTP server, and a demo app
laya-jev-api
smallnest
Serves local Laya Core ML inference through a Jev-compatible HTTP API
laya.axera
AXERA-TECH
Exports and calibrates Laya models for AXERA NPU deployment
laya-serve
stiermid
Serves local Laya decision models through a Jev-compatible HTTP API
openzl-laya
madeye
Adds optional on-device Laya routing to OpenZL compression on Linux and macOS
laya-windows
Zuhair-01
Ports Laya typed-decision inference to Windows using ONNX Runtime and DirectML
laya-cli
MIt9
A CLI runs Laya typed decisions with batch, routing, and resident-daemon modes
Decis
chaitin
Serves open decision models through a self-hosted TypeSafe System One API
laya-cuda
Alexw1111
Provides a lightweight CUDA inference library for Laya
laya-rs
Fanaperana
Reimplements Laya in Rust for zero-shot text classification with verified numerical parity
laya-ternary-lite
xixi3548942758-design
Quantizes Laya to ternary weights for smaller, lower-memory inference
decidealot
psyb0t
Serves local Laya and Von decision models through TypeSafe-compatible HTTP and MCP APIs
laya-candle
mannlohchab
Runs English Laya decision-model inference using the Candle framework
laya-multilingual-gguf
fr0stbit3
Provides F16 and quantized GGUF conversions of the multilingual Laya model for llama.cpp
open-jev-laya-multilingual-onnx
killkli
A browser-ready ONNX export of Laya multilingual for typed decisions
laya-neutron-gguf
wigcheng5566
A GGUF package of Laya's encoder and decision head for CPU and Neutron NPU inference
More ways to use Laya