system-one-triage-eval
by Pedro-Yataco
Evaluates ticket-triage performance across Laya, Jev, and language-model backends
Evaluation harness for "System One" models (typed decisions with calibrated probabilities) on classification and triage tasks. Compares accuracy, calibration, latency and cost across backends such as Laya, Jev and an LLM baseline.
Platforms
Use cases