sys1bench
by rssr25
A benchmark measures calibration, robustness, and other properties of typed decision models including Laya
Benchmark for typed System One decision models (Jev, Laya, and whatever comes next): calibration against a noise floor, framing sensitivity, selective prediction, ordinal fidelity, interference, robustness. pip install sys1bench.
Platforms
Use cases