sys1bench

by rssr25

A benchmark measures calibration, robustness, and other properties of typed decision models including Laya

Benchmark for typed System One decision models (Jev, Laya, and whatever comes next): calibration against a noise floor, framing sensitivity, selective prediction, ordinal fidelity, interference, robustness. pip install sys1bench.

Platforms

Related projects