laya-ternary-lite

by xixi3548942758-design

Quantizes Laya to ternary weights for smaller, lower-memory inference

1.58-bit ternary quantization of the Laya decision model -- 9.17x smaller, 81.2% agreement

Related projects