mlx-quantization-lora-eval
by D-Harshith
Benchmarks MLX quantization and LoRA alongside Laya encoder classifiers on GitHub issue classification
On-device LLM evaluation on Apple Silicon with MLX: Laya-mlx, Llama 3.2 3B at 4-bit, 8-bit and bf16 quantization, LoRA fine-tuning and adapter fusing, compared on accuracy, speed and memory for GitHub issue classification.
Use cases