mlx-quantization-lora-eval

by D-Harshith

Benchmarks MLX quantization and LoRA alongside Laya encoder classifiers on GitHub issue classification

On-device LLM evaluation on Apple Silicon with MLX: Laya-mlx, Llama 3.2 3B at 4-bit, 8-bit and bf16 quantization, LoRA fine-tuning and adapter fusing, compared on accuracy, speed and memory for GitHub issue classification.

Related projects