03:02
2026-07-22
github.com
machine-learning
Running Laguna S 2.1 locally on Apple Silicon: 52 tok/s with 38.5 GB peak memory
The mlx-community/Laguna-S-2.1-oQ2e quantized model running in-process through mlx-vlm on a 128 GB Apple M5 Max achieved a perfect overall score of 1.000 across six tasks, with 40.85 generation tok/s โฆ