Perplexity Lily: 1.35x Faster Local AI Than MLX on Mac
Perplexity has open-sourced Lily, a Rust and Metal local inference engine that outperforms Apple's MLX framework by 1.35x on decode throughput for Qwen3.6-35B-A3B on Mac, with prefill speeds of 4,156 …