{"type": "article", "title": "Apple Silicon and macOS VMs: 11–16× Faster LLM Inference with Llama.cpp", "publisher": "Web Pulse", "url": "https://wpnews.pro/news/apple-silicon-and-macos-vms-11-16x-faster-llm-inference-with-llama-cpp", "original_source": "https://github.com/trycua/cua/blob/main/blog/gpu-passthrough-macos-vms.md", "published": "2026-08-11T14:50:33+00:00", "accessed": "2026-08-11", "id": "apple-silicon-and-macos-vms-11-16x-faster-llm-inference-with-llama-cpp"}