{"slug": "show-hn-single-file-gguf-inference", "title": "Show HN: Single-File GGUF Inference", "summary": "A developer released a single-file GGUF inference engine, tested with the Qwen2.5-Coder-0.5B-Instruct-Q3_K_L.gguf model (369 MB), enabling local inference from a single file.", "body_md": "Tested on:\n\n      Qwen2.5-Coder-0.5B-Instruct-Q3_K_L.gguf (369 MB)\n    \n\n```\nSelect a Qwen2.5 GGUF model file to start inference...\n```\n\nEngine Console", "url": "https://wpnews.pro/news/show-hn-single-file-gguf-inference", "canonical_source": "https://wasm-gguf.netlify.app/", "published_at": "2026-09-03 05:01:19+00:00", "updated_at": "2026-09-03 05:22:07.927327+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "developer-tools", "ai-tools"], "entities": ["Qwen2.5-Coder-0.5B-Instruct"], "alternates": {"html": "https://wpnews.pro/news/show-hn-single-file-gguf-inference", "markdown": "https://wpnews.pro/news/show-hn-single-file-gguf-inference.md", "text": "https://wpnews.pro/news/show-hn-single-file-gguf-inference.txt", "jsonld": "https://wpnews.pro/news/show-hn-single-file-gguf-inference.jsonld"}}