Tested on:
Qwen2.5-Coder-0.5B-Instruct-Q3_K_L.gguf (369 MB)
Select a Qwen2.5 GGUF model file to start inference...
Engine Console
source & further reading
wasm-gguf.netlify.app — original article
A developer released a single-file GGUF inference engine, tested with the Qwen2.5-Coder-0.5B-Instruct-Q3_K_L.gguf model (369 MB), enabling local inference from a single file.
Tested on:
Qwen2.5-Coder-0.5B-Instruct-Q3_K_L.gguf (369 MB)
Select a Qwen2.5 GGUF model file to start inference...
Engine Console
EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.