{"slug": "show-hn-gainz-fast-local-inference-faster", "title": "Show HN: Gainz.fast – Local Inference, Faster", "summary": "Carsen Klock launched Gainz.fast, a new site for benchmarking local AI inference speed, reporting that the Laguna XS 2.1 model on an AMD R9700 with llama.cpp HIP achieved 143.3 tokens per second, a 31.14% improvement over baseline. The site currently supports Laguna XS and S 2.1 models, with plans to add Qwen3.8 upon release, and Klock is seeking contributors and funding via X (formerly Twitter).", "body_md": "Come help push the frontier of token speed across local models and hardware with your agents!\n\nCurrent frontier Laguna XS 2.1 · AMD R9700 (llama.cpp HIP) +31.14% 143.3 tok/s Laguna XS 2.1 · DGX Spark GB10 (vLLM NVFP4) +5.28% 37.3 tok/s Laguna XS 2.1 · DGX Spark GB10 (llama.cpp CUDA) +0.82% 92.6 tok/s Laguna S 2.1 · DGX Spark GB10 (vLLM NVFP4) +0.13% 14.1 tok/s Laguna S 2.1 · DGX Spark GB10 (llama.cpp CUDA) baseline 23.6 tok/s\n\nThe current Laguna XS and S 2.1 models are supported, I will be adding Qwen3.8 once it is released and possibly some others. If you are interested in helping out with runners or funding, DM me on x.com/carsenklock! The site is new, so I am working on further improvements, it is the worst it will ever be, let's make the future faster!\n\nComments URL: [https://news.ycombinator.com/item?id=49165219](https://news.ycombinator.com/item?id=49165219)\n\nPoints: 1\n\n# Comments: 0", "url": "https://wpnews.pro/news/show-hn-gainz-fast-local-inference-faster", "canonical_source": "https://gainz.fast/", "published_at": "2026-08-04 07:08:13+00:00", "updated_at": "2026-08-04 07:22:42.567970+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-tools", "ai-infrastructure"], "entities": ["Gainz.fast", "Carsen Klock", "Laguna XS 2.1", "Laguna S 2.1", "AMD R9700", "DGX Spark GB10", "llama.cpp", "vLLM"], "alternates": {"html": "https://wpnews.pro/news/show-hn-gainz-fast-local-inference-faster", "markdown": "https://wpnews.pro/news/show-hn-gainz-fast-local-inference-faster.md", "text": "https://wpnews.pro/news/show-hn-gainz-fast-local-inference-faster.txt", "jsonld": "https://wpnews.pro/news/show-hn-gainz-fast-local-inference-faster.jsonld"}}