10:49
2026-10-08
forum.level1techs.com
ai-infrastructure
Advice on a local AI server before buying
A forum commenter running Qwen 3.8 Flash Next on dual AMD R9700 GPUs reported roughly 200 tokens per second decode and 6200 tokens per second prefill at 4-bit precision, and advised buyers to use an x…