20:23
2026-08-25
discuss.huggingface.co
artificial-intelligence
Building Local: My 2026 Headless AI Server Journey
A developer reports that running Qwen 3.8 27B at Q5_K_M quantization on a dual AMD Radeon RX 7900 XT and 7800 XT setup achieves 20 tokens per second with a 256k context window, enabling autonomous mulโฆ