09:02
2026-08-18
dev.to
large-language-models
I got Qwen3.8-27B running on dual RTX 3090s (no NVLink) under WSL2 โ every pitfall I hit
A developer detailed the process of running Qwen3.8-27B on dual RTX 3090s without NVLink under WSL2, achieving 170-210 tok/s on code/JSON. The setup required specific CUDA 13.0 toolchain, SGLang 0.5.1โฆ