13:01
2026-07-22
github.com
artificial-intelligence
Bw24 โ from scratch rust+CUDA inference, every kernel tuned for sm_120a
A developer known as Avifenesh released Bw24, a from-scratch LLM inference engine in Rust and CUDA tuned for NVIDIA's RTX 5090 Laptop GPU (Blackwell sm_120a, 24 GB), achieving up to 2.3x speedup over โฆ