17:38
2026-09-28
dev.to
ai-infrastructure
A weekend with TensorFold on a MacBook: the engine mattered, the quant did not
A developer benchmarked TensorFold, a speculative-decoding inference engine, on a MacBook Pro with an M5 Max and 128 GB of unified memory, finding it decoded Qwen3.8-27B at 220 tokens per second on a …