02:46
2026-09-12
twitter.com
large-language-models
DeepSeek v4.1 flash runs 23 seconds/token on a 2020 16gb M1 Mac Mini
A developer posting as FP4 Brain on X reported running DeepSeek V4.1 Flash locally on a 16GB M1 Mac Mini using original FP4/FP8 weights, SSD streaming, and a custom MLX runner built on the pipenetwork…