23:58
2026-08-04
promptcube3.com
artificial-intelligence
Maple-Preview: 20B MoE Hits 120 tok/s on iPhone
Maple-Preview, a ternary 20B mixture-of-experts model, achieves 120 tokens per second on an iPhone, as announced on Hacker News. The model's performance highlights the potential for local AI on low-enβ¦