23:48
2026-09-08
runtimewire.com
artificial-intelligence
Inception ships Mercury 2.5 at 1,107 tokens per second, by its count
Inception Labs launched Mercury 2.5 on September 8th, a diffusion-based language model API that the company claims generates 1,107 tokens per second on widely available Nvidia GPUs, with a 260,000-tokβ¦