19:11
2026-09-12
rheisen.me
ai-agents
Hardware x Model x Harness
Cerebras is now serving OpenAI's GPT-5.6 Sol at up to 750 output tokens per second, according to a Cerebras blog post, a speed the author says shifts the bottleneck from the model to the surrounding a…