cd /news/ai-chips/nvidia-says-its-groq-3-lpx-is-four-t… · home topics ai-chips article
[ARTICLE · art-110142] src=the-decoder.com ↗ pub= topic=ai-chips verified=true sentiment=· neutral

Nvidia says its Groq 3 LPX is four times faster than Cerebras, but the math is more complicated

Nvidia has moved its Groq 3 LPX inference chip into full production, claiming 3,400 tokens per second on Gemma 4 31B, which it says is four times faster than Cerebras. However, Nvidia requires at least 64 accelerators to achieve this performance, while Cerebras needs only one or two, according to The Register, leaving questions about scalability with large Mixture-of-Experts models.

read1 min views2 publishedAug 25, 2026

Nvidia is moving its Groq 3 LPX inference chip into full production and reports 3,400 tokens per second on Gemma 4 31B, four times faster than Cerebras. But the numbers don't tell the whole story. Nvidia needs at least 64 accelerators to get there, while Cerebras needs only one or two, according to The Register. How well the architecture scales with large MoE models remains an open question.

The article Nvidia says its Groq 3 LPX is four times faster than Cerebras, but the math is more complicated appeared first on The Decoder.

── more in #ai-chips 4 stories · sorted by recency
── more on @nvidia 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/nvidia-says-its-groq…] indexed:0 read:1min 2026-08-25 ·