Grok 4.6 catches up to frontier models while costing far less. According to the Artificial Analysis Intelligence Index, SpaceXAI's new model scores 61 points, tying OpenAI's GPT-5.6 Sol. Only Anthropic's Claude Opus 5 (63) and Claude Fable 5 (62) score higher. That's a five-point jump over its predecessor, Grok 4.5.
Grok 4.6 performs especially well on agentic tasks, where models independently carry out multi-step workflows. On the GDPval-AA v2 benchmark, which aims to measure real-world knowledge work on a computer, it ranks second with an Elo score of 1,753, trailing only Claude Opus 5. It completes complex tasks in about 53 steps. Claude Opus 5 needs roughly 103.
Pricing stays at $2/$6 per million tokens. That's more than 60 percent cheaper than Claude Opus 5 ($5/$25) and GPT-5.6 Sol ($5/$30). Grok 4.6 is available now through the API, Cursor, Grok Build, and partners like OpenRouter, Vercel, and Cloudflare. For the first week, x.ai is offering double the usage quota in Grok Build and Cursor.
AI News Without the Hype – Curated by Humans
Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.
Subscribe now