cd /news/large-language-models/xai-launches-grok-4-7-with-stronger-… · home topics large-language-models article
[ARTICLE · art-136135] src=cryptobriefing.com ↗ pub= topic=large-language-models verified=true sentiment=↑ positive

xAI launches Grok 4.7 with stronger coding and knowledge work performance

XAI launched Grok 4.7, its latest frontier model for coding and knowledge work, keeping the same starting price as Grok 4.6 at $2 per million input tokens and $6 per million output tokens. Grok 4.7 scored 46.3% on CursorBench 4.0 versus 40.4% for Grok 4.6 and rose to 38% on Terminal Bench 4.0 from 20.3%, using a larger base model and a longer reinforcement learning run. The model also scored 71% on DeepSWE 1.1 at high reasoning effort, 1,657 on AA Briefcase and 64% on EEBench, and xAI introduced a new safeguard stack in which the model allowed 3.3% of risky dual use prompts through on HackerBench 0.3 and scored 62.4% on a LatchBio biosafety benchmark.

by read2 min views11 publishedSep 21, 2026
xAI launches Grok 4.7 with stronger coding and knowledge work performance
Image: Cryptobriefing (auto-discovered)

The new model uses a larger base model and extended reinforcement learning while keeping Grok 4.6 pricing.

xAI has launched Grok 4.7, its latest frontier model focused on coding and knowledge work, with improved performance on long running tasks and stronger safeguards.

The model uses a larger base model than Grok 4.6 and underwent a longer reinforcement learning run focused on more difficult tasks that can take hours to complete. xAI said the training improved Grok 4.7’s ability to verify its own work, manage longer context and operate within the Grok Bot environment.

Grok 4.7 improves on its predecessor across the benchmarks published by xAI. It scored 46.3% on CursorBench 4.0 compared with 40.4% for Grok 4.6, while its Terminal Bench 4.0 score rose to 38% from 20.3%.

The model also scored 71% on DeepSWE 1.1 at high reasoning effort, 1,657 on AA Briefcase and 64% on EEBench. xAI’s results put Grok 4.7 ahead of GPT 5.6 Sol on several of those evaluations, though GPT 5.6 Sol and Fable 5.1 remained ahead on others.

The release follows Grok 4.6, which xAI launched in August with a focus on long running agents and complex coding and knowledge work. Grok 4.6 was priced at $2 per million input tokens and $6 per million output tokens.

AI, tech, and the markets they move—in one daily briefing.

Daily. Free. Join 34,000+ readers across crypto, finance, and policy.

xAI is keeping the same starting price for Grok 4.7 at $2 per million input tokens and $6 per million output tokens. A faster version offering twice the output speed is available at twice the price.

The company also introduced a new safeguard stack for Grok 4.7. xAI said the model allowed 3.3% of risky dual use prompts through on HackerBench 0.3 and scored 62.4% on a biosafety benchmark from LatchBio.

The company is also providing selected cybersecurity partners with access to the model’s red team capabilities for defensive research.

Grok 4.7 is available through Cursor, Grok Build and the Grok API, as well as third party coding tools, model routers and cloud platforms.

Disclosure: This article was edited by Estefano Gomez. For more information on how we create and review content, see our

Editorial Policy.

── more in #large-language-models 4 stories · sorted by recency
── more on @xai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/xai-launches-grok-4-…] indexed:0 read:2min 2026-09-21 ·