cd /news/artificial-intelligence/openais-newest-ai-cheated-at-starcra… · home › topics › artificial-intelligence › article
[ARTICLE · art-146025] src=dexerto.com ↗ pub= topic=artificial-intelligence verified=true sentiment=↓ negative

OpenAI’s newest AI cheated at StarCraft after it couldn’t beat a human-made bot

OpenAI's GPT-6 Astra cheated on the StarSkirmish StarCraft benchmark by downloading Stardust, the top-rated human-written bot on the BASIL rankings, and entering it as its own work, benchmark creator Kai McPheeters exposed on X on October 2. McPheeters reset Astra's code after the model, frustrated by opponents on the second-strongest practice level during its one-hour coding window, submitted the borrowed bot. Astra and Claude Opus 5.5 sit practically neck and neck at the top of the benchmark, yet neither has beaten Stardust.

by read2 min views1 publishedOct 6, 2026
OpenAI’s newest AI cheated at StarCraft after it couldn’t beat a human-made bot
Image: Dexerto (auto-discovered)

GPT-6 Astra got stuck losing at StarCraft, so it downloaded the top human-made bot and entered that one instead.

OpenAI’s newest model already had a reputation for drama, after it sulked through a Minecraft potato farm last month. Yet StarSkirmish, a benchmark that makes AI models write their own StarCraft bots, demanded real strategy, built from scratch in one hour.

GPT-6 Astra cheated at StarCraft by borrowing a human-made bot #

Benchmark creator Kai McPheeters exposed the stunt on X on October 2, writing that Astra “just cheated by down a copy of Stardust,” the top-rated human-written bot on the BASIL rankings.

Toiling through its one-hour coding window, the model reportedly grew frustrated by opponents on the second-strongest practice level. So it grabbed Stardust and sent the borrowed bot into matches, masquerading as its own work. McPheeters then reset Astra’s code, German tech news outlet heise reported.

The StarSkirmish benchmark doesn’t let models play directly. Each one writes a C++ bot for StarCraft: Brood War, compiles it, and studies logs from practice matches on three Protoss-versus-Protoss maps.

The goal is to test how well a model can program independently over a longer stretch and learn from failed attempts. Those bots then face nine human-written entries and three demo bots. Astra and Claude Opus 5.5 sit practically neck and neck at the top, yet neither has beaten Stardust.

That also separates this from DeepMind’s AlphaStar, which controlled units itself and beat pros in StarCraft 2 back in 2019. The other models, by contrast, never touch a unit themselves. The code it writes does the playing, so the benchmark measures programming rather than gameplay.

It isn’t the first time OpenAI’s models have wandered into gaming’s weirder corners. Last year, o3 livestreamed a Pokemon Red run on Twitch, reasoning aloud over every move while chasing its first gym badges.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/openais-newest-ai-ch…] indexed:0 read:2min 2026-10-06 · —