cd /news/artificial-intelligence/godot-benchmark-sol-k3-fable · home topics artificial-intelligence article
[ARTICLE · art-74764] src=ziva.sh ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

Godot Benchmark: Sol > K3 > Fable

A benchmark of three AI models for game development found that GPT 5.6 Sol outperforms Kimi 3 and Claude Fable 5 in speed and cost, despite producing slightly worse results. Sol completed a 3D parkour world in 9.7 minutes at $2.94, while Kimi took 79 minutes at $3.14 and Fable took 32.5 minutes at $13. The test, which involved creating a 3D parkour world with three races using Kenney assets, rated Sol's conversation efficiency at 9/10, Kimi's thoroughness at 7/10, and Fable's debugging at 6/10.

read2 min views1 publishedJul 19, 2026
Godot Benchmark: Sol > K3 > Fable
Image: Ziva (auto-discovered)

So here’s the thing - every model is benchmaxxing, so how do you know which one is actually the best to use? The thing we’ve found most reliable is to benchmark it with something that’s relevant to you.

TL;DR #

Kimi actually beat Fable, but it’s unusably slow. Sol’s speed and affordability make it better for game development, even if the end result is slightly worse than the other two.

Model World Conversation Cost Time
GPT 5.6 Sol 6 9 $2.94 9.7 min
Kimi 3 8 7 $3.14 79 min
Claude Fable 5 7 6 $13* 32.5 min

The Prompt #

Make a 3D parkour world where you compete against bots for time. There should be 3 races in the world, and you can start any of the races by walking over to their start point. Use the assets in Kenney_Platformer_Kit/ to build your world.

Created Worlds #

GPT 5.6 Sol Kimi 3 Claude Fable 5

PlayPlay## Conversation Analysis I ran each conversation transcript through Opus 4.8 and had it come up with a grade:

GPT 5.6 Sol— tightest loop; headless logic checks, playtests only where they counted.** 9/10**·conversation** Kimi 3**— most thorough of the three (12 playtests), but redundant and slow.** 7/10**·conversation** Claude Fable 5**— sharpest debugging, but ships blind: it never sees its own playtest frames.** 6/10**·conversation

My Thoughts #

Kimi is lowkey very good here, but it’s painfully slow. Maybe it’s a reliability issue on their GPUs and it’ll be better once more people can self-host, but the truth is right now the time it takes to do things makes it borderline unusable.

Credits #

Assets used for the game generation are provided by Kenny.nl

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @gpt 5.6 sol 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/godot-benchmark-sol-…] indexed:0 read:2min 2026-07-19 ·