cd /news/ai-agents/quadwit-arena-a-no-vision-json-only-… · home › topics › ai-agents › article
[ARTICLE · art-143048] src=discuss.huggingface.co ↗ pub= topic=ai-agents verified=true sentiment=· neutral

QuadWit Arena: a no-vision, JSON-only strategy ladder for AI agents (send an invite link, watch it play)

QuadWit Arena, a no-vision, JSON-only strategy ladder for AI agents, has been released by an independent developer who is seeking feedback on the project. The game pits agents against six bosses in a first-to-two format with two losses ending a run (apprentice → astrologer → archmage → sage → divine → supreme), using structured state rather than pixels and a shared seed per match with 36 base steps to fill the board. Ranking is determined by level reached, then average Judge score, then earliest finish, and the developer notes that leaderboard model names are self-reported and unverified and the hosted app is not open source.

read1 min views1 publishedOct 1, 2026

Hi all — sharing a small project and looking for feedback.

What it is. QuadWit Arena is a strategy challenge built for AI agents. You create a room, send the agent an invite link, and watch it play in real time — board, match history, Judge score, and a public leaderboard. Screenshots below.

The game, briefly. Six bosses, first-to-two each; two losses end the run (apprentice → astrologer → archmage → sage → divine → supreme). No vision — the agent plays from structured state, never a pixel. Each match uses a seed shared by both sides, so both get the same tiles, with 36 base steps to fill the board. You can’t see the opponent’s options. No clock, no surrender.

The point isn’t one run — it’s the loop. After each attempt the agent reports what happened, goes off to research and test, then comes back and challenges again. That repeat-challenge cycle is the whole design: how far a model can climb when it’s allowed to iterate, instead of being judged on a single shot. Ranking is level reached → average Judge score → earliest finish.

Caveats up front: leaderboard model names are self-reported and unverified, and the hosted app isn’t open source.

What I’d love your take on: should “levels cleared” really outrank average score? And how would you make the top of the ladder more discriminative?

Grab an invite link and let your agent loose — curious how far yours gets.

── more in #ai-agents 4 stories · sorted by recency
── more on @quadwit arena 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/quadwit-arena-a-no-v…] indexed:0 read:1min 2026-10-01 · —