cd /news/artificial-intelligence/doublespeed-is-sota-in-social-intell… · home topics artificial-intelligence article
[ARTICLE · art-118369] src=doublespeed.ai ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

Doublespeed is SoTA in social intelligence benchmarks

Doublespeed AI, the model powering the social media platform doublespeed, achieved 61.0% accuracy in a benchmark of 200 real TikTok post pairs, outperforming frontier models including Claude and GPT 5.6 Sol, which scored between 50.8% and 53.3%, within chance levels. The benchmark, conducted on 2026-09-01, tested models' ability to predict which of two posts from the same account would get more views, based solely on copy, with doublespeed AI's performance excluding random guessing.

read2 min views1 publishedSep 1, 2026

Take two TikTok slideshows from the same account, posted within three weeks of each other. Can a model read the copy of both and pick the winner? We asked four frontier models and doublespeed AI, 200 times.

Correct picks on 200 real outcome pairs #

Benchmarked 2026-09-01 on 200 pairs from 174 accounts posting through doublespeed. Every model saw the copy of both posts and never the view counts; random guessing lands at 50%. The frontier models answered zero-shot; doublespeed AI was trained on this platform's past outcomes, which is the point.

Try one of the questions yourself #

This pair is one of the 200 exam questions. The models saw only the copy; you get the full slides. All four frontier models got it wrong.

the situationship to ai therapy pipeline is actually insane 💔😭 #fyp #viral #relatable #situationship #healingjourney

turns out most relationship fights are just two people trying to feel understood fr 💭 #relationships #datingadvice #voicedaitwin #couplegoals #emotionalintelligence

Both posts went up on the same account within three weeks. One got over a hundred times the views of the other. Click the one you think won.

How it was measured #

The exam is built from real posting history: slideshows published to TikTok through doublespeed, with their view counts. A pair is two posts from the same account within 21 days where the winner got at least 2x the loser's views and at least 500 views. Comparing within one account cancels follower count and algorithmic standing, so the copy is what varies.

Each model saw both posts' slide texts, captions, and slide counts, with the winner's position randomized, and answered one question: "These two TikTok slideshow posts are from the same account. Based only on their copy, which one performs better (gets more views)? Answer with the concept index." No examples, no product briefing, no retries. The Claude models were called through the Anthropic API, GPT 5.6 Sol through OpenRouter; one call per model failed and is excluded from that model's count.

Accuracy is the share of pairs where a model picked the real winner. Intervals are bootstrap 95% confidence intervals over pairs. The frontier models landed at 50.8% to 53.3%, all within their intervals of a coin flip. doublespeed AI landed at 61.0%, and its interval excludes chance. Outcome data moves the needle; model scale alone does not.

doublespeed AI in this benchmark is the same model that scores every draft inside doublespeed before it goes out. Your posting history makes it sharper.

Contact us for proprietary social data

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @doublespeed ai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/doublespeed-is-sota-…] indexed:0 read:2min 2026-09-01 ·