cd /news/ai-products/measure-the-queue-not-the-bot · home topics ai-products article
[ARTICLE · art-118698] src=amitkvint.com ↗ pub= topic=ai-products verified=true sentiment=· neutral

Measure the Queue, Not the Bot

A customer support leader argues that AI support agents should be measured by queue-wide resolution time and satisfaction, not deflection rates, citing that deflection metrics can be gamed and don't reflect actual problem-solving. The author reports that after rolling out an AI agent to a queue of 5,000 tickets per month, average resolution time dropped from 24 hours to 10 hours while satisfaction stayed above 95%, and notes that industry voices including Automattic's CX lead and a Smart Customer Service column have echoed similar concerns.

read2 min views1 publishedSep 2, 2026
Measure the Queue, Not the Bot
Image: source

When we rolled out an AI support agent to a queue of around 5,000 tickets a month, we never built a deflection dashboard. Not because we had some strong opinion against it. We just couldn’t see what it would actually tell us.

Deflection measures the conversations that never reach a human. It doesn’t tell you whether the customer’s problem was solved. Someone who gives up is not a successful deflection. They might be a refund next week, or another ticket with a completely different subject line.

This August, I noticed three very different parts of the industry making essentially the same point. Automattic’s CX lead argued that deflection is not a support strategy. A vendor-side column in Smart Customer Service argued that resolution rate should replace it. And China’s consumer-protection coverage quoted a legal expert saying AI support should be judged on resolution and escalation efficiency, with a regulator behind the idea.

So, problem solved? Not quite.

The problem is that a resolution rate can be gamed too.

Close-on-silence is one common example: if a customer doesn’t reply within 72 hours, the ticket is marked as resolved. Then there are bot-confirmed resolutions, where the agent asks “did that help?” and treats the customer disappearing as a yes. And if reopened tickets are logged as new tickets, one failure can conveniently turn into two “successful” resolutions.

Basically, any metric that only looks at the bot’s part of the queue can be made to look good by carefully choosing what counts.

For me, the answer is to not give the AI its own scoreboard at all. You look at one queue and one set of numbers. AI-handled and human-handled tickets should be measured together, using average resolution time and customer satisfaction across the whole queue. If the AI said it had solved something when it hadn’t, the customer would come back, and that would show up in our numbers. There should be no separate AI success metric to hide behind.

That also gives you a much safer way to scale it. You start the agent on 10% of the queue and only increase its share when the overall numbers hold up. Six months later, it'll be covering the entire queue. In our case average resolution time had dropped from around 24 hours to around 10, while satisfaction stayed above 95%.

And those numbers were for everyone, not just the tickets where the AI happened to perform well.

I understand why deflection dashboards are attractive. They’re easy to set up, easy to explain, and, conveniently, they usually go up.

But I think the metric you use to judge a support team should be something the customer would recognize as a success too.

Nobody ever wrote to us to say, “Thanks for deflecting me.” :)

── more in #ai-products 4 stories · sorted by recency
── more on @automattic 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/measure-the-queue-no…] indexed:0 read:2min 2026-09-02 ·