cd /news/artificial-intelligence/google-just-made-agents-3x-cheaper-t… · home topics artificial-intelligence article
[ARTICLE · art-68883] src=the-ai-corner.com ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

Google just made agents 3x cheaper to run. Here is the playbook

Google released three Flash-tier models on Monday, including Gemini 3.6 Flash, which cuts cost per task from $0.59 to $0.50 and time per task from 2.7 minutes to 1.3, while Gemini 3.5 Flash-Lite runs at 350 tokens per second for $0.30 input and $2.50 output. The models prioritize operational efficiency over intelligence, with Gemini 3.6 Flash scoring the same 50 on the Intelligence Index as its predecessor, as product lead Tulsee Doshi said developers need "higher token efficiency, lower latency, and more reliable performance.

read2 min views1 publishedJul 22, 2026
Google just made agents 3x cheaper to run. Here is the playbook
Image: The-Ai-Corner (auto-discovered)

Gemini 3.6 Flash cuts your cost per task ~18% and your time per task in half, and Flash-Lite runs at 350 tokens a second for 30 cents. The full build guide: the routing table, the migration steps, and

On Monday, Google shipped the least glamorous release of the year, and the one most likely to change your infrastructure bill.

No new flagship. No frontier score. Instead, three Flash-tier models built for one job: running agents at scale, cheaper. And the numbers underneath are the kind that rewrite unit economics without a headline:

▫️ Gemini 3.6 Flash ships 17% fewer output tokens than 3.5 Flash at a lower price. Artificial Analysis measured cost per task falling from $0.59 to $0.50, and time per task dropping from 2.7 minutes to 1.3.

▫️ Gemini 3.5 Flash-Lite runs at 350 tokens per second, the fastest model Artificial Analysis clocked at launch, for $0.30 in, $2.50 out.

▫️ Gemini 3.5 Flash Cyber, a defensive security model gated to governments and trusted partners, found 55 confirmed vulnerabilities in the V8 JavaScript engine against Claude Opus 4.6’s 36.

3.6 Flash is barely smarter than the model it replaces. Its composite Intelligence Index sits at 50, the exact same score as 3.5 Flash. Google spent this release cycle making the model cheaper, faster, and more token-efficient instead of smarter, on purpose. Tulsee Doshi, who leads Gemini product, said it plainly: developers building production agents need “higher token efficiency, lower latency, and more reliable performance,” over another benchmark point.

Which reframes the whole thing. The win here is operational rather than intellectual, and it only shows up if you route correctly, migrate cleanly, and stack the cost levers Google buried in the docs. Get it right and the same agent workload costs a fraction of what it did Friday. Get it wrong and you leave most of the savings on the table.

The routing table, the migration steps, the cost math, the cost levers, and the Cyber guide, in one system.

Get The Gemini Flash Agent Playbook below 👇

Keep reading with a 7-day free trial #

Subscribe to The AI Corner to keep reading this post and get 7 days of free access to the full post archives.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @google 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/google-just-made-age…] indexed:0 read:2min 2026-07-22 ·