🤖Google Ships Gemini 3.6 Flash, Cuts Agent Costs 65%
TL;DR
Google DeepMind released Gemini 3.6 Flash plus 3.5 Flash-Lite and 3.5 Flash Cyber on July 21, skipping the flagship 3.5 Pro. The workhorse model uses 17% fewer tokens and claims up to 65% lower cost on long-horizon agent tasks.
Google DeepMind released Gemini 3.6 Flash plus 3.5 Flash-Lite and 3.5 Flash Cyber on July 21, skipping the flagship 3.5 Pro. The workhorse model uses 17% fewer tokens and claims up to 65% lower cost on long-horizon agent tasks.

Key Points
Three models shipped July 21: Gemini 3.6 Flash, 3.5 Flash-Lite, 3.5 Flash Cyber
Pricing: $1.50 per 1M input tokens, $7.50 per 1M output tokens
Google claims ~17% token reduction and up to 65% cheaper long-horizon agent runs
3.5 Pro flagship still in partner testing, not released
Why It Matters
Google is competing on cost-per-agent-task, not benchmark bragging rights, a sign the frontier fight has moved to unit economics for teams running agents at scale.
Quick Facts
Frequently Asked Questions
Why does this matter?
Google is competing on cost-per-agent-task, not benchmark bragging rights, a sign the frontier fight has moved to unit economics for teams running agents at scale.
What happened?
Google DeepMind released Gemini 3.6 Flash plus 3.5 Flash-Lite and 3.5 Flash Cyber on July 21, skipping the flagship 3.5 Pro. The workhorse model uses 17% fewer tokens and claims up to 65% lower cost on long-horizon agent tasks.
Comments
Be the first to comment
Enjoyed this article?
Get it daily. 7am. Free. Reads in 5 minutes.
Join 2,179 builders reading daily.