Skip to content
Gemini 3.6 Flash: Cut Agent Tokens With thinking_level — ContentBuffer guide

Gemini 3.6 Flash: Cut Agent Tokens With thinking_level

K
Kodetra Technologies··10 min read Intermediate

Summary

Build a token-thrifty tool-calling agent on Gemini 3.6 Flash using the new thinking_level control.

Gemini 3.6 Flash: Cut Agent Tokens With thinking_level

Google shipped Gemini 3.6 Flash on July 21, 2026, and the headline is not a benchmark score, it is the bill. This is the cheaper, more token-efficient Flash tier, tuned for agentic and coding work. Google says it burns about 17% fewer output tokens than 3.5 Flash on the Artificial Analysis Index, with reductions as high as 65% on some DeepSWE tasks, and it dropped the output price from $9 to $7.50 per million tokens while keeping input at $1.50.

Keep reading — it's free

Enter your email to keep reading — plus the best of AI & tech, daily. Free, forever.

Also get
or

Already a member? Sign in

Comments

Subscribe to join the conversation...

Be the first to comment