Skip to content
Fable 5 Effort: Cut Thinking Token Costs in Python — ContentBuffer guide

Fable 5 Effort: Cut Thinking Token Costs in Python

K
Kodetra Technologies··9 min read Intermediate

Summary

Claude Fable 5 always thinks. Use effort, display and max_tokens to control reasoning cost.

Fable 5 Effort: Cut Thinking Token Costs in Python

Anthropic shipped Claude Fable 5 on June 9, 2026, and it is now the most capable model the company offers to everyone. It also behaves differently from every Claude before it in one way that catches teams off guard the first week: thinking is always on and you cannot turn it off. Send a one-line prompt and the model may still burn hundreds of reasoning tokens before it answers. At $50 per million output tokens, and with thinking tokens billed as output, that surprise lands on your invoice.

Keep reading — it's free

Enter your email to keep reading — plus the best of AI & tech, daily. Free, forever.

Also get
or

Already a member? Sign in

Comments

Subscribe to join the conversation...

Be the first to comment