
Fable 5 Effort: Cut Thinking Token Costs in Python
K
Kodetra Technologies··9 min read Intermediate Summary
Claude Fable 5 always thinks. Use effort, display and max_tokens to control reasoning cost.
Fable 5 Effort: Cut Thinking Token Costs in Python
Anthropic shipped Claude Fable 5 on June 9, 2026, and it is now the most capable model the company offers to everyone. It also behaves differently from every Claude before it in one way that catches teams off guard the first week: thinking is always on and you cannot turn it off. Send a one-line prompt and the model may still burn hundreds of reasoning tokens before it answers. At $50 per million output tokens, and with thinking tokens billed as output, that surprise lands on your invoice.
Keep reading — it's free
Enter your email to keep reading — plus the best of AI & tech, daily. Free, forever.
Also get
or
Already a member? Sign in
Comments
Subscribe to join the conversation...
Be the first to comment