First off - Opus 4.6 is impressive and the Claude Code team has been shipping at an incredible pace. This is meant as constructive feedback, not a complaint.
Opus 4.6 launched yesterday and the upgrade was automatic. Within hours, Max plan subscribers across Twitter and Reddit started reporting that their quotas were depleted far faster than with Opus 4.5 - in some cases before they even realized they'd been upgraded.
This isn't a bug report and I'm not claiming entitlement to higher quotas. This is feedback: the current quota allocations that worked fine for Opus 4.5 don't seem to be calibrated for 4.6's token consumption patterns. People who had sustainable workflows are now hitting walls.
Community reports (all from Feb 5-6, within 24h of launch)
Max plan users:
- "not even 12 hours and already at 20% weekly usage ON the $200 max plan [...] This model is token HUNGRY and may not be worth it if I run through limits like that" - @Dallenpyrah
- "After a full day of using Claude Code with Opus 4.6, the difference feels minimal [...] But man, 4.6 chugs tokens. My Claude Max account copped a beating" - @SeanDoesLife
- "claude code just upgraded to 4.6 without me changing anything. And it freaking eating all the tokens!" - @atla_
Comparative observations:
- "i never hit the limit with opus 4.5" - @Jungle_Fren
- "Opus 4.6 can use 20-30% more tokens. It is not as efficient as predecessor models" - @VQuinones
- Back-to-back testing on Reddit showing 6-8% session quota per prompt (4.6) vs ~4% (4.5) - r/ClaudeCode thread
Cross-platform:
The ask
If 4.6 is genuinely more capable and that capability costs more tokens, that's understandable. But it would help to:
- Acknowledge the consumption difference so users can make informed model choices
- Consider adjusting quotas for 4.6, or at minimum document the expected increase
- Don't auto-upgrade to a model that burns quota faster without warning
Thanks for building a great product. Happy to provide more data if helpful.
Related
First off - Opus 4.6 is impressive and the Claude Code team has been shipping at an incredible pace. This is meant as constructive feedback, not a complaint.
Opus 4.6 launched yesterday and the upgrade was automatic. Within hours, Max plan subscribers across Twitter and Reddit started reporting that their quotas were depleted far faster than with Opus 4.5 - in some cases before they even realized they'd been upgraded.
This isn't a bug report and I'm not claiming entitlement to higher quotas. This is feedback: the current quota allocations that worked fine for Opus 4.5 don't seem to be calibrated for 4.6's token consumption patterns. People who had sustainable workflows are now hitting walls.
Community reports (all from Feb 5-6, within 24h of launch)
Max plan users:
Comparative observations:
Cross-platform:
The ask
If 4.6 is genuinely more capable and that capability costs more tokens, that's understandable. But it would help to:
Thanks for building a great product. Happy to provide more data if helpful.
Related