Anthropic attributes Claude Code usage drain to peak-hour caps and large context windows
Anthropic has identified two primary causes for Claude Code users hitting usage limits faster than expected: stricter rate limiting during peak hours and sessions with context windows exceeding 1 million tokens. The company also recommends switching to Sonnet 4.6 instead of Opus, which consumes limits roughly twice as fast.
Anthropic Attributes Claude Code Usage Drain to Peak-Hour Caps and Expanding Contexts
Anthropichas identified the root causes behind complaints from Claude Code users depleting their usage limits faster than anticipated, according to Anthropic's Lydia Hallie.
The two primary factors are:
- Tighter peak-hour rate limits — Anthropic implements stricter usage caps during periods of high demand
- Ballooning context windows — Sessions with 1-million-token contexts or larger consume limits significantly faster
Bug Fixes and Billing Accuracy
Hallie confirmed Anthropic fixed several bugs but stated none resulted in incorrect billing. The company has deployed efficiency improvements and added in-product notifications to help users understand their usage patterns better.
Anthropic's Recommendations
To manage usage more effectively, Hallie recommends:
- Switch to Sonnet 4.6 instead of Opus, which consumes limits approximately twice as fast
- Disable Extended Thinking when not required for specific tasks
- Start fresh sessions rather than continuing previous conversations to reduce accumulated context
- Limit context window size to reduce per-request token consumption
Users experiencing usage depletion they believe is unusual should report the issue through the in-product feedback function.
What This Means
The clarification addresses a widespread concern among Claude Code users and reveals that usage consumption is largely a function of rate limiting mechanics and user behavior rather than billing errors. The recommendation to use Sonnet 4.6 over Opus suggests a significant performance-per-token trade-off between the two models. Anthropic's focus on user transparency through pop-ups and clearer guidance indicates the company is attempting to manage expectations around usage consumption before users encounter hard limits.
Related Articles
Anthropic adds Mods to Claude Code, a plugin system that hooks into tool calls, prompts and UI rendering
Anthropic released Mods for Claude Code, a plugin system built on JavaScript and TypeScript functions that hook into events such as tool calls, user prompts and UI rendering. Mods are not sandboxed and run with the user's permissions. They work in the CLI, the desktop app and, partly, the VS Code extension.
AWS adds managed Web Search to Claude Desktop via Bedrock AgentCore Gateway in three Regions
AWS published a walkthrough for connecting Claude Desktop on Amazon Bedrock to a managed, MCP-compatible Web Search capability through Amazon Bedrock AgentCore Gateway. According to AWS, the search is backed by an Amazon web index spanning tens of billions of documents, and query traffic stays within AWS infrastructure. Web Search is available in three AWS Regions; pricing is not disclosed in the post.
Graphite: Opus 5.5 uses 'this matters' 116x more than humans as AI writing tells persist
Marketing firm Graphite identified 13,000 phrases that appear at least twice as often in AI-generated writing as in human writing. Claude Opus 5.5 uses "this matters" 116 times more than humans, while OpenAI's Astra favors "corrective framing" more than 100 times as often. Em-dash use has collapsed across frontier models, but total tells are holding steady, according to Graphite.
Anthropic launches Claude for Government for US civilian agencies in FedRAMP High environment
Anthropic is now offering Claude for Government to US federal and state agencies. The platform has been in open beta since July and runs in a FedRAMP High environment. The launch comes as the company's legal fight with the Pentagon continues.
Comments
Loading...