Skip to main content

Anthropic "Raised" Claude Code Limits by 25%. It's Actually a 17% Cut.

Anthropic
Photo by Vitaly Gorbachev on Pexels

Anthropic announced on September 14 that Claude Code weekly limits are getting a permanent 25% increase over the original baseline. The blog post framed it as a capacity boost. It is not. If you've been using Claude Code any time this summer, your effective weekly quota just dropped by roughly 17%. Here's the math that the headline buried.

What Actually Changed

Back in the spring, Anthropic ran a temporary 50% capacity boost on top of the original baseline — call that baseline 100 units. The temporary boost put you at 150 units per week. That ended on September 14.

The permanent change Anthropic announced replaces the temporary boost with a 25% permanent increase. So the new normal is 125 units — better than the original 100, but 17% less than the 150 you've been working with all summer.

Period Weekly Capacity vs. Original
Original baseline 100 units —
Summer 2026 (temporary boost) 150 units (+50%) +50%
From September 14, 2026 125 units (+25%) +25% vs. original, −17% vs. summer

Anthropic also doubled the five-hour rate limits earlier this year and confirmed those are not changing. So the hit is specifically to weekly capacity — the longer-horizon budget that determines whether you run out of quota mid-sprint. This applies to Pro, Max, Team, and seat-based Enterprise plans. The change is documented on Anthropic's announcements page.

The Numbers That Matter for Your Workflow

If your Claude Code usage was already bumping against limits during the summer, you'll feel this immediately. The typical pattern for heavy users is: front-load usage on Monday and Tuesday during intensive coding sessions, then hit the wall by Thursday. Under the temporary 50% boost, that wall was 17% further away than it is now.

For teams on seat-based Enterprise plans, the per-seat weekly budget shrinks accordingly. If you have a 10-person engineering team and each seat was running close to the 150-unit ceiling, the aggregate weekly throughput just dropped by roughly 170 units — the equivalent of losing more than one full seat's worth of capacity every week. That's not trivial if you've been structuring sprint workloads around the boosted availability.

What doesn't change: the five-hour rate limits (the short-burst throughput caps), and the availability of Claude Fable 5.1 for Max and Team subscribers. Anthropic added Fable 5.1 access to those tiers this month as a separate upgrade, so the capability level is higher even as the weekly budget tightens.

How to Manage the Tighter Budget

A few practical adjustments worth making before you run short mid-week:

  • Audit your heaviest tasks. Claude Code isn't equally expensive per prompt. Long context windows, multi-file edits, and repeated tool calls against large codebases burn through weekly quota faster than single-file questions. Identify those and batch them at the start of the week.
  • Move bulk automation to the API. If you're running Claude Code for scripted tasks — automated code review, batch refactors, test generation — shifting those to direct API calls with prompt caching can cut token costs sharply. On a workload where 80% of tokens are repeated context, caching can turn a $1,000 monthly bill into roughly $250, saving about $750 a month while freeing your interactive weekly quota for hands-on work.
  • Reserve interactive sessions for judgment work. Use your weekly Claude Code budget for the debugging, architecture, and review tasks where the back-and-forth actually pays off. Offload the predictable, repetitive jobs elsewhere.
  • Watch your Thursday cliff. If you historically ran out on Friday under the 150-unit ceiling, expect that to move up by roughly half a day. Plan the week's biggest tasks accordingly.

Why Anthropic Framed It This Way

Comparing the new limit to the original 100-unit baseline is technically accurate — it is a 25% increase over where things started. But almost nobody's frame of reference is the original baseline. The frame of reference is what you had last week, which was 150. Measuring a cut against a number users stopped experiencing months ago is a familiar move, and it's worth naming plainly: for anyone active this summer, this is a reduction dressed as an increase.

To be fair, temporary boosts are temporary by definition, and Anthropic is under no obligation to make them permanent. Compute is expensive, and a 25% permanent lift over baseline is a real, funded commitment rather than a promotional one. The problem is purely one of framing versus lived experience.

Bottom Line

If you're a light or occasional Claude Code user, this change is genuinely an upgrade — you're at 125 units instead of 100, and you probably never hit the summer ceiling anyway. For heavy daily users and engineering teams that calibrated their sprints against the boosted 150, this is a real 17% haircut on weekly capacity, and you should plan around it.

The practical response is the same regardless of how you feel about the messaging: audit your token-heavy tasks, push repetitive automation to the cached API where the savings can run into hundreds of dollars a month, and protect your interactive quota for the work that needs a human in the loop. Do that, and the tighter weekly budget stops being a wall you hit on Thursday and becomes a number you barely notice.

Comments

Popular posts from this blog

AWS vs Azure vs GCP in 2026: Which Cloud Platform Should You Choose?

The cloud platform decision is one of the most consequential technology choices an organization makes, and in 2026 it's also one of the most misunderstood. Most of the debate I see in enterprise architecture forums reduces to "we're an AWS shop" or "we go Azure because of Microsoft" — neither of which is a strategy. A platform choice made primarily on inertia or existing vendor relationships is a choice that will cost you for years. I've spent significant time in all three major cloud environments — AWS for scale workloads and data engineering, Azure for enterprise SAP and Microsoft-integrated architectures, and GCP for AI-intensive and analytics-heavy use cases. My goal in this guide is to give you a genuine, nuanced comparison that goes beyond feature lists and into the practical realities of choosing and running a cloud platform in 2026. I'll cover market position, each platform's honest strengths and weaknesses, how to match workloads t...

EU AI Act Compliance in 2026: What Every Enterprise Needs to Do Now

EU AI Act Compliance in 2026: What Every Enterprise Needs to Do Now The EU AI Act entered into force on August 1, 2024. The first provisions took effect six months later, and the full implementation timeline runs through 2027. If you're building, deploying, or using AI systems in or for the European Union, this law applies to you — and the window for being caught unprepared is closing fast. I've spent the past year working with enterprise clients on AI governance programs, and one pattern shows up again and again: organizations badly underestimate how much operational work compliance actually takes. It's not a checkbox exercise. It's a rethink of how you develop, document, deploy, and monitor AI systems. This guide is what I wish someone had handed me when I started — the substance of the law, the practical requirements, the deadlines that matter, and the mistakes I keep watching enterprises make. Photo by Petrit Nikolli on Pexels Photo by Karolina Gra...

GPT-6 Astra Is Here: $10/M Tokens, 100% on ExploitBench, and What It Actually Means for Developers

Photo by Michał Robak on Pexels Photo by Tara Winstead on Pexels OpenAI launched GPT-6 Astra on September 3rd, and unlike the usual cadence of incremental updates, this release ships with benchmarks that are hard to look past: 100% on ExploitBench, 98–99.9% on FrontierMath Tier 4 and ARC-AGI-3, and 72.6% on OSWorld 2.0. OpenAI is calling it their most capable model yet — and specifically their best for computer use, coding, and professional work. If you manage an API budget, the real question isn't whether the benchmarks look impressive. It's whether switching your workloads over saves money or burns it. Here's how the numbers actually shake out. The Numbers Behind the Launch Here are the headline specs from OpenAI's announcement: FrontierMath Tier 4: 98–99.9% (previous frontier models sat in the 60–70% range) ARC-AGI-3: 98–99.9% — a benchmark built specifically to resist memorization ExploitBench: 100% — which tripped OpenAI's Preparedness F...