<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Claude Code Usage Credits on RockB</title><link>https://baeseokjae.github.io/tags/claude-code-usage-credits/</link><description>Recent content in Claude Code Usage Credits on RockB</description><image><title>RockB</title><url>https://baeseokjae.github.io/images/og-default.png</url><link>https://baeseokjae.github.io/images/og-default.png</link></image><generator>Hugo</generator><language>en-us</language><lastBuildDate>Thu, 01 Oct 2026 04:58:20 +0000</lastBuildDate><atom:link href="https://baeseokjae.github.io/tags/claude-code-usage-credits/index.xml" rel="self" type="application/rss+xml"/><item><title>Claude Code Limits in 2026: The 50% Promotion Ended, the 25% Raise Stayed</title><link>https://baeseokjae.github.io/posts/claude-code-weekly-limits-promotion/</link><pubDate>Thu, 01 Oct 2026 04:58:20 +0000</pubDate><guid>https://baeseokjae.github.io/posts/claude-code-weekly-limits-promotion/</guid><description>The 50% Claude Code limits promotion ended Sep 13, 2026. A permanent 25% raise replaced it — here is what heavy users keep and how to stretch it.</description><content:encoded><![CDATA[<p>Claude Code limits today are permanently 25% higher than the pre-promotion baseline for Pro, Max, Team, and seat-based Enterprise plans — but roughly 17% lower than the 50% promotional peak that ran from May 13 to September 13, 2026. Anthropic kept part of the boost, ended the rest, and never touched the separate five-hour session limit. That single paragraph is the whole story, and it is the reason so many heavy users simultaneously feel &ldquo;the promotion ended&rdquo; and &ldquo;my plan got better this year.&rdquo;</p>
<p>Below is the exact math, what actually drains a weekly allowance in October 2026, and the levers — official ones, not folklore — that stretch it.</p>
<h2 id="what-changed-the-2026-claude-code-limits-timeline">What changed: the 2026 Claude Code limits timeline</h2>
<p>Anthropic changed Claude Code usage limits three separate times in 2026. Heavy users who noticed only one of those changes have a distorted picture of what they are entitled to.</p>
<table>
  <thead>
      <tr>
          <th>Date</th>
          <th>Change</th>
          <th>Who it applied to</th>
          <th>Source</th>
      </tr>
  </thead>
  <tbody>
      <tr>
          <td>May 6, 2026</td>
          <td>Five-hour window limits doubled; peak-hours limit reduction removed</td>
          <td>Pro, Max (tied to a SpaceX Memphis data-center compute deal)</td>
          <td>arstechnica.com</td>
      </tr>
      <tr>
          <td>May 13, 2026</td>
          <td>Weekly limits raised 50% (promotion begins)</td>
          <td>Pro, Max, Team, legacy seat-based Enterprise</td>
          <td>support.claude.com</td>
      </tr>
      <tr>
          <td>Sep 13, 2026</td>
          <td>Promotion expires at 11:59 PM PT</td>
          <td>—</td>
          <td>support.claude.com</td>
      </tr>
      <tr>
          <td>Sep 14, 2026</td>
          <td>Weekly limits permanently raised 25% above the pre-promotion baseline</td>
          <td>Pro, Max, Team, legacy seat-based Enterprise</td>
          <td>support.claude.com</td>
      </tr>
  </tbody>
</table>
<p>Two of those three changes are still in force. The five-hour doubling from May 6 was never rolled back, and the permanent 25% weekly raise from September 14 is now the standard. Only the 50% promotional window closed.</p>
<p>The promotion itself ran automatically. There was no opt-in, no coupon, and no action required: if you were on an eligible subscription between May 13 and September 13, 2026, your weekly Claude Code allowance was 50% larger, and <code>/usage</code> in the CLI showed the elevated numbers. Free plans and consumption-based Enterprise seats were excluded from the start, because neither of them meters usage the way a subscription allowance does.</p>
<p>If you want the longer version of how that promotional window was announced and extended through the summer, our <a href="/posts/claude-code-weekly-limits-promotion-2026/">earlier breakdown of the weekly limits promotion</a> walks through the announcement-by-announcement sequence.</p>
<h2 id="what-the-50-promotion-actually-covered--and-what-it-never-touched">What the 50% promotion actually covered — and what it never touched</h2>
<p>Misreading the scope is the most common reason heavy users misjudge their headroom. The promotion was narrower than &ldquo;everything got 50% better.&rdquo;</p>
<p><strong>What was included.</strong> The boost applied to the weekly usage allowance in Claude Code across every surface it runs on: the CLI, IDE extensions, the desktop app, and the web/agent experience. It covered Pro, Max, Team, and legacy seat-based Enterprise.</p>
<p><strong>What was excluded.</strong> Free plans have no Claude Code allowance to boost, and consumption-based Enterprise seats are metered on consumption rather than on a subscription pool, so neither was eligible.</p>
<p><strong>What was never part of it.</strong> The five-hour session limit. That meter was doubled on May 6 and left alone by the promotion. Claude chat and Claude Cowork limits were also untouched — the extra 50% was Claude Code only, in every surface Claude Code runs on, but nowhere else.</p>
<p>That last point resolves a genuinely common confusion. If your Claude Code weekly pool felt enormous in July while your Claude chat allowance in the browser felt unchanged, that was by design, not by bug.</p>
<h2 id="the-honest-math-a-25-raise-and-a-17-cut-at-the-same-time">The honest math: a 25% raise and a 17% cut at the same time</h2>
<p>The numbers are only confusing if you mix up the baseline. Normalize your pre-promotion allowance to 100 units — the units are opaque, since Anthropic does not publish token counts for subscription plans:</p>
<table>
  <thead>
      <tr>
          <th>Period</th>
          <th>Allowance index</th>
          <th>vs. pre-May baseline</th>
          <th>vs. promotional peak</th>
      </tr>
  </thead>
  <tbody>
      <tr>
          <td>Pre-May 2026 baseline</td>
          <td>100</td>
          <td>—</td>
          <td>−33%</td>
      </tr>
      <tr>
          <td>Promotion (May 13 – Sep 13)</td>
          <td>150</td>
          <td>+50%</td>
          <td>—</td>
      </tr>
      <tr>
          <td>Since Sep 14, 2026</td>
          <td>125</td>
          <td>+25%</td>
          <td>−16.7% (≈17%)</td>
      </tr>
  </tbody>
</table>
<p>Anthropic conceded the framing publicly: &ldquo;Compared to today, this works out to a 17% reduction in weekly limits on Claude Code.&rdquo; That statement is accurate relative to the promotional peak and simultaneously consistent with the permanent 25% raise relative to the original baseline. Both descriptions point at the same number, 125.</p>
<p>The practical consequence for a heavy user is that you should stop sizing your workflow around August&rsquo;s ceiling. Your real budget is 25% above where you started the year and about a sixth below the temporary high point — and the weekly allowance was never a fixed count of prompts anyway. It varies with conversation length, model choice, tool usage, and effort level, which is why two developers on identical plans can have wildly different limits remaining at the end of a week.</p>
<h2 id="how-claude-code-limits-actually-work-in-october-2026">How Claude Code limits actually work in October 2026</h2>
<p>&ldquo;Claude Code limits&rdquo; is shorthand for at least four distinct meters that behave differently and reset on different clocks.</p>
<table>
  <thead>
      <tr>
          <th>Meter</th>
          <th>What it governs</th>
          <th>Reset cadence</th>
          <th>Affected by the promotion?</th>
      </tr>
  </thead>
  <tbody>
      <tr>
          <td>Weekly allowance</td>
          <td>Your plan&rsquo;s overall budget across Claude Code</td>
          <td>Fixed weekly window</td>
          <td>Yes — now +25% permanently vs pre-May</td>
      </tr>
      <tr>
          <td>Five-hour session limit</td>
          <td>How much you can run in one working session</td>
          <td>Every five hours</td>
          <td>No — doubled May 6, unchanged by the promotion</td>
      </tr>
      <tr>
          <td>Per-model weekly caps</td>
          <td>Model-specific weekly ceilings (e.g. Fable 5.1)</td>
          <td>Weekly, tied to the model</td>
          <td>Not as a headline change, but new models moved the math</td>
      </tr>
      <tr>
          <td>Length limit (context)</td>
          <td>How large the conversation can grow before degradation</td>
          <td>Per session, managed by you</td>
          <td>No — this is the context window, not your plan allowance</td>
      </tr>
  </tbody>
</table>
<p>Two structural facts deserve emphasis. First, on subscription plans all Claude surfaces draw on one shared usage limit: claude.ai, Claude Code, and Claude Desktop are not separate pools. Second, API keys are metered per token and never touch your subscription allowance — so &ldquo;I&rsquo;ll just switch to an API key&rdquo; is a billing change, not a way to unlock more subscription headroom.</p>
<p>The distinction between the length limit and the usage limit is the one heavy users most often conflate. The context window is a technical ceiling you manage with <code>/compact</code> and <code>/clear</code>; the plan allowance is a commercial budget you can only refill by waiting, paying, or upgrading.</p>
<h2 id="why-your-weekly-limit-drains-faster-than-you-expect">Why your weekly limit drains faster than you expect</h2>
<p>If your allowance is 25% larger than in January but still evaporates by Tuesday, the promotion is not the culprit. Five documented mechanisms are.</p>
<p><strong>1. Prompt-cache expiry forces re-caching.</strong> When the prompt cache expires mid-session, Claude Code re-reads context it had already paid for. Developers have reported single re-cache events of roughly 800,000 tokens, charged against both the five-hour and the weekly meters.</p>
<p><strong>2. Subagent fan-out multiplies the burn.</strong> A parent agent delegating to many subagents consumes allowance in parallel rather than serially. One report saw roughly 70% of a weekly Fable 5.1 allowance consumed mostly by subagents that were then killed by the rate limit; another burned 39% of a Max 20x weekly limit in under 24 hours by running about 16 maximum-effort reviewer subagents.</p>
<p><strong>3. Long-lived sessions accumulate context.</strong> Because drain scales with conversation length, a session left open across days is the most expensive way to work. This is a hygiene problem, not a plan problem.</p>
<p><strong>4. Per-model weekly caps are invisible until they bite.</strong> The status-line JSON exposes only <code>five_hour</code> and <code>seven_day</code> windows, so a per-model weekly limit — Fable 5.1, for instance — cannot be seen there at all. As of September 29, 2026 that visibility gap was still an open feature request.</p>
<p><strong>5. The meter is shared and your model mix changed.</strong> Anthropic shipped Fable 5.1 and Mythos 5.1 on September 1, Opus 5.5 on September 22, and Sonnet 5.5 on September 28, 2026. New models on the same subscription change the effective cost per task even when the headline allowance does not move.</p>
<p>Two field reports put scale on the problem. A Max 20x user burned 9.6 billion tokens across 34 sessions in seven days (September 16–22, 2026), exhausted the weekly limit in about 2.5 days, and did 93.9% of that work on Opus 5. Separately, one developer measured a roughly 3.6x worse weekly consumption rate after the September 25, 2026 reset — about 93 responses per 1% before, roughly 30 after — on an unchanged workflow.</p>
<h2 id="seven-tactics-that-actually-stretch-a-weekly-allowance">Seven tactics that actually stretch a weekly allowance</h2>
<p>Anthropic&rsquo;s own reduction playbook is the honest starting point, because it targets consumption rather than the ceiling: manage context proactively, choose the right model, cut MCP overhead, move standing instructions from CLAUDE.md into skills, adjust extended thinking, delegate verbose work to subagents deliberately, and manage agent-team token costs. Here is what those translate into in practice.</p>
<ol>
<li><strong>Run <code>/compact</code> before you think you need it.</strong> Compaction summarizes prior turns into a shorter form. Waiting until the context is full means you already paid for the bloat.</li>
<li><strong>Run <code>/clear</code> between unrelated tasks.</strong> Carrying yesterday&rsquo;s context into today&rsquo;s task is the single largest avoidable drain.</li>
<li><strong>Default routine I/O work to a cheaper model and reserve the premium model for reasoning.</strong> Spotify&rsquo;s engineering team documented a 90% cut in Claude Code token usage by routing non-reasoning &ldquo;I/O&rdquo; work — reading files, generating boilerplate, updating docs — to a cheaper worker model. Most agent work is I/O, not reasoning.</li>
<li><strong>Cap subagent fan-out deliberately.</strong> Give subagents narrow scopes and a maximum-effort ceiling that matches the task. Sixteen max-effort reviewers is a budget decision, not a quality decision.</li>
<li><strong>Watch MCP overhead.</strong> Every connected server&rsquo;s tool definitions occupy context on every turn. Disable the servers you are not using in a given session.</li>
<li><strong>Move instructions out of CLAUDE.md and into skills.</strong> Standing instructions load constantly; skills load on demand.</li>
<li><strong>Right-size extended thinking.</strong> Maximum effort on a task that does not need it is the cheapest way to burn a weekly allowance quickly.</li>
</ol>
<p>Two older habits are now obsolete: there is no peak-hours reduction to schedule around (removed May 6, 2026), and switching to an API key is a billing model change rather than a way to extend your subscription pool.</p>
<h2 id="when-you-hit-the-cap-wait-usage-credits-or-upgrade">When you hit the cap: wait, usage credits, or upgrade</h2>
<table>
  <thead>
      <tr>
          <th>Option</th>
          <th>How it works</th>
          <th>Cost signal</th>
          <th>Best when</th>
      </tr>
  </thead>
  <tbody>
      <tr>
          <td>Wait for reset</td>
          <td>Weekly window resets on a fixed cadence; five-hour window resets continuously</td>
          <td>Free</td>
          <td>You hit the cap sporadically</td>
      </tr>
      <tr>
          <td>Enable usage credits</td>
          <td>Keep working past included limits, billed at standard API rates</td>
          <td>API rates; $2,000/day redemption limit; monthly spend cap and auto-reload configurable</td>
          <td>You hit the cap occasionally and the overage is small</td>
      </tr>
      <tr>
          <td>Upgrade the plan</td>
          <td>Move to a larger multiplier (Max 5x → Max 20x)</td>
          <td>Fixed monthly price</td>
          <td>You hit the cap during normal, non-exceptional work</td>
      </tr>
  </tbody>
</table>
<p>Usage credits are available to Pro, Max 5x, and Max 20x subscribers. You enable them under Settings → Usage → Usage credits on claude.ai and prepay (&ldquo;Add funds&rdquo;), with optional auto-reload. Note that credits are billed separately from the subscription and appear as additional charges, and the daily redemption limit is $2,000. If you enable credits and then upgrade plans, verify the meter after switching — reported upgrade-timing bugs mean the displayed allowance is worth re-checking rather than assuming.</p>
<h2 id="which-plan-a-heavy-user-should-actually-buy">Which plan a heavy user should actually buy</h2>
<p>Pricing: Claude Pro is $20/month ($17/month billed annually); Max starts at $100/month with 5x and 20x tiers. The official Enterprise benchmark for Claude Code usage is about $13 per developer per active day and $150–250 per developer per month, with 90% of users below $30 per active day.</p>
<table>
  <thead>
      <tr>
          <th>Plan</th>
          <th>Price</th>
          <th>Claude Code capacity</th>
          <th>Fits</th>
      </tr>
  </thead>
  <tbody>
      <tr>
          <td>Free</td>
          <td>$0</td>
          <td>None</td>
          <td>Claude chat only; Claude Code excluded</td>
      </tr>
      <tr>
          <td>Pro</td>
          <td>$20/mo ($17 annual)</td>
          <td>Baseline multiplier; modest weekly pool</td>
          <td>A few hours a day, mostly Sonnet</td>
      </tr>
      <tr>
          <td>Max 5x</td>
          <td>$100/mo</td>
          <td>5x Pro-class capacity</td>
          <td>Daily work on large repos with Opus in the loop</td>
      </tr>
      <tr>
          <td>Max 20x</td>
          <td>$200/mo</td>
          <td>20x Pro-class capacity</td>
          <td>Continuous or concurrent batch work</td>
      </tr>
      <tr>
          <td>Team / seat-based Enterprise</td>
          <td>Per seat</td>
          <td>Seat multipliers; weekly limits now +25% vs pre-May</td>
          <td>Teams needing shared, governed capacity</td>
      </tr>
  </tbody>
</table>
<p>The economics get uncomfortable fast at the top. A quarter of engineering leaders already report spending $200–500 per developer per month on tokens, with some above $2,000, and by 2028 AI coding costs are projected to exceed the average developer salary. That is the argument for treating model routing (tactic 3 above) as a budget control rather than a preference: if a Max 20x subscription is not enough, the correct first fix is usually to stop spending premium-model tokens on non-reasoning work, not to buy a bigger plan.</p>
<h2 id="how-to-monitor-claude-code-usage--and-what-each-tool-cannot-see">How to monitor Claude Code usage — and what each tool cannot see</h2>
<table>
  <thead>
      <tr>
          <th>Tool</th>
          <th>What it shows</th>
          <th>What it misses</th>
      </tr>
  </thead>
  <tbody>
      <tr>
          <td><code>/usage</code></td>
          <td>Plan usage bars, Day/Week toggle, activity stats, usage breakdown</td>
          <td>Per-model weekly caps; the &ldquo;Session&rdquo; block is API-oriented, not your subscription allowance</td>
      </tr>
      <tr>
          <td><code>/usage</code> (rate-limited)</td>
          <td>Last-known bars from the past 60 minutes with a &ldquo;Showing last-known usage&rdquo; note (v2.1.208+); press <code>r</code> to retry</td>
          <td>Live numbers while the endpoint is throttled</td>
      </tr>
      <tr>
          <td><code>/usage-credits</code></td>
          <td>Opens the right settings page by role (v2.1.211+; v2.1.248+ for Enterprise)</td>
          <td>Only present while credits are enabled</td>
      </tr>
      <tr>
          <td><code>/insights</code></td>
          <td>Analyzes up to 200 recent sessions, writes ~/.claude/usage-data/report.html</td>
          <td>Nothing below the session level</td>
      </tr>
      <tr>
          <td>Status line</td>
          <td><code>five_hour</code> and <code>seven_day</code> windows</td>
          <td>Any per-model weekly window (e.g. Fable 5.1)</td>
      </tr>
      <tr>
          <td>Third-party trackers</td>
          <td>Local trend lines and burn-rate estimates</td>
          <td>Anything the CLI does not expose; per-model caps remain invisible</td>
      </tr>
  </tbody>
</table>
<p>The honest summary: no single surface shows your whole budget in October 2026. The weekly bars are authoritative, the status line is structural only, and per-model weekly limits stay invisible until they stop you.</p>
<h2 id="frequently-asked-questions">Frequently Asked Questions</h2>
<p><strong>Are Claude Code weekly limits higher now than they were before the promotion?</strong>
Yes. Since September 14, 2026, weekly limits are permanently 25% higher than the pre-promotion baseline for Pro, Max, Team, and legacy seat-based Enterprise plans. They are simultaneously about 17% lower than the promotional peak that ran May 13 to September 13, 2026.</p>
<p><strong>Is the 50% Claude Code promotion still active?</strong>
No. The promotion expired at 11:59 PM PT on September 13, 2026, and was replaced the next day by a permanent 25% raise to standard weekly limits. Anthropic itself described the change as &ldquo;a 17% reduction in weekly limits on Claude Code&rdquo; relative to the promotional period.</p>
<p><strong>When does the Claude Code weekly limit reset?</strong>
The weekly allowance resets on a fixed weekly window, and the five-hour session limit resets continuously every five hours. Per-model weekly caps follow their own weekly schedule, which the status line does not display — run <code>/usage</code> for the windows it can show.</p>
<p><strong>How do I check my Claude Code usage?</strong>
Run <code>/usage</code> in the CLI for plan usage bars with a Day/Week toggle, activity stats, and a breakdown. If the usage endpoint is rate limited, v2.1.208+ shows last-known bars from the past 60 minutes with a &ldquo;Showing last-known usage&rdquo; note — press <code>r</code> to retry. <code>/usage-credits</code> opens the credit settings page when credits are enabled, and <code>/insights</code> writes a report from up to 200 recent sessions.</p>
<p><strong>Do subagents and <code>/compact</code> affect my weekly limit?</strong>
Yes to subagents, and that is often the surprise. Subagents bill against the same weekly allowance, and reported cases include roughly 70% of a weekly allowance consumed mostly by subagents, plus 39% of a Max 20x weekly limit in under 24 hours via about 16 max-effort reviewer subagents. <code>/compact</code> works the opposite way: it reduces future drain by shrinking accumulated context, and <code>/clear</code> starts a clean, cheaper session.</p>
<h2 id="bottom-line">Bottom line</h2>
<p>The 50% era is over and the 25% permanent raise is what you actually have. Anthropic&rsquo;s five-hour doubling stayed, the weekly promotion did not, and the meters that will stop you now are the ones almost nobody can see: subagent fan-out, prompt-cache re-reads, long-lived contexts, and per-model weekly caps. Keep the hygiene — <code>/compact</code>, <code>/clear</code>, cheaper models for I/O work, deliberate subagent budgets — and treat usage credits and the Max 5x → 20x step as escalations rather than defaults. Run <code>/usage</code> after any plan change and verify the number instead of trusting it.</p>
]]></content:encoded></item></channel></rss>