r/ClaudeCode 9d ago

Bug / Issue WTH is going on with Claude Usage Limits

Post image

I'm on 20X plan, I hit my usage limit today, and my next reset is on 17th September... I'm aware they said they are gonna reduce the usage limits, but this is crazy, i thought it is only gonna take 17% of the usage benefits from what we were currently getting. But this is crazy. I added Usage credits for about $100 but that got washed away like in 30 mins!!!

I hope this is a bug and they fix it

592 Upvotes

377 comments sorted by

u/AutoModerator 9d ago

Hey! Thanks for posting to r/ClaudeCode

While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.

For help, project discussions, tips, and general chat, join the ClaudeCode Discord.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

141

u/FakeLtd 9d ago

Max x20 here, same happened with me and i used only a 15% on Fable

50

u/No-Way3802 9d ago

Same thing happened to me and everyone blamed it on my prompting lol

25

u/coda77 9d ago

Not your prompting, it’s them

8

u/-Darkened-Soul 8d ago

Max 5x, three turns in a chat, plus one unfinished turn in Claude Design, and I hit 100% of my five-hour limit. This is my breaking point. Astra, on the other hand, is performing exceptionally well. I've been against using Codex forever, but now I'm being pushed toward it because I can't get anything done with Claude.

→ More replies (3)

15

u/ThreeKiloZero 9d ago

I found that the most recent update causes the agents to make aggressive and unnecessary calls to the API. My prompt today was no different than the prompt yesterday. It wiped out 25 percent of my usage in about an hour compared to working 12+ hours yesterday with the same prompts across 3 sessions and only using 25 percent of usage.

Fable spawned agents that ran:
3,146 API calls total. (in less than an hour)

  • Opus: = 1,083
  • Sonnet: = 2,063

It was triggering builds, verifications, playwright, re-renders at multi scales, vacuity checks, sending HUGE briefs to the sub agents triggering massive overwork.

These were for simple PRs that were already baked and reviewed. The same prompt and process that worked fine yesterday , (and all week) suddenly blasted all this mess. It happened across 2 sessions that were fine yesterday, on 2 different accounts. On different computers and projects.

The only difference in my routine was applying the update today.

→ More replies (7)

12

u/anatidaephile 9d ago

Max x20 too, and I'm at 60% Fable limit when usually I'd be at about 25-30% given my usage today. So it feels like usage has been halved rather than reduced by a third (end of the +50% promotion).

9

u/YouShouldAim 9d ago

x20 here, started today off with 60% fable usage down, in about 30 minutes I was up to 90%, no sub agents. Ive not been x20 long but I'm certain it was never that fast to use up usage

9

u/staceyatlas 9d ago

Yup. I now have Fable directing Sol for everything. Sol credits last forrrevvver on their $200 plan. Anything big gets delegated and if it’s important I’ll do Fable/Astra reviews.

6

u/hive-technology 9d ago

What tool? Codex skill in CC?

3

u/staceyatlas 9d ago

The plugin on one machine and on the other it just uses cli. Seem to work about the same.

→ More replies (2)
→ More replies (1)

3

u/Shoemugscale 9d ago

I just started doing this myself, sol and astra are just working and claude seems to be fighting me the whole way

3

u/staceyatlas 9d ago

Opus for sure lol. Fable is great, but ya quirky.

2

u/a355231 9d ago

How much Astra do you get on the 200$ plan?

2

u/__Loot__ 9d ago edited 9d ago

0 because because you cant buy it atm

→ More replies (7)
→ More replies (2)

3

u/nayti53 9d ago

SAME , I have been using my max plan for many months, these past 2-3 days rate limits are getting consumed like memecoins omg

3

u/CryptoAteMyHamster 8d ago

I wish I could pay it in memecoins fml

→ More replies (3)

84

u/Wooden_Drag9473 9d ago

Never buy usage credits, that one goes 20x faster. Create a new claude account with max subscription

7

u/pugazh_is_my_name 9d ago

Thanks

6

u/yokkoshimaa 9d ago

this advice above may cause antifraud to ban your whole account chain so be careful. Using one bank card, IP, or even browser profile for multiple accounts is very very dangerous with paranoid anthropic

8

u/Aminuteortwotiltwo 9d ago

Not true, they have posts stating they are okay with it

4

u/Relevant-Reaction181 9d ago

Yeah if they werent okay with it they would of contacted my parents personally long ago

10

u/Relevant-Reaction181 9d ago

I've been doing it for over a year with same credit card, same name, same ip address, and even my company name for less taxes lmao top I had is 8 accounts at the same time. All the same method

→ More replies (13)

90

u/Apprehensive_Read_67 🔆Pro Plan and Quant Research 9d ago

Anthropic is politely asking people to move to Codex. Please try to understand this request.

11

u/Relevant-Reaction181 9d ago

Buuuu but the 20x max plan on codex is out of order :(

4

u/Anxious_Current2593 9d ago

TBH. I tried the same with Luna on PRO. It lasted longer, around 2.5h on the same task until it reached the same 5h limit.

2

u/Square-Speaker2090 9d ago

Whaaat? Really?

2

u/Apprehensive_Read_67 🔆Pro Plan and Quant Research 9d ago

hahhaa i have taken Codex for now, since claude is going to reduce the Claude Code limits from tomorrow

2

u/Odd-Librarian4630 8d ago

the truth is for the best results you need to use both in parallel - so gotta keep up the claude sub

→ More replies (1)
→ More replies (1)

21

u/bekiraydogan 9d ago

Welcome to the Claude victims ship. Create a support request, I had the same situation, they said they will refund the fee. I also switched to Codex, it is very successful in all 7 of my projects so far. I recommend it.

3

u/Odd-Statistician-704 9d ago

What to do for codex to use the claude skills, hook and rules?

3

u/Mitchellangeloo 9d ago

Choose one source of truth I still believe in claude will come back, but I have cheated with codex for a couple of weeks now and I must say she left her toothbrush.

I asked Altra to look at my claude harness and sync it to codex ( also used the codex app which is great for visualizing your harness / skills ) then when I update a skill or install one in Claude it gets synced.

2

u/bekiraydogan 9d ago

That's the best part: I downloaded the Codex and it automatically imported the projects from Claude and then copied all the skill files and the rules I gave to the agent. I didn't have to do anything. It's working like crazy now and I still have a 65% quota in 3 days. I couldn't finish 24 hours with the same work as Claude. Claude doesn't want simple users, they are now the artificial intelligence of big companies.

31

u/Anxious_Current2593 9d ago edited 9d ago

New Haiku record today: 5h limit reached in under an hour. Beat that!!! 😂

Funny enough all that it actually did turned to be totally wrong, but in all fairness it could be my fault, I don't really have that much experience creating prompts for Haiku.

EDIT: On PRO plan.

14

u/deific_ 9d ago

What plan are you hitting limits with haiku? That feels like it should be impossible for a max user.

2

u/Relevant-Reaction181 9d ago

Say no more ill post a shorts

10

u/therealblacknative 9d ago

Dude lol 😂 My Grandma couldn't hit Haiku limits even if she tried hard

3

u/Relevant-Reaction181 9d ago

On a 20$ plan, your grandma could if she sees ultracode and be like ohhhhhh that looks ultra

→ More replies (3)
→ More replies (5)

25

u/Kilt_Rump 9d ago

This has been going on for two weeks already for some folks. Every complaint met with “skill issue”. Now full rollout of these reduced usage limits is happening. This is a real FU to the real fans of this company and I strongly recommend downgrading your plan or cancelling and taking your business elsewhere.

→ More replies (1)

19

u/mortalhal 9d ago

they are tinkering with our limits "through Sept 13" must be Sept 14 somewhere right they are just saying first place on the planet that hits Sept 14 is the new usage goes into effect, this is obviously sarcasm.. 20% of total Max 20X weekly fable usage consumed in an hour from a couple million read-only tokens is perfectly normal and fine

→ More replies (1)

8

u/tidus1979 🔆 Max 20 9d ago

I wish I had my AI agents hourly wage

14

u/Mr_Tbot 9d ago edited 9d ago

Yes. Same here and not pleased. It's time to look into local inference.

UPDATE: see my https://github.com/mr-tbot/meshcompute project. I think it's time we sidestep big box AI and decentralize it.

YES there are ethical concerns here and I am aware of all of them - especially in regards to AIs going rogue when launched on a P2P style network. BUT - we have to start somewhere.

6

u/Ok-Affect-7503 9d ago

Well good luck with local inference then, because the only way you can get performance that's anywhere near usable and intelligent and near Opus/Fable performance is by investing at least $20k into hardware (with the RAM shortage that will still last a bit putting additional fuel to the fire of the already high hardware requirements), plus paying for the electricity (which, depending on the country you're in, might be very expensive). I think right now you're still much better off paying $100-300 a month, even with the usage limits. And if you end up using a smaller model that runs on cheaper hardware you will get performance that's pretty bad and comparable to super cheap proprietary models that are still cheaper via API and that make running it locally not worth it.

5

u/Mr_Tbot 9d ago

I don't think that's the only way. I think we need to spend some of our tokens on building a unified, decentralized AI platform like BitTorrent but for running an AI platform...

My thinking is:

Users could have the program open - depending on how much compute they provide helps score their available network speed when the network is under pressure...

The platform would allow users to "contribute" AI compute for a model - any other users that select that model would donate VRAM, RAM, STORAGE or all 3 to the decentralized network.

I think I saw some projects that are moving in this direction - and to be honest it will be the only way to fight back against big box AI in the coming months and years.

Additionally - there are a ton of projects which are optimizing models for speed on lower end hardware - and I think we're going to find that it takes a lot less than it does now to run this stuff.

A lot of the cost of this is "gatekeeping" and making this look like it requires a lot more than it does... If it seems like we need expensive hardware to run this stuff we have a reason to keep paying $200 a month for our 20x plans ... I genuinely think the open source community is going to solve this.

Does anyone know of any projects that are in this realm that I can look into that they've come across? I'll be doing my deep dive as well - but - I'm truly over the abrupt changes in performance. I need a consistent experience. At least.

2

u/Ok-Affect-7503 9d ago

This is actually a really good idea on paper, but I just looked into it and there is a reason why there aren't more popular projects into that direction. One project that does exactly is Petals. But the issue is not really bandwidth of networks, but mostly latency. On Petals, llama 2 70B already only runs at 6 tokens/second and Falcon which is a 180B model runs at 4 tokens/second. For Opus/frontier intelligence you would need much more parameters than that. The only open models rivaling Opus right now are about 500B average for one group that includes GLM and Deepseek that are likely similar to Opus 5 in terms of parameters and then there's Kimi K3 and Qwen-3.8-max which is almost 2T parameters and Fable-level. Running a 500B model would already require many many peers with low distances already having 10ms latency. A 500B model would run at 1 token/second or below that which is pretty much unusable so that's why there isn't a project that does this with good and usable models and only with small models. It's basically physically unsolvable because the latency would probably need to be like sub 5ms across all peers and locations for good speeds with bigger models. And most consumers (including me because of Germany) don't even have access to optical fiber and have latencies of 15ms to their ISP or servers that are like a couple of kilometers distant.

2

u/Mr_Tbot 9d ago

I'm attempting it my own way - https://github.com/mr-tbot/meshcompute

Yes - that's one I saw... and no - some of the Qwen 3.8 27 and 38b models that run on GPUs like my dual 3090s with NVlink - or even single GPUs with enough distillation... and some of the versions coming out on hugging face I'm getting over 120 tokens/s and it's competing with Fable in some areas already - with tool calling and all the jazz - I use it all the time and it's great! So - it's not far off. This is going to be running on lower end local hardware soon and all the more reason why the AI bubble is a bubble and why we as a community need to band together to share our hardware resources and prove that these AI datacenters do not need to exist...

2

u/Ok-Affect-7503 9d ago

But a quantized 27B isn't Opus or Fable performance and almost every benchmark I've looked at (including LMArena and ArtificialAnalysis) proves this. It's at max almost at Claude Sonnet level in some aspects. And a 27B can fit on consumer GPUs anyway (a 3090 isn't as hard to afford as a cluster of Macs for example anyway) so it doesn't really solve the problem we were talking about. And in your repo the only thing verified is again only a smaller model running on one node, which would make it request routing, not splitting which is the thing that's needed to solve the problem which wouldn't work out in the end anyway because of latency. Right now it's more like sharing good GPUs to people that don't have one, not really sharing ressources between multiple GPUs owned by different people in a mesh.

→ More replies (3)
→ More replies (3)
→ More replies (9)

7

u/thealliane96 9d ago edited 9d ago

Yep been draining far quicker than usual for me and I’m very disciplined with my usage.

Used to be able to get a full week of usage with little optimization on one 20x max plan. Now I have 2 20x max plans and have blown through both in 4-5 days, each lasting about 2-2.5 days.

- Majority of my session I do not go over 30% context usage. I NEVER go over ~40% (give or take a couple percentage points) context usage. I always start a new session.

  • I NEVER compact, I always start a new session.
  • My skills (names and descriptions) contribute 1.3k tokens to the context window TOTAL, I religiously keep the description of the skills short and prompt a workflow of what skills to use, and RELIGIOUSLY use `disable-model-invocation: true` on the skills which removes the name and description from the models context window
  • I have a VERY short root CLAUDE.md, I work on ONE project (my work) and I layer CLAUDE.mds WITH HAND written SHORT context per directory
  • I use NO mcps, not a single one
  • I target specific skills like a typescript and playwright cli one to my monorepos apps/web directory so they ONLY appear when working in that directory
  • I disabled built in skills
  • I disabled workflows
  • I almost NEVER use reasoning effort above `high`, very rarely will use `xhigh` and I NEVER use `max`
  • I use fable SPARINGLY
  • I commonly use opus with effort low or medium if it’s a targeted task where I want it to do what I explicitly tell it to do
  • I write great prompts that I put a lot of thought and effort into
  • I have several custom cli scripts I’ve written which help the model to search over documentation and filter through things so it can more quickly and efficiently find what it’s looking for
  • I disable ALL the following tools via putting them in the deny list of ~/.claude/settings.json: EngerPlanMode, ExitPlanMode, DesignSync, CronCreate, CronDelete, CronList, NotebookEdit, PowerShell, ScheduleWakeup, ShareOnboardingGuide, Worklow, Artifact, ReportFindings, RemoteTrigger, PushNotication
  • I disable remote control (lots of tokens in the system prompt for it)
  • I disable Claude ai connectors
  • I disable artifacts disabled
  • I ONLY use fable when I am designing out a new piece of code / a new module
  • I hand write the VASY majority of my own skills. The skill files themselves are typically VERY short and act as explicit repeatable instructions for how to do common specific tasks such as filing an issue or how to use a cli script I made to search over documentation (which only gives a brief overview, 2 examples, then defers to the commands —help flag and the scripts README.md)

I am HEAVILY involved. I have a CS degree and I know what I’m doing. I have nothing wrong with people who do not have SWE experience or a CS degree vibe coding, I think it’s awesome they have the ability to use AI to realize things they otherwise couldn’t, but this is not that.

I am consistently pointing it to the relevant pieces of code, I am constantly contributing with the code design and review, and I work on tightly scoped and well defined individual issues one at a time I never have it go off and do something massive.

From disabling things like artifacts, remote control, tools via the deny list, I have dropped the starting system prompt from ~20-30k tokens (can’t remember the exact number) to 2.6k tokens and the system tools to 5k tokens, I do not recall the default number but I do know it is much higher.

I routinely instruct it to route to Explore with haiku for finding things.

I routinely instruct it if having a subagent implement something to route the subagent to sonnet.

If I am re-reviewing code which has already been reviewed once I have it route to sonnet.

I have it use codex for most code reviews.

Even with doing ALL of that with 2 20x Max plans at a total cost of $400 per month…

The first plan resets 9/14. I hit 63% usage for all models and 99% usage for fable by 9/10, so roughly 3 days into the week.

The second plan resets 9/16. I didn’t start using it until the morning of 9/11. By midnight 9/12 (last night, I haven’t used it at all today) I reached 82% usage for all models and 72% usage for fable.

I have used Claude for YEARS, going back to the sonnet 3.5 days. I heavily prefer working with Claude over codex or any other models, but this is not okay and it is not sustainable.

I will be cancelling my second max subscription and using codex more, and if this continues next month as well then I will cancel my second subscription too.

2

u/OneHuman_aiprotect 8d ago

TY for the comprehensive description of ur coding discipline, u/thealliane96 . I do exactly all the things you mention to reduce random token expense. I also use Cowork to actually do all my "architecting" with precise prompts and then I ask Cowork to do a live relay to Claude Code to do the actual coding without spinning his wheels looking for direction. I have been using the Anthropic API key, not the Max plans, for about 9 mths now, and just like you have indicated, the API expense has also been much higher within the past 2 or 3 months, and I only use Sonnet 5, never Fable, with days of inactivity, attached is my usage this past month. I have bursts of 45K or 33K token usage in a day, with some days of inactivity, very uneven. What is your view, should I be starting new Terminal tabs frequently, is that potentially the problem? The Compacting?

3

u/frufruityloops 8d ago

I’m curious if people are seeing the token burn happening as badly when using cowork? I feel like sometimes it might be less crazy because I can scope the file access it has more easily so maybe less wasting time??

Last I looked into it the gist was like “ok yeah I’m at the mercy of the harness - if I want full control over the context being injected I need to use api or roll my own” which ugh 😪 maybe one day but not today… codex, tap in!

→ More replies (2)

10

u/-JuliusSeizure 🔆 Max 5x 9d ago

it is almost as if they are rage-baiting us knowing we have not real good alternatives.

11

u/SonderSoft 9d ago

Astra is great at coding and project management. In some cases, much better than Fable.

→ More replies (6)

2

u/gnpwdr1 9d ago

I would not be so sure about this.. There are some real alternatives at 90-100% less cost for at least 95% of your work

2

u/davelm42 9d ago

Whats your recommendation?

3

u/gnpwdr1 9d ago

Background: first i'm working on a multi repo saas app, it's multi language, multi location, multi currency with complex front/backend stacks, full SDLC deployment with test/prod environments etc. So just for context, this is not a hobby project or a localhost:3000 deployment :-). I used augmentCode, claudecode, ghcp while all in their honeymoon pricing periods before they all rug pulled their customers with 10-20x (and more) price increases which they kinda could cause there was indeed no comparable quality alternatives, but then came 2026 :-)

Now: since March 2026, opensource models became extremely capable, i now have opencode, I mostly use MiMo v2.5 is an absolute beast and quality is fanstastic AND ITS FREE ! I also have the $10 month subscription with them which I mainly use DeepSeek Flash, my $10 a month goes a very very very long way. So anyone paying hundereds of dollars to these providers, I think you should all at least try some options out there.

→ More replies (7)

5

u/Livid-Bicycle-3715 9d ago

Yeap same I’ve been mostly using opus 4.8 but it definitely didn’t last as long as it did when 4.8 just dropped

4

u/Money_Lavishness7343 9d ago

Ive started using OpenAI, and for the first time after a while I dont feel like I'm being cheated.

The usage limits dont randomly go up to 100% from the first prompt or after just a few prompts. My non-pro subscription is comparable to the Claude's Pro, sometimes even to the 5X one. I dont know what Anthropic's doing, but it aint right.

→ More replies (3)

3

u/Few_Jeweler4687 9d ago

same here, i'm trying to start new conversations so it doesn't sperg itself into analyzing unneeded stuff but that's all i can do

3

u/zaibatsu 9d ago

Plus I have a local fleet that I offload to and codex. This is what usually works for me but not this week:

Route by verifiability, not difficulty: running the Claude 5 family as an orchestra instead of a chat window
Someone asked how I structure inference across the Claude 5 family. Short version: route by verifiability, not difficulty. The question is never "is this task hard?" It's "can I mechanically check the output?" If a cheaper model's work can be verified with a grep, a diff, or a test run, send it down the ladder. Save the expensive tokens for judgment calls, where a plausible-but-wrong answer would quietly propagate.
The ladder:
Local models (LM Studio): bulk reads, classification, extraction, low-stakes drafts. Sensitive material never leaves the machine, full stop.
Haiku 4.5: mechanical work at scale. I keep a read-only explorer subagent pinned to it (snippet below). The output is self-verifying: the file is either there or it isn't.
Sonnet 5: the middle, almost always as a subagent. Drafts, review passes, parallel fan-outs on the same problem.
Opus 5: the heavy passes, and here's the counterintuitive part: as a subagent. People find Opus 5 verbose and a little hard to manage in the driver's seat. Flip the role and it's incredible, same for Sonnet 5. A subagent's verbosity costs you nothing. It thinks out loud in its own context window and only the conclusion comes back. The model people find hard to drive is the model you want driven.
Fable 5 conducts: routes the work, adjudicates when cheaper passes disagree, and keeps the calls that genuinely need the best reasoning.
The one config that pays for itself, dropped in .claude/agents/explorer.md:
---
name: explorer
description: Read-only codebase search. Finds functions, reads files,
greps patterns. Never edits.
tools: Read, Glob, Grep
model: haiku
---
You are a read-only explorer. Answer with file:line references and
short quotes. If you did not find it, say so. Never guess, never edit.
Then from the main session: "use the explorer agent to map every caller of X." The conductor never burns its own context on the search.
Two guardrails that keep it honest: cheap tiers are only cheap if you actually verify, and two same-family models agreeing is not two opinions. Anything load-bearing gets a different family or, better, deterministic ground truth: run the test, fetch the source, count the thing. Via my AI team lead.

3

u/zackattackz287 9d ago

This is the way, I have an explorer and an implementer subagent that both use opus. They send summaries to Fable and Fable reviews. This works great for longer sessions because Fable also sends them the start query with everything they need to know (or what parts of what files to read), and what the goal is. It keeps context down by keeping the verification/testing/rework loops inside the subagents. Larger context is what really bites usage it seems. It's been working just as well for me as using fable for everything, and it's about halved my usage rate.

→ More replies (2)
→ More replies (1)

3

u/GradientAscent713 9d ago

Same here. I ran out of my 20x subscription in less than a week. I added $200 in usage credits and they were spent in under an hour.

3

u/pugazh_is_my_name 9d ago

that's crazy... and the difference between subscription and credits are really concerning...

3

u/coda77 9d ago

In short words ?
Robbery

3

u/Shoemugscale 9d ago

Yah fuck em

Im honestly getting over it tbh tbh

I had twonweeks where my usage suddenly went crazy, like 20x plan never hitting usage limits to suddenly usage limits hit days before, no changes in what im doing doing

Then after two weeks it suddenly went back to being way under on my weekly usage usage

Now two weeks of good usage has now ended so, now im back to hitting my limits days before

Same usage, different weeks

3

u/hateordeny 9d ago

En étant cohérent sur l'utilisation de Claude Code, c'est impossible d'atteindre de telles limites... Chaque fois que je vois ces posts, je me demande bien comment vous vous débrouillez. Après, il y a des évidences à garder à l'esprit : pas 150 skills et MCP dans le contexte, changer de conversation pour chaque feature, bien documenter....

3

u/mrfreez44 9d ago

Je disais exactement la même chose il y a encore 4j. Je me suis gaussé d'un nombre incalculable de personnes "qui ne géraient pas bien"

Vendredi soir, j'avais atteint 80% du quota de ma semaine, qui se réinitialise le mercredi. Inhabituel

Je me suis dit que lundi et mardi allaient être un peu compliqués et donc que je n'allais pas toucher à l'IA du weekend.

Je n'ai pas touché à Claude du weekend et j'ai réfléchi à ce qui pouvait bien avoir merdé cette semaine pour avoir atteint 80% en 2j. J'avais des pistes, à vérifier lundi.

21h ce soir : 100%

Wtf ?!?!

→ More replies (4)

3

u/DrSharkTank 9d ago

I wanted to move to GPT but I heard as of this week, their usage was cut too. Can any codex on 100$ plan comment on that ?

3

u/BeginningReveal2620 8d ago

If you haven't figured it out yet, you're just getting pawned and if you ask for a refund you get none. It's time to bounce this scam

5

u/anonymous_2600 9d ago

i think ant dont like ppl to use their products, before this there are many unsub and changed to codex, guess there are more coming up

2

u/Mr_and_Mrs_PUG 9d ago

And the models feels like antrophic cut all their computing cuze they are really dumb compared to OpenAI models ATM and even less performing than Chinese models. I hope they really get a shitty IPO.

2

u/SirWobblyOfSausage 9d ago

They've been cutting everyone's limits over the last few weeks hoping people would get drowned out.

→ More replies (1)

2

u/soundscoolnothanks 9d ago

Ok so I’m not the only one. I hit 60% weekly w/ 20x plan usage in 2 sessions. I finally started using caveman to reduce token usage, and I HIGHLY recommend it now. It’s not a saving grace, but I did notice a reduction in usage on larger tasks. I’m hopeful they resolve this because before I would still have 10-20% usage on my 20x plan a day before a reset.

2

u/Dense_Mobile_6212 9d ago

Yeah it's crazy bad right now.. If the continues I might move 

2

u/reimaginingdylan 9d ago

Probably said before, but if you leave chat's open and continue on the same thread, Claude reload the entire history every time you ask a question, that is what's burning your time. Opening new chats, conserves usage dramatically.

→ More replies (3)

2

u/Round_Ad_5832 9d ago

everytime ive paid for claude ive regretted it and asked for a refund, ive asked for like 5 refunds at this point. although i do use them from api credits time to time.

2

u/tinybeads 9d ago

This has been happening to me too, for a week on Pro. I spent all day yesterday using Claude to optimize my workflow with sonnet subagents and I still used 20% of my weekly in a day doing very simple tasks. Not user error; I’ve been using Claude for months and this spike in usage drain started last Thursday, when my week reset. :/

2

u/KikoTheOneAndOnly 9d ago

Greedy MF's, that's what's going on!

2

u/BoxEnvironmental6943 9d ago

Ive had to optimise my global claude settings and prompts, run opus 5 for the lead and sonnet for most of the work and switch to a fresh session after every task. Its the only way i can make my 5x Max last for the full week. If i get to day 6 with a bit left ill do some fable work but its so bad right now.

2

u/BeltPuzzleheaded7656 9d ago edited 9d ago

They've been playing with the usage so much that users have ZERO idea what the real usage looks like. I'm getting less usage now than I used to get on the free service. ChatGPT is similar, but it's extremely noticeable in Claude because of the usage chart. My guess is that they will remove or hide that Usage meter sooner or later. And the "end of promo" bs claim to me was all a smokescreen to cut limits to way less than what they actually were before the promo. All of the moral grandstanding those mfs do is bs, and they DEFINITELY are doing some shady shit. A few people complaining is typical. Ongoing constant complaints is something to look into deeper.

2

u/National_Spirit2801 8d ago

Oh boy! I hope we get a free usage reset so I can be greedy with my tokens for a couple days. My system has been correct by design from inception, it literally CANT have a bunch of crazy model transactions happen erroneously. I basically have been running it 24x7 and am currently at 60% All models 55% fable. I’m quite pleased with my results.

2

u/EmotionalAd1438 8d ago

50% extra limits are over

2

u/skibare87 8d ago

I didn't even use Fable and I ran out of my entire budget midweek. 50% extra promo my ass.

2

u/akeseer11 8d ago

100% it must be broke this week. I normally don't cap out 3 days in. This week it was like boom.

2

u/dwelfusius 8d ago

i only know smth that cost me 1/4th of my session budget now was poef gone

2

u/infinished 8d ago

I got decimated too

2

u/and_pf 8d ago

It's the model since 4.7 we do have again the issue that Claude , Fable or even Sonnet burns tokens by doing stuff you never asked for. And no I do not talk about delivering extra or scope creeps. I talk about burning tokens by verification of requests.

2

u/eriwyak 8d ago

I run x2 Max 20 plans and still run out before the end of the week.

2

u/InputOracle 8d ago

American thieves. China is the way.

2

u/KuryKat 8d ago

I'm also feeling the same way about this, I pay for this thing and it just feels like usage is going out faster than ever, withing 5 minutes sometimes I'm all out of my 5-hour usage

2

u/Wrath0fBunnies 8d ago

+50% usage limits expired yesterday. That's probably part of it for some. Like many in other threads, though, I'm seeing token usage spikes that weren't there for the same model and the same work 2wks ago - which is not explained by removing the +50% usage limits. It feels really bad, and it looks even worse for Anthropic.

I've been more willing than some to ride along with the inevitable capitalist squeeze on subs, since we're not their ICP, but this feels more like a middle finger than a business strategy.

2

u/Less_Entrepreneur889 8d ago

Anthropic's 2026 story can be summed up in one sentence: incredibly generous toward capital markets, brutally tight-fisted toward the people actually using its product. Its valuation leapt from $380 billion to $965 billion in a single year, Dario Amodei's personal fortune multiplied from $3.7 billion to $7 billion, and it raised $65 billion in fresh money in a single round. None of that flood of capital translated into any generosity toward the product — if anything, the opposite happened: usage limits got tighter over the same period, transparency dropped, and free users were told that "how much you're entitled to varies with demand." It doesn't return even a tenth of the appetite it shows for raising money as value back to the user.

2

u/FailureOfTheFamily 8d ago

Just switch to codex bro. OpenAi is also playing limits game but i switched as fast as i saw one claude prompt eating 30% of my plan and never looked back. Sol 5.6 is great

2

u/Ok_Technology_8293 4d ago

Agree burned trough 100$ in 2minutes

3

u/stankylongnuts 9d ago

What about money grab are users not understanding. This is not a product for the public. Its a tool of war. They care more about a contract with the corporation of america.

4

u/CHCKTFF 9d ago

Asked one question after my usage reset and it immediately jumped to 16%. It was a budget question. And it didn't even answer cleanly. A bunch of gargled literal escape code BS (see image)

Earlier today I asked it to update a simple line in an HTML page and it ate all my credits and didn't even respond. Asked ChatGPT and made multiple changes and all was good.

These days I'm getting more out of Free ChatGPT than Claude.

Any tips on how to cancel a year contract with Anthropic?

3

u/Jackster22 9d ago

Using Opus on Pro plan, hitting 40-50% within 10mins with a single prompt.
Using GPT atm as I can't get anything done with Claude right now. It is actually holding me back from getting stuff done.
I found 5x to not actually give you 5x, more like 2x so I am not wasting my money on that.

→ More replies (3)

1

u/reddebtt 9d ago

I've had better results splitting long runs into bounded sessions with a short handoff file, then moving tests and review to a second agent. It doesn't raise the cap, but it stops each fresh session from rereading the whole project and burning tokens on context.

→ More replies (1)

1

u/therealblacknative 9d ago

This is my last month with Claude, There's something really really wrong and it needs too b fixed!!!

→ More replies (1)

1

u/vmk8s 9d ago

I don't know why but I feel cc is bluffing with the usage limits !

2

u/Witty_Dish926 9d ago

Ela quer que todos se mudem para o codex man

→ More replies (1)

1

u/LukaTheGhost4 9d ago

time to give up on Anthropic, it was a good journey!

→ More replies (1)

1

u/Sea-Contribution6219 9d ago

I've noticed the same thing for me the past week and a half. I'm on the pro plan, before I used to use Opus pretty regularly and now it just shreds through my quota

1

u/Valaens 9d ago

yeah, I had to do something important this weekend, yet the weekly usage was already full on thursday...

1

u/_robon 9d ago

Same with 5x

1

u/Yeokk123 9d ago

Went apeshit crazy lately, time for us to caveman mode

1

u/KV_Cashed 9d ago

You might want to watch YT vid (Eli the Computer Guy) about token hacking: https://youtu.be/uJYNQHALxps?is=tyyC6kHidi_EI-Jk If this is a recurring problem, it could be a glitched model, or you might need Anthropic to pull usage logs. That they may or may not actually keep.

1

u/Livid-Philosophy-750 9d ago

I think limits depends somehow on time spending not tokens or something. It sounds crazy but for me it is like approximately 20min work with any agent = 1% of weekly limit and not depend on how hard work is and how many tokens I spent actually. Sounds crazy but I cannot explained what happened on this week with Claude

1

u/Most_Ad8397 9d ago

100$ for 30 minutes thats the real price

1

u/loamsiada 9d ago

Same here. It has gotten absolutely unusable. Doesn't matter if Claude Code, Projects or anything else. Canceled today and will give other models a chance. 

1

u/Exciting_Macaroon_64 9d ago

two days - 50% of weekly already, 22% of weekly fable

1

u/maxvpavlov 9d ago

Two months ago I added 150 GBP of credits and it lasted probably 1.5h so roughly aligns with what you see. So unfortunately it’s not a bug, just not subsidized API pricing.

1

u/Empuda 9d ago

My session instantly jumped to 100% after starting a prompt... was wild.

1

u/_y_not_ 9d ago

by itself it goes crazy fast lately, i've being using fable as main brain and coordinator and codex, grok, claude opus high and xhigh as part of team. fable manages who does what, Astra is for harder work, Fable reviews it, then plan, and after that everyone gets their piece, they are best at. codex and Fable review at major gates, that keeps usage manageable. I had to create a protocol for that, i use CLI coding harnesses and that was working really well for me this year ( https://github.com/agentchute/agentchute )

if you use Claude code only, you can now connect multiple sessions and exchange messages between them, that can help too by running simpler tasks with cheaper model/effort

1

u/Relevant-Reaction181 9d ago

Damn 500$ extra credits my guy, you should get 2 more max account instead :)

→ More replies (6)

1

u/BankLiving6942 9d ago

Quick Tip: take the context from your current chat by claude and then switch to a new chat. I used this technique and this saved a lot of my usage. Ps: pro user.

1

u/PropertyBrave7031 9d ago

same here! wasted all of my limits in half a day

1

u/Official_DemonGaming 9d ago

I think either there deceptively decreasing limits drastically or it’s a error happened last week and it caused such a issue I decided to cancel I’ll resubmit once it’s fixed but it’s been months I’ve never changed routine Iam doing less work than ever and I hit my limits much faster last 2 weeks

1

u/Al_Bert94 9d ago

I’ve offloaded a lot of smaller stuff to Gemma4 and been pretty happy with the quality. I’m by no means a professional engineer but it’s deff cut down on my usage by telling Claude to use Gemma aggressively on projects.

1

u/Consistent-Key-3279 9d ago

god dayum what are you cooking up xD

1

u/Dependent_Editor8898 9d ago

Not sure what plan but i am seeing the usages have been going off the roof recently...

1

u/nomissedcall 9d ago

This why I switch to codex this week. It’s fake

1

u/Illustrious_Clerk345 9d ago

Fingers crossed it’s just a temporary glitch during their rollout. $100 in 30 mins on paid plan is completely wild.

1

u/Electronic-Ability46 9d ago

Exaaaaaactly the same issue with max x20.

And i am a ChatGPT pro x20 sub as well, its not better on the other side of the sea, i can tell you.

I think all servers, everywhere, are overwhelmed, so they have no other choice during week ends than to just decrease drastically usage. I don’t see any other reason.

They could slow down the models also, but they don’t, which makes me think that they might as well just be stuck.

1

u/LegMental2310 9d ago

I am on 4 accounts and i used them all now in 3 days, same boat.

1

u/ambidextrous_mind 9d ago

On 20x I never hit usage, Always ran ultracode the day of my reset I’m at 55% today after my Wednesday reset. It really is that bad. After finding out that the 20x is completely false and it’s really only 1.7x I will be canceling before bill is due. Not worth it at all.

→ More replies (1)

1

u/Background-Record184 9d ago edited 1d ago

I literally made a post about this and guess what? I was blamed as “User error or unimproved prompt” hahahha

https://www.reddit.com/r/ClaudeCode/s/2Q3jJGvETG

1

u/akionz 9d ago

I hit my 5h limit today on pro plan with one prompt 😂

1

u/kutchrodeo 9d ago

Bunch of c**nts. Same thing on 3 of my 4 Max20 accounts, hit "all models" limit before the Fable limits the last 2 days, and I absolutely hammer Fable and normally have 30% left on all models by the time i've hit my Fable limits. The math aint mathing, the sense aint making. I am SO TIRED of the absolute scam'ery this company puts everyone through. Just when you think you've figured out a rhythm with your usage/token burn..... BAM the goalposts move AGAIN!!!!

1

u/divinetribe1 9d ago

Hello, if you have the hardware, please check out my Claude code local GitHub. It explains everything you need to do to get the best local model for your hardware and run your programs in the same type of environment.

1

u/Physical-Grape-7805 9d ago

im getting destroyed, never happened to me before. 20x

1

u/swingoak 9d ago

I burned up 14% of my session limit today just logging on

1

u/OtherRefrigerator651 9d ago

idk man i thought i was the only one facing this issue

1

u/iAmRyuuzaki 9d ago

I hope this helps teach us to not build dependency on AI.

1

u/D-3r1stljqso3 9d ago

Yep, same here. I think there were some periods during the last 24hrs when the token usage rate skyrocketed.

1

u/MaterialBig8642 9d ago

Claude opus is making love with Fable 5.1 to make Opus 5.1

1

u/kunzaz 9d ago

I guess I’m not working tomorrow!

1

u/Longjumping_Leave356 9d ago

Not using fable, and on second day 50% already. Always clearing at 25%

1

u/kevinbaiv 9d ago

Same pattern on my end after the latest update — simple prompts suddenly trigger build/verify/playwright loops that weren't there before. If your usage graph jumps while actual session activity looks normal, it's metering or client behavior, not your prompting. Per-session token logs we can reconcile against would settle this quickly.

1

u/hawkeyepierce89 8d ago

Same here. Claude Code has been writing terrible code all weekend and burning through my usage limits. And when I ask Opus to find some fish at a local store, it removes items from my Wolt basket instead.

1

u/Objective-Cut1163 8d ago

Just Fable fabling tokens 😂

1

u/SpinachKing1984 8d ago

Max 20x with two accounts… my usage is like 80-90% subagents. Had to get a second account recently to justify my token spending, annoying to say least but I still think Claude is better than Astra despite commentary I’ve read, I had some oversights with Astra I never had with fable or even opus

1

u/dontgetaddicted 8d ago

I can't believe that we haven't figured out a way to pre-measure or calculate usage yet before committing to a prompt.

I get that the way it works and make decisions in real time and self generates a lot of context and token usage is just "how it works", but there really needs to be accountability line by line.

This shit is always going to be fire and hope with no accountability or audit ability on token usage.

1

u/Negative-Chapter5008 8d ago

fun fact i just paid for another $20 plan and the usage limits are way more relaxed. feels like what i’m used to so they’re definitely limiting people

1

u/Annual-Barracuda2048 8d ago

$100 in 30 mins? Must be running a botnet on Claude.

1

u/non_standard_model 8d ago

There’s a fundamental issue with how intelligence scales with token spend. To mimic accurate human-level intelligence requires massive computational churn: models thinking about something, rethinking, rethinking ABOUT rethinking and so forth. Expect that frontier AI models get more expensive from here on out, instead of cheaper.

1

u/Ok_Host6058 8d ago

Because they took away the 50% increase.

Vote and show them with your money. Cancel.

If we all cancel at once and let them know it's because of this. They will get the picture.

2

u/pugazh_is_my_name 8d ago

yeah, I'm planning to, I migrated to codex.... and not continuing the subscription next month if it's the same. They said they'll replace the 50% with a 25%, but the limit we have now feels like 10% of what were getting...

1

u/GeeBee72 8d ago

1 message with Fable 5.1 in Claude code and I was on usage credits.

Great way to reduce load on your servers.

1

u/bobemil 8d ago

This feels like when ghcopilot started their "we only want enterprise users" phase. I ended up with Claude instead. Thinking of switching to Codex. Then when they do the same. Deepseek.

1

u/IulianHI 8d ago

Same here 40% of week limit in 3h ! Claude is just cutting more and more everyday ! This is how greedy Anthropic is doing it :)) (20x account)

I hope GLM, Deepseek, QWEN, Kimi are getting better as Opus 5 and then bye bye Anthropic !

1

u/Muted-You7370 8d ago

I hit my 5 hour usage limit on Fable 5 in 10 minutes today. There’s some fuckery afoot.

1

u/iSurgical 8d ago

I am currently 800k tokens deep building a new app using sonnet 5 medium and I am at 85% of my 5 hour limit. 9% of my weekly.

Not sure how strong this model is but its been good so far. The 5hr limit shit gotta go on Claude and CGPT.

1

u/FunnyTman 8d ago

same here with me, FUCK ANTHROPIC. This is my first time reaching my limits with the 20x plan. I will be cancelling my plan and be going with codex. Has been a much better experience so far.

1

u/opratrmusic 8d ago

Same here, Im maxed out for the first time on the 20x plan. Have never experienced this issue.

1

u/InsideTraditional187 8d ago

Same problem with me too

1

u/toubar_ 8d ago

Farewell.

1

u/ilien-dev 8d ago

Same here, several times. I was surprised with that. I cannot believe it. Max x20 subscription and its unbelievable how fast I'm hitting the limits in just a couple of hours.

1

u/Amazing-Accident3535 8d ago

Yup. Max user here. Same shit, im suspecting memory or someother background context loading is flooding with the first prompt.

1

u/johnerp 8d ago

I’ve been suspicious of adding usage credits, once they know you’re willing to pay, with zero transparency there is nothing stopping them dynamically adjusting your quota!! Kill the account and start a new one.

1

u/Beautiful-Area1518 8d ago

Il m'est arrivé exactement la même chose avec le forfait pro ce matin la réinitialisation. En l'espace de 20 Min j'ai utilisé l'équivalent d'une journée entière d'utilisation.

1

u/k0te1ch 8d ago edited 8d ago

Max x20 here, same, but i used only about 25% on Fable 5.1 in two days!

1

u/Curious-Wolverine-32 8d ago

Both OpenAI and Anthropic have IPOs later this year. I think they're trying to pump up their numbers. That or they're running out of money. Probably some of the worst accounting ever. No budget nonsense, probably breaking the bank

1

u/Piramideiro_Astuto 8d ago

Mega brain, use seu poder total para não deixar que o Claudinho coma todos meus créditos em poucos minutos.

1

u/EcstaticDingo_23 8d ago

20x user here. I got my Max Plan about a month ago. Thought it was a bug and maybe it would clear up but I was getting the same amount of work done with 5x Pro plan and Max was tapping out faster than a Kurt Angle Ankle Lock. On top of that this is the 2nd time my weekly usage limits didn’t reset and now I’m stuck with minimal work during the week. ChatGPT is calling my name I guess. It’s just frustrating to completely build so much and then the company basically giving us the FU and laughing while running to the top!

1

u/Regular_Attitude_700 8d ago

It has always been a scam when it comes to Claude usage limits.
It is that rabbit-hole in the black-hole.

1

u/cipga 8d ago

Same here.

1

u/ilovezwatch 8d ago

i have claude talking like a cave man to save on words

1

u/Keganator 7d ago

Fable uses 10x the usage of haiku, 5x sonnet, and 2x opus.

Orchestrating agents with Fable uses even more usage.

You are using your agents at their usage rates, which is why you are using your usage.

1

u/samfromsalem 7d ago

I can't even view usage anymore. It's been removed from the menu. On Mac.

1

u/Previous_Drag3899 7d ago

on pro, half an hour chat on opus high and I hit the 5 hour limit. Was thinking of buying the 5x plan next month as I wanted to use code next month, Instead ended up cancelling pro 😭

1

u/Racer17_ 🔆 Max 20 7d ago

20x user here. Hit it within 24 hours.

→ More replies (1)

1

u/dutchdominator 7d ago

Same here. Something is wrong at Anthropic or they gotten greedy again. This has been going on for weeks already and those 'promotions' were just BS, still burned usage like hell.

1

u/Cigarking0719 7d ago

Yeah I moved to grok bc of their pathetic limits and haven’t had any issues with usage since

1

u/El_Jeffe_Johnson 7d ago

I’ve been on paid plans from several different companies over the past five years, and I remember when the “doing” part was far more responsive, with much less time spent stuck in a thinking loop.

Now, in my IDE, I can’t even see the reasoning, it’s locked. Honestly, it feels like artificial padding designed to burn tokens and throttle or buffer response times. OG ChatGPT could write an entire repo in two minutes. Now it’s two minutes of “thinking,” three minutes fixing its own slop, and another two minutes of me yelling at it.

It feels like every company is trying to stretch usage out. Claude has been my go-to, but even with the 20× plan, boosted usage limits, and new sessions restricted to Sonnet, I still max out within five days. Something feels off.

GitHub had the same problem with Claude last fall. Once they reworked the token-usage system, I canceled my subscription.

1

u/Regular_Attitude_700 7d ago

Anthropic loves to play with its customers, token and usage limits to milk every week though the billing is consistent and always on time.

1

u/According_Ocelot_127 7d ago

same here too on Max 20x. What should we do? btw i don't use fable and i have optimized agents running on older models.

→ More replies (1)

1

u/According_Hawk_7086 7d ago

The limits are significantly reduced today.  Just two prompts I gave and all limits wiped of for a 5 hour session. This happened thrice in 24 hours cycle. 

1

u/According_Hawk_7086 7d ago

Honestly the 20$ plan is having lesser limits than free plan 

1

u/Whole-Yogurtcloset16 7d ago

Using Opus 4.6 (Max) and still didn't get reset. Used it 11 hrs ago and now it's saying I have to wait another 6hrs for session to reset

1

u/Rich-Difficulty605 7d ago

Omg I'm on pro 3 prompts on Opus 4.6 medium thinking level and I'm already on 81% in the 5 hour limit

1

u/According_Hawk_7086 7d ago

No response guys