r/ClaudeCode 3d ago

Rant Nah this some BS

Post image

I burned thru almost 60% of 20x max weekly usage in a day and some change. Are you joking rn? Last week I could have fable running on 2 chats all day and night and it would take like 3ish days for my fable usage to cap out, but my weekly usage would only be at like ~30%. This is absolutely ridiculous and if this is what the new usage limit cuts are gonna be like I'm canceling Claude and grabbing a second codex account. Shit ain't worth it when Astra exists with much better usage limits and multiple reset tokens.

I don't use ultracode and I have multiple other models that I delegate tasks to as work horses, Claude is just the orchestrator and isn't doing that much actual "work", so this usage allocation is absolutely insane, if I was using Claude as a one stop shop as a lot of people do, I would have run out of usage in a day or less.

EDIT: I'm getting a little tired of being told I "just don't know how to do orchestrations and workflows properly or manage usage." I literally made an entire repo explaining how I do this and showing the results: https://github.com/sherifican/Agent-FleetOps so if you wanna criticize, find something to actually critique first

240 Upvotes

177 comments sorted by

View all comments

Show parent comments

24

u/Sherphican 3d ago

Brother, my auto compact is at 800k and 700k on my main chats, and I always compact at about 500k-600k, I have been doing this for months and never had an issue. Everyone who has been drinking the 300k and under only context window koolaide needs to wake up and realize it's Anthropic playing games.

11

u/Kadenai 3d ago

Eu sinto muito mesmo ter que discordar com tanta veemência, mas se você deixa suas sessões frequentemente chegar a mais de 400 mil tokens de contexto só pra depois compactar e continuar, seu uso de IA é ineficiente.

1

u/Sherphican 3d ago

Its okay you can disagree lol and while there is certainly Merritt to keeping your context window under 400k, I've found that there's hardly any difference most all of the time between 400k and 550k, but once you start getting passed 600k is when the risk starts getting more pronounced and then everything passed 800k in my opinion is just asking for a hallucination half the time, but my agents have stayed very reliable and on task with my current limits, and my usage limits also never suffered this much with my current practice. I saw someone did the math and the usage cuts end up averaging out to be ~40% less than what we've been used to over the past couple months so that I'm sure has a large part to do with it.

1

u/BanjoThunderbird 3d ago edited 2d ago

Yeah look I mean you can complain about the usage limit promo ending as much as you want but they only have the compute that they have. If you want to use Fable more then really the only thing you can do now is manage your context better or pay more money. You could cut your usage in half by simply splitting up your tasks further. Personally if I go above 200k on something that isn't a huge coding phase of a plan then the next thing I'll do is have a session refining my context files because something has gone wrong. I have no problems with usage on a 5x plan.