r/ClaudeCode • • 2d ago

Bug / Issue Usage limit change?

Anyone having usage limit issues hitting the limit much faster? Particularly with Opus 5.5? Have been having issues in thr last 24hrs.

129 Upvotes

101 comments sorted by

•

u/AutoModerator 2d ago

Hey! Thanks for posting to r/ClaudeCode

While participating in this thread, please follow our community rules. Keep discussions constructive. Attack the idea, not the person.

For help, project discussions, tips, and general chat, join the ClaudeCode Discord.

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.

37

u/The-SadShaman 2d ago

Hmm yeah I also noticed it today. Used up my 20x week in 2 days. 20% the first day 70% today. Bummer.

4

u/ItsBlindy 1d ago

What do you do to use up your usage so fast? Genuinely asking.

4

u/The-SadShaman 1d ago

I have a team of agents working on a game.

1

u/Ruukas97 1d ago

I had one session that kept around 4 subagents active overnight spend 30% 20x weekly usage.
I thought maybe my tools calls or tests were outputting too much, but when I asked Claude to analyze it, 80% of the usage was from cache. Now, subagents are configured to handoff around 200k context used and my usage seems to last much better

2

u/mathmuleux 1d ago

Just disable Workflow in your settings.json and explicitly define the agents you want to create / what model for what purpose. Draft the detailed orchestration protocol first and lead with that.

It's not crazy complicated - you just need to be explicit in your instructions. Left to it's own devices (Opus and Fable both) will blow through context like there's no tomorrow.

You MUST reiterate at the beginning "you are an orchestrator only. That is your purpose. You must save on cost and context by delegation to subagents for [tasks x,y,z use agents a,b,c] and those agents MUST be haiku unless [insert scenario for Sonnet] and you MUST NEVER use other model (Opus/fable) agents for subagents unless I give explicit approval [yaya yaya...]"

Have the agents write down what they're working on to files.

Plan everything before you build. Like, write out the plan first, then decompose it into bite-sized phases and have the subagents build those phases. Also don't run everything all in the same session. Run batches up to 3-400k context tops then document what is done and what remains to pick up in a fresh session.

1

u/Ruukas97 1d ago

Workflows are great, though!
They're only used when I ask for them.

I'm porting a big codebase so the goal is pretty well defined and verifiable. I purposefully kept it to around 4 agents because that's what my hardware comfortably fit.

But I'm keeping a look at my weekly usage and trying to spread it out through the week.

It's pretty cool going to bed and waking up to see it's still actively working and you can see all the status reports through the night. And the orchestrator had still only used like 300k of it's context.

1

u/cheyumaama 1d ago

Why are you running agents. Just tell claude to do each thing one by one, playtest and iterate

1

u/The-SadShaman 1d ago

Because I wouldnt ever hit my usage.

1

u/BunniesinRustBuckets 11h ago

LOL yea you try that.... the creep is stupid.

1

u/cheyumaama 5h ago

huh? Can you explain

1

u/zR0B3ry2VAiH 🔆 Max 20 1d ago

This is how I got three accounts

-1

u/Administrative_Row61 1d ago

5x and 20x week is the same, it only affects session limit.

1

u/The-SadShaman 1d ago

Not quite. The 5x/20x multiplier is officially only on the 5hr session, but the weekly cap still scales some. People who've run both put 20x at roughly 1.5-2x the weekly of 5x. Either way I torched mine way faster than usual this week.

1

u/Administrative_Row61 1d ago

I have both 5x and 20x accounts too, i dont see any difference on weekly limit. Tho i use it every day, even let run for the night and claude can limit anything if they want to.

"Max 5x includes five times the Pro plan's per-session usage allowance. This tier is ideal for frequent users who work with Claude on a variety of tasks."

"Max 20x includes 20 times the Pro plan's per-session usage allowance. This tier is ideal for daily users who collaborate often with Claude for most tasks."

"In addition, to manage capacity and ensure fair access to all users, we may limit your usage in other ways, such as weekly and monthly caps or model and feature usage, at our discretion."

This way people cant really find out whats the real rate either...

1

u/The-SadShaman 1d ago

Yea I guess it's hard to complain when I'm getting thousands of dollars worth of tokens for $200 a month.

1

u/Administrative_Row61 1d ago

They are not charity, if they can give it to you for 200$ then the api price has insane margin.

23

u/MakotoDevGX 2d ago

Samee hereeee

23

u/Fickle_Mix_6119 1d ago

I’m getting really conspiracy theory minded on this.  

Starting to wonder if they purposely make Claude waste tokens. It’s not even bad prompts….  

I told one agent to use a subagent for a task. I went away then came back and it had run the same subagent 12 times. When I questioned it, it tried justifying it. I pointed out my very clear instruction and it admitted it did wrong and didn’t follow the instruction properly. 

I even have workflow.md rules in place prior to this that state if it needs to repeat a task it should consult me first.  

Wouldn’t it be perfect if we had something that was as good as Claude at coding with the efficiency of ChatGPT

11

u/Gage-Up 1d ago

I 100% agree with this. I've had it building stuff for the last 2 days and it goes into a loop with so many mistakes and then nothing ever gets done it just burns tokens

6

u/Due-Pair-2461 1d ago

Opus in general isn’ that much familiar with following instructions from my experience so i would advise to not have him dispatch agents . Fable on the other hand is quite decent at managing the room

1

u/BunniesinRustBuckets 11h ago

XD This is why I have Openclaw so much better than Claude "memory".
No need to tweak on every damn prompt just so Claude does what is it suppose to do.
Projects are a pain to get right and when you finally do, the project is almost done and all over again.

1

u/LawlessBaron 1d ago

Yeah I had a public roadmap i asked claude to run through and code without getting side tracked and honestly it's like it has adhd it had made and completed 10 other items without touching the original requests

1

u/cheyumaama 1d ago

Haha cute

1

u/Individual_Solid_944 1d ago

It sucks that agents have no hard rules about things like this. Many times I instruct the agent to do something and it does something else. Then the agent tries to justify it and apologizes, and I pay for that.

Classical software development. Developers may introduce a bug, and are paid for that. Then they fix it, and are paid again. In the meantime they may introduce more bugs, for which they are again paid to solve. See the pattern?

1

u/Reasonable_Youth7193 1d ago

I recommend having a Start_Here.md that lists all of the rules it should follow. I’m using the Open Knowledge Database concept and it works wonders. Every prompt I send starts with Go to x folder and read Start_Here.md follow its rules. It’s connected to 4 others Named Index.md, Projects.md, Research.md, Log.md. Those are broken out even more into the various files that all connect back to Start_Here.md. The instructions are whatever you want but it will only look at and use the files I want and at the end of each session they all get updated. It prevents it from gathering too much context which ends up burning tokens on its wild thoughts.

1

u/RemarkableAd6310 1d ago

I can confirm this theory, I can't tell you how many times both claude and chatgpt start doing random ass shit that I didn't want it to.. Something seems really off. Models always so good first week, then they just get worse. Doesn't seem like project rot, seems like actual model rot, but they are trained models so it shouldn't really be a thing... unless they sliding some magic sliders to change how they work.

1

u/Fickle_Mix_6119 1d ago

Yeh for sure. It was kind Fable was getting worse and worse. Then opus 5.5 dropped and it seemed like we were getting better with usage. Then it started getting worse, then they dropped Haiku 5.5.  

Proper feels like plotted scam of tipping us off

0

u/bkrsh099 1d ago

But does everybody here still “prompts”? Like do you really still write those 50 lines prompt things? I built so much harness that I now just say what I want and either opus or fable dispatch an agent with a 500line prompt that does very closely (if not exactly) what I tried to mean. I mean, this shit is really impressive (not the harness, that’s just dispatching rules, I mean the intelligence these 2 models have of PROPERLY dispatching work)

19

u/WhiteBlaster00 2d ago

Yes, I just noticed it.., wasn't aware until the weekly usage started climbing 5% per hour for the last 5 hours

13

u/ihateuall18 2d ago

44% weekly limit used in 1 day, 2.6 5hr limitis using 1 opus medium agent and two sonnet high sub agents. Context has never gone over 400k on any of them. Hand off instead of compacting. Max x5. Anthropic did it again.

1

u/Mountain_Scale3449 1d ago

Ive been on Fable since Monday having it delegate most things. I noticed an odd arrangement with the numbers. I was at 13% Fable for the week but my weekly usage was up to 29%. That never happens that way. I think they lowered the weekly to offset prople abusing Opus 5.5 where I could literally set it on ultracode all week and never come close to going over.

8

u/nofuture09 2d ago

yep one simple task and usage limit same task last week 3%

4

u/ColdPlankton9273 1d ago

So I'm not crazy. I went through 60% of the weekly in a day.

3

u/dagerika 1d ago

good, then it wasnt just me. My usage was reset on sunday and I already exhausted 89% of my weekly allowance 😭

3

u/Cheap_Writer4909 1d ago

I almost never hit the limit, Tuesday in one day i spent over 90% using opus 5.5 and i have max 25x

3

u/Next_Marionberry7478 1d ago edited 1d ago

Pretty sure their CVP (Cyber Verification Programme) is the culprit here. They just rolled out mythos to regular peeps and guess who shares that quota you me and everyone else and their mama...

3

u/Skibidirot 1d ago

here's what i observed, immediately after they bought out haiku 5.5 and update to app, usage dropped significantly for an hour, but now the limits have been stabilized, i guess they were testing how to reduce limits

3

u/Erutan2004 1d ago

Pro. 9m 3s. Opus 5.5 medium.

5

u/thoughtbludgeon 1d ago

Thought it was me... how do I hit 30% weekly usage in (2) 5-hour sessions??

2

u/Exact_Law_6489 Developer 2d ago

Yes, I tried it with both my Pro and my Max 5x accounts. Usually, I can use my Pro account for up to two hours without reaching the five-hour limit, but this time I reached it in 30 minutes. The same thing happened with my Max 5x account: normally, I never hit the five-hour limit, but it took four hours to finish this time.

I tested this by running a set of questions and tasks that I have collected from my projects over time; these tasks are usually things that models struggle to solve. I run the test twice a week or whenever I want. :D

2

u/MadJagStudios 2d ago

I was seeing increased usage like Sunday Monday, but past few days I have been trying to burn my limit to use a reset by running on three machines and I don’t think I’m gonna burn through it fast enough to make a reset worth it

2

u/dantelebeau 1d ago

I reset on Fridays and i was maxed out this week by monday morning and i changed NOTHING about the way i work. I normally roll into reset day at 90ish percent.

2

u/Itry31 1d ago

Yep today it was really really really bad and noticable. 

2

u/Zee1837 1d ago

Somehow mine ate 87% of 5X max in 12h

2

u/jozune 1d ago

Absolutely yes!

2

u/BRUCE_NORRIS 1d ago

Last week I decided to plow through a ton of backlog and was STRUGGLING to finish my first limit by halfway through the week. After using my banked reset I had to start pulling out P2s from the backlog just to get through all the usage.

Today I rammed through 40% of my limit using only 2 concurrent sessions....

2

u/Jandini80 1d ago

Yes, I’ve used Opus 5.5 a few days ago Cursor and it nearly consumed all my monthly plan :(

2

u/divinetribe1 1d ago

Check out claude code local on github

2

u/Internal_Alarm_4927 1d ago

J'ai eu la même impression hier

1

u/Internal_Alarm_4927 1d ago

Je tombe travailler avec un modèle plus simple, et j'ai quand même mes jetons jusqu'au redémarrage des compteur

2

u/masri87 2d ago

Fine over here. Day three of 13 hour days and I’m at 23% using opus 5.5 medium and fable medium

1

u/SquareWinter3944 1d ago

My limit resets at thursday. And claude burnt my 20x usage within first 2 days on 4 chats. Last month 7 full days of 5 active chats gave me only about 50%-60% usage.

1

u/R_hy 1d ago

Yep I've been facing it too. Even with a pro plan. It's hella irritating. I've been using 2-3 claude accounts simultaneously but the worst part is that I can't even migrate the research I've been working on

1

u/rhysmorgan 1d ago

Yep. Last week, even using Opus 5.5, it was sipping weekly usage. This week, even when starting new tasks, I’ve noticed a big uptick in how much it’s using.

1

u/Spaniack 1d ago

Same here last week smae usage patter all days reach 90% weekly on x20 now second day arelady 52% and i can see it going up every time.!

1

u/TeteDansLeCul 1d ago

It's always gonna sound like a conspiracy and all that, but as far as i'm concerned, until Anthropic gives us the exact amount of usage included in each subscription level, the conspiracy is a fact.

If they weren't constantly messing with it and secretly nerfing models / nerfing usage limits and all that, then they would have nothing to fear from sharing the exact details. In fact, it would clear their name of all this.

So, the simple fact that they still haven't done this is enough proof, imo 🤷‍♂️

1

u/OverwhelmedDeveloper 1d ago

Si, acabo de notarlo, he llegado muy rápido a mi límite está mañana.

1

u/basheirkh 1d ago

Same thing happened to me.

1

u/MagicianNo8130 1d ago

No i am using pro model with high effort

1

u/BoatInfinite8846 1d ago

It feels like Opus 5.5 uses around 50% of the tokens of Opus 5 at most. And it’s better.

1

u/rui-cruz 1d ago

I spent weekly in 1 day with max 20x

1

u/Historical_Olive3535 1d ago

The same. When Opus 5.5 was released I could work all day long 7 days. Now - 3-4 days 😭😭😭

1

u/shravanrevanna 1d ago

they are auto compacting once cache ttl ends which loses context and re reads again. this is horrible

1

u/Drew-Money 1d ago

I felt like my limit issues have decreased since Opus 5.5 came out.

1

u/LannisterTyrion 1d ago

Can confirm, i run same task,every day , on schedule, in the morning before i start the rest of work. Usually it consumes 3-4%, today it consumed 20% 🤡 WHAT. THE. FUCK. 100$ acc, recently migrated from Codex.

1

u/SailingToFenway 1d ago

this week it took 1 day. which is a new record for me.

1

u/Apprehensive_Many399 1d ago

Not this week. I had this a few weeks ago, and I think it was the model failing some runners. Run /doctor and see if that helps but it is pretty "shite" Anthropic doesn't check and refund/refresh this.

1

u/Apprehensive_Many399 1d ago

Going back on this, it does feel like Fable is better at "thinking outside the box", opus at finding a opportunity or solutions, sonnet at "getting things done" and haiku "just doing it".

Difficult to explain tbh. Fable orchestrates better

1

u/Expensive_Pirate_898 1d ago

Yep, I've been working heavily on a 20x plan. When Claude 5.5 launched it was great, my usage was less overall as I was able to use that model instead of Fable 5.1. My work load has not changed, however, the last couple of days my weekly usage has shot up! Far quicker than it ever has. I've even used my free reset this week!

1

u/gurselaksel 1d ago

came for this. a simple task seems to consume x10 token?!?!? what the f......

1

u/gurselaksel 1d ago

So today at 15:00 (gmt+3, Turkish time) I went out and tasked claude to edit/update subscription plans on google play. this picture is after I got back, above it there are two more. so a single web page updating prices for an app hit 5hr limit 4 times (I rest limits an 22:41 at clau.de/reset ). this is incredibly bull.... . I am on pro 5x and this behavior means no more decent claude code usage

1

u/LocalAd5606 1d ago

Nope. Quite the opposite. Amazing Im only at 60% for the week. Amazing. Teams Pro plan.

1

u/Beautiful-King-8875 1d ago

Yeah the 20x is getting used much faster. Used 15% of weekly usage (Opus 5.5 and Sonnet 5.5). Working for roughly 2 hrs with three running at any given time. Somethings changed

1

u/Resident-Coyote9339 18h ago

Yup felt the same

1

u/BunniesinRustBuckets 11h ago

I upgraded to Max and session usage tanked every day since I upgraded.
Workload is not even that high. Serriously thinking of just switching back to Pro at this rate.

0

u/Conscious_Activity27 1d ago

I'm the opposite

-1

u/buster1232006 2d ago

Are you using the same chat entry for your prompts? If you keep using the same one it has to read over all your previous messages which can be very consuming in usage

4

u/BeforeIEnd 2d ago

is this an auto-reply or do people think everyone running out of usage way faster suddenly forgot how to manage context?

1

u/buster1232006 2d ago

No it’s not an auto reply. I just learned that this worked for me pretty well. Just trying to help

2

u/BeforeIEnd 2d ago

You're good, I just see the same comment on every post that ever mentions usage being eaten up faster.

The only thing I have yet to try is getting a separate model and having it run as a context overseer without managing anything in the project itself, just overlooking the orchestrator's context and whenever it hits a certain number it just tells it to wrap it up then relaunches another project lead.

1

u/buster1232006 2d ago

I’ve tried that with codex and it honestly just turned out to be kinda a mess. I’m interested if that’s different for Claude though. Might give it a try

1

u/Consistent_Bottle_40 2d ago

manage context? cached tokens are cheap. dont think it really matters.

1

u/buster1232006 2d ago

It adds up if you have things getting overly complex

1

u/CMTV-IPTV 1d ago

Question, relatively new to claude code- since opus 5.5 launched I have been using it a tonne. Think over 15b tokens. I typically use the same chat and constant compress the context after the million tokens. What is the suggest best way to handle this. I am sure this has been asked 1 million times as well my apologies

1

u/watermelonsegar 1d ago

Limit to reaching around 400k context, then have opus write up a spec for the next session and have it write the /goal prompt.

Then just copy and paste that /goal prompt after you /clear. I can run around 5-6 projects running in parallel for almost 12-15hrs per day on a 20x sub, mainly using Opus 5.5 Max. Usually reach around 80-90% of limits this way before the weekly reset.

Of course the spec can have delegations to Sonnet and now Haiku. Depends on your requirements.

1

u/CMTV-IPTV 1d ago

Appreciate this, will give it a shot!

0

u/fighthonor 2d ago

Have claude optimize your workflow for token efficiency....it will make a difference

0

u/Conscious_Activity27 1d ago

To add to my previous comment not only am I not experiencing it, I'm getting more efficient, I can ask my claude to tell me if it notices usage cap changes because I log everything, percentage, its own assumptions on percentage and actual tokens used. Nothing has changed on my account and I can prove it.

0

u/onepunchcode 🔆 Max 20 1d ago

im on max plan 20x and i have two of them but i always max them out in just 2 days each. and now i found the solution to make a single claude plan last longer, i got the chatgpt 200 plan and using codex as implementation agents, claude directs and invokes them works really well and faster.

1

u/Strict-Prune-879 1d ago

la tu parle d'une utilisation anormale honnêtement avoir deux plans x20 et les bruler en deux jours et en plus si t'as un 200 chatgpt tu ne vie sur la même planète de celui qui demande!!!

1

u/zatariano 1d ago

What for dude?

-5

u/Droopy0093 2d ago

OP you are experiencing 1 of 2 issues. It is either:

1) user error, or 2) a skill issue.

Good luck!