r/ClaudeAI Full-time developer Jul 14 '25

Other Anthropic didnt rate limit us, they got too popular

Post image

Lot of people have been accusing anthropic of making Claude models dumber, or changing how much we get on 5x or 20x plan etc. Lots of pretty wild speculation. This is the first time ive started seeing this from Claude and its a symptom of what I beleive has been happening lately, the backend is just overloaded so all work is costing more tokens and there is a quality dip due to lack of resources.

I could be TOTALLY wrong but I don't think Anthropic as a company has been doing anything nefarious or underhanded, I just dont think they were prepared for the absolute RUSH of use that has come with the latest press about Claude and the garbage with other ai based IDEs and their cost models changing, so people have been jumping ship and coming here.

Hopefully they will be able to build up infrastructure quickly to take on the load, but that is always a risky proposition for big tech companies that I don't envy.

159 Upvotes

138 comments sorted by

View all comments

208

u/Illustrious-Ship619 Jul 14 '25

We don’t need “speculation.” We need accountability.

I’m on the $200/month x20 Max plan, and right now — we’re getting x5 performance at best. When I subscribed, Anthropic clearly promised ~800+ messages per 5h window. That was the selling point. Now? Hard limit after 1.5 hours, even with light usage, no agents, one terminal.

This isn’t “backend overload.” This is a silent downgrade. No communication. No warning. No transparency.

If you charge for Premium, deliver Premium. Don’t shift the burden onto loyal paying users.

12

u/TheOneNeartheTop Jul 14 '25

It cracks me up when people use an AI to write their complaints about another AI. Just the mental image of someone flipping their keyboard at Claude and then going to chat.openai.com and saying write me a Reddit comment about how Claude code sucks so much. Then copying and pasting it into Reddit.

3

u/SpecialistWinter4376 Jul 15 '25

It cracks me up even more is that we are just becoming glorified messengers every second by second. What disturbs me even more that I have had more conversations with ai than I had with a human in the last 6 months. And the darkest truth. It’s like Heinz doofentmertz getting access to gru’s minion with phineas and ferb like brains. It’s just everything everywhere all at once. It was chaotic before now it’s On a different plane. I am feeling like just like we diversify our stock portfolio. We need to do the same for all ai models and ux. Like instead of getting a anthropic 200$ subcribtion. We get cursor+ cc+ copilot+ cline+ codex. 100$ at max. Vscode to integrate all that. Use all the superpowers from them. Probably loses all the weakness as well. And the most important one. Treat it as a coworker. Not as a slave. Think we are close to a perfected problem solving machine.

1

u/budz Jul 15 '25

You're right. I think I see the issue.

1

u/squareboxrox Full-time developer Jul 15 '25

People are becoming too lazy to write their own comments now

1

u/KendrickCP Nov 12 '25

Probably because theys uses bads grammars.

22

u/Mr_Hyper_Focus Jul 14 '25

I just find it hard to believe that’s light usage. I was on the 5x plan last month and never hit a limit.(

Now as of recent I’m on the $20 plan, and I still haven’t hit a hard limit.

I feel like everyone’s perception of light use is very different. What does your ccusage app say?

22

u/[deleted] Jul 14 '25

[deleted]

4

u/oneshotmind Jul 15 '25

Single terminal, 2 hours of opus on 200 is hitting limit hard limit for me

3

u/Mr_Hyper_Focus Jul 14 '25

I’m not saying youre imagining less usage. I think really only Anthropic can confirm if that’s true.

What I’m saying is that we need to go off of actual Usage numbers not just what people feel. Last week some dude on the Cursor forum said he did “light edits”, and his account page said he used 60 million tokens lol.

8

u/AbsurdWallaby Jul 14 '25

I went back and ran the same prompts on past projects to directly compare apples to apples and there's an obvious issue.

-12

u/Mr_Hyper_Focus Jul 15 '25

Lay it out in detail otherwise this comment is no better than. “cLaUdE bAD! PApa AnThrOpIC liE!”

What was the prompt? How many tokens was it? Why was it worse? What’s the issue? How many message did you sent and how long were they before you hit the limit? What’s your ccusage say? In detail, how did it perform different from what you expected?

5

u/AbsurdWallaby Jul 15 '25 edited Jul 15 '25

With all things being constant since this is the same exact prompt for the same exact project a month ago, the prompt was to execute a particular sprint from a specific phase of a development plan file using the todo tools and task tools with a sub agent per sprint task. One single prompt to execute the sprint by asking Claude to "execute sprint x of @devplan.md"

7 folders created named CDCUsersComputerDesktop

Reaching Opus limits at the fifth task in that todo list for this sprint, each task averaging 70k tokens and taking about 15 minutes to complete. This isn't even a complicated sprint, it's just to make a UI mockup in a fresh project folder.

This particular type of skeptical tonality you are using would be vastly more helpful if you also contribute your own test results along with the requested methodology.

-1

u/Mr_Hyper_Focus Jul 15 '25

Thanks actually a lot more information and that’s really helpful information that we can use as a community to set a baseline.

I don’t have an issue with the way it’s working so what would I even demonstrate? And I definitely cant offer any sort of specific advice without knowing the issue in detail. How could anyone?

Model degradation delusion syndrome is so prevalent in the ai community that it’s become a community meme. So when these things happen, deciphering the two situations becomes almost impossible if the OP/user/person having the issue just says that it didn’t do what it did last week. Nobody knows what that means.

So I’m not trying to have a tone with anyone.

Maybe some prompt adjustments in the “sprint” or spec prompt could be adjusted to be more clear for the model to follow. You could run some quick tests to see the success rate differences in the different prompts.

1

u/AbsurdWallaby Jul 15 '25

Thanks I figured the only real way to know if there's been a limitation issue is by using the task tool with assigned sub agents to get 'official' metrics and comparing them to the exact same previous project workflows. Response quality is a different beast entirely and significantly more subjective, so there's no use rallying there.

Even if you aren't noticing limit issues, perhaps try and open a past project and run the same prompts multiple times to test your new upperbounds.

7

u/Illustrious-Ship619 Jul 14 '25

You're missing the point.

This isn’t about “feelings” — this is measurable. People on x20 used to get 800+ Opus messages over 4–5 hours, consistently. Now it’s hard stop after 1.5 hours — same project, same terminal, no agents, no abuse.

We’re not just talking about “some dude using 60M tokens.”
We’re talking dozens of users reporting the same degradation across different setups, many of whom track usage carefully with /model opus and ccusage.

So either:

  1. Anthropic silently nerfed usage
  2. Or their backend is failing and burning tokens 10x faster

Either way — it’s unacceptable for a $200/mo product.

8

u/Mr_Hyper_Focus Jul 15 '25

I appreciate you trying to make your words clear. But if I wanted to talk to Claude or ChatGPT about this I woulda done it.

I’m not fighting you on this though….all I said was if you were going to complain it had to be clear how much usage you’re using, since that’s clearly subjective.

2

u/Crafty-Wonder-7509 Jul 15 '25

You sound like a die-hard fanboy. No idea where ur ego comes to defend a multi billion dollar company for ripping you off.

1

u/Mr_Hyper_Focus Jul 15 '25

I’m not a fan boy lol. I have no loyalty to any of these brands I use them all. I’ve just been in these communities enough to know what I’m seeing.

3

u/Crafty-Wonder-7509 Jul 15 '25

I get that, but you're clearing defending that what apparently most people are witnessing is "subjective". I don't know about you, but some people here, including me use CC along with some other tools on a daily basis, and I personally witnessed a drastically reduced limit. It doesn't even take me half of the time it used to, to hit limits. I also work on hobby projects on the side and there its easier to halt it, as I usually have it process a few csv/json files and restructure some files, and it hits limits way earlier. It's not subjective if you can prove it, if not through personal views, through ccusage or any other token calculation tool. But hey, thats my take on it.

1

u/d33mx Jul 16 '25

lets just admit anthropic makes it random.

end of afternoon for a 2 hours or so, london time, i know claude will invariably suck. like, not sonnet 1.5 mode. could just be me. but it freaking does happen invariably.

beyond that, I rarely reach opus limit. And when I do (happened yesterday), it also feels random.

-6

u/Illustrious-Ship619 Jul 14 '25

Exactly. This is not a "perception" issue — it's a silent downgrade.

We paid for x20 Max, and Anthropic promised ~800+ Opus messages per 5 hours.

Now? We get 1.5 hours max, working in one terminal, no agents, no abuse — and then get cut off.

This isn’t premium. It’s a crippled x5 pretending to be x20. And Anthropic stays completely silent.

No warning. No notice. No transparency.

If x20 now gives us less than before — what the hell are we paying for?

#Claude #Anthropic #x20 #Opus #usagelimit #LLM #nerf

6

u/Revotheory Jul 14 '25

I was on the $20 plan last month and upgraded to 5x this month to try out Opus. Never hit a limit on the $20 plan. Yesterday and today I hit my Opus limit within 30min. Today it was like 3 prompts. If it stays the same I’m going back to the $20 plan. Barely being able to use Opus isn’t worth an extra $80.

5

u/Mr_Hyper_Focus Jul 14 '25 edited Jul 14 '25

I have always hit the opus limit after a short number of messages on the 5x plan. You’re only allowed to use OPUS for the first 20 percent of your usage, which really isn’t very much at all considering opus is 5x the price at $75 per 1M output tokens. You don’t get a lot of opus use(prompt wise) on the 5x plan.

If you do the math there, that means that you get 10x more opus usage on the $200 plan than you do on the $100 plan.

1

u/Revotheory Jul 14 '25

That’s good to know. So really you need the 20x plan to get much out of it. I’ve had one issue Sonnet was going in circles on and I was struggling to understand and Opus 1 shot it. It’s nice but not sure I’ll run into stuff like that on the regular.

2

u/Mr_Hyper_Focus Jul 14 '25

Yea you really need that 20x plan to get a lot of opus.

I always just keep it set to sonnet, that way when I have a hard problem I can put it in plan mode, switch to OPUS. And then let Sonnet execute.

Default mode stays on OPUS until you use your 20 percent allotment for the session.

1

u/Severe-Video3763 Jul 15 '25

Same experiences with Opus. On paper it seems to be so close in quality with Sonnet but my experience is that it’s vastly superior. I’d rather get 2h of Opus than 5h of sonnet (which is roughly what I get on the max 20 plan using it non stop)

1

u/Illustrious-Ship619 Jul 14 '25

Sure, I did the math.

And guess what? Even on the $200 x20 plan, Opus now hits the cap in 1.5 hours. That’s nowhere near the 800+ messages per 5h window Anthropic originally promised.

So either:

  • We’re not getting the 10x value anymore
  • Or Anthropic silently changed the policy

It’s not just about pricing per million tokens — it’s about broken expectations and a silent downgrade across the board. That’s what users are upset about.

4

u/Mr_Hyper_Focus Jul 15 '25

You were never promised 800 messages per session. It says 200-800 per session. It’s always been based on token usage. And it still says that’s for an average user.

There might actually be an issue, but people acting like this are part of the problem. And they cloud the actual issue around BS

2

u/stormblaz Full-time developer Jul 15 '25

Maybe they exploded in popularity recently, or people left cursor since its gotten messy.

1

u/RunJumpJump Jul 14 '25

Opus is still very limited on 5x. I've never hit a limit with Sonnet on 5x, though.

2

u/Longjumping-Bread805 Jul 15 '25

They definitely nerf down Opus for either plan.

1

u/No-Row-Boat Jul 15 '25

Really? I spent my entire quote in 15 mins on a $20 plan just to get the assignment clarified for it to give meaningful results. It's the reason why I dropped the subscription, never had that once happen at Chatgpt.

5

u/m0strils Jul 14 '25

You are feeding in a huge amount of tokens each time I assume. Let's be truthful with the critiques.

3

u/KenosisConjunctio Jul 15 '25

Some people don’t realise that it’s sending wayyyyy more than just what you type in. The whole history of the chat will be sent in some setups

3

u/Efficient_Ad_4162 Jul 14 '25

The accountability (or lack thereof) is you continuing to pay them.

5

u/CrazyFree4525 Jul 14 '25

It's worth pointing out that the "800 messages" thing is not as generous as it sounds. Each message Claude sends you or you send it counts so one prompt can easily generate many dozens of messages.

800 messages != 800 prompts. 800 messages is actually FAR less than 800 prompts.

6

u/Illustrious-Ship619 Jul 14 '25

Yes, we know it's not 800 prompts — but the point is: a few weeks ago we could work 3–5 hours straight in Opus without hitting any limits.

Now? Even x20 plans are capped after 1.5–2h in a single terminal, with no abuse, no agents.

So either the token cost silently increased, or the usage limits were quietly slashed. Either way — this isn’t what we paid for.

1

u/Jolly_Painting5500 Jul 16 '25

Learn how to code buddy

1

u/ScaryGazelle2875 Jul 15 '25

I'm thiking the latter. Many in subred warned against abuse of people running hundreds of agents and possibilities that it might trigger overload or worse, Anthropic silently introduce harder usage limits.. Seems its happening

1

u/Crafty-Wonder-7509 Jul 15 '25

Those 800 messages, if true, would be higher than you think, it would mean it could send a message for a good more than 12 min straight using a message per second. My Claude on 5x, roughly sends a "message" (if you consider the cli texts) every other good 10 seconds, sometimes or or less. So this would suddenly give you a good 2h anyway. But, at the end of the day, never ever blame paying customers, they got more customers? They're welcome, but scale up the operations then, there is 0 loyalty in this industry.

4

u/njmh Jul 14 '25

Your comment reads exactly like it was written by AI. Does every little bit of text you write have to go through an LLM?

Why can’t people just chat like real people to other real people ffs?

2

u/256BitChris Jul 14 '25

I've still not hit any limits you're ranting about and I'm a moderate to heavy user.

2

u/fynn34 Jul 15 '25

Those 800 messages weren’t promised as opus which I’m sure you are using. I know because I checked before I bought it

2

u/Lorevi Jul 15 '25

Anthropic clearly promised ~800+ messages per 5h window.

There was nothing 'clear' about their promises. They were incredibly vague about the rate limit from the start, so they could arbitrarily change the rate limit when it suited them. People just didn't care about the complete lack of transparency when the rate limit felt high. 

Their promises have always been in the context of number of 'messages', made meaningless by the caveat that depending on how large your messages are will change how many messages you get. 

I.E the real token rate limit is anywhere between saying hi to the model 45 times and sending your entire project over with complex queries 45 times. (for pro. 5x 20x for max). 

This entire pricing bait and switch was so obviously going to happen because every other ai coding app has also done it lmao. You operate at a loss for a while to get customers then throttle usage when your runway runs out. And you just have to compare basic Claude code usage to api pricing to know they were losing money on it. Cursor did it last month now it's cc's turn. 

Taking advantage of the product while it's cheap knowing it won't last is one thing. Acting entitled like the company should lose money on your behalf of another lmao. And unlike cursor CC didn't actually promise to deliver anything to you idiots who bought the annual plan. 

1

u/garnered_wisdom Jul 15 '25

Not my experience at all. Heavy opus user, with agents.

1

u/shadyringtone Jul 15 '25

Then don’t renew until they fix it. I don’t know why you’re expecting them to be perfect

1

u/moretti85 Jul 15 '25

Have you tried using ccusage to check your usage report?

Your assumptions might be off, it's never been about "hours" or "messages". Usage is based on token consumption, and some tasks consume far more tokens than others.

Take a look at the past week when your usage was higher and compare the token usage, that might give you a better sense of what's going on

1

u/Severe-Video3763 Jul 15 '25

We don’t need speculation we need proof…

1

u/No_Pressure_3675 Jul 16 '25

Yep, bought it a week ago. Went from coding 8-10 hours a day to now coding for 1-2 hours then having to wait 4 hours.

Only using sonnet 4 too, splitting the load between cursor and Claude CLI, I just want to code all day mang :(

1

u/[deleted] Jul 14 '25

[deleted]

-1

u/Illustrious-Ship619 Jul 14 '25

That sounds completely fake, honestly. Right now x20 users get locked out after just 1.5–2h of moderate Opus 4 use — in a single terminal, no agents, no abuse.

There’s no way you spent 12+ hours on Opus 4 today without hitting a wall — unless you were idling or using it like once per hour.

Let’s be real.

3

u/fujimonster Experienced Developer Jul 15 '25

Let’s be real then — I pay for the $200 plan , I use it all day and everyday .  I beat the hell out of it and have yet to hit any limits .  Opus 24x7.    Don’t care if anyone believes me , I know my usage and it’s not being limited in anyway .  There might be issues in the lesser plan , but here on mt Olympus I’m not seeing any . 

-1

u/WholeMilkElitist Jul 15 '25

They really need to pull something like a Geforce Now where if they can't maintain capacity for existing customers then they pause new subscriptions for a bit.

This shit is mad annoying for us who use this tool daily (and consistently pay for it) while some chuckle fuck vibe coders are just trying the new hyped thing out until they move on next month.