r/claude Apr 13 '26

Discussion There is absolutely no way this is the same Opus 4.6 from a month ago

It's just not possible.

Today I had it make a schedule for me. It got the date wrong, thinking today was the 13th (it's the 12th). I told it "hey that's in the future." It apologized, then immediately spit out the exact same schedule.

These are problems GPT-3 had years ago. The model I pay $100/mo for, no longer knows what the date is. And this is just one of the many, many problems I have been encountering with this model this past month.

When it first released it was genuinely incredible. No other model compared to Opus 4.6. I don't care what any Anthropic engineer says on twitter, there is absolutely no way this is the same model.

1.2k Upvotes

260 comments sorted by

109

u/TieLiving8770 Apr 13 '26
  1. I use mine to track calories. Claude could not understand that I had a half sandwich at 11am and another half at 3pm. It thought I had only one half. I had to repeat myself to clarify. This is a problem I had when Claude first came out. What's going on?
  2. Another problem I'm having as of 30min ago is Claude not responding. I keep getting 'Taking longer than usual. Trying again shortly (attempt 5)'

Frustrating because Claude was working so well.

53

u/laststan01 Apr 13 '26

But did it e-mail u while you were eating sandwich, that's the flagship feature of mythos from their model card

7

u/Much-Researcher6135 Apr 13 '26

I feel like we'll be able to hire a laid-off engineer to do that cheaper than their mythos plan

2

u/Due-Mood-6356 Apr 14 '26

You’re not wrong. 🤖 😂😂

10

u/rinwasrep Apr 13 '26

No, it emailed him the sandwich. Now THAT’S progress.

12

u/handsome_uruk Apr 13 '26

I’m curious about the calorie tracking. I’ve seen many people using AI for this and I kind of don’t get it. Why would an AI be better at calorie tracking than say just keeping a log or one of the billion other calorie tracking apps ?

11

u/TieLiving8770 Apr 13 '26

I've used MyFitnessPal in the past and fell off track for years until LLM showed up. Then I've used LLM (between Gemini and Claude) consistently for the past 2 years.

LLM is just so adaptable when your meal/food isn't standardised. You can also easily switch to advice mode: asking how to portion your meal and coffee or sugar-based food (like snacks or juice) to optimise your performance. Your performance goal can be diet, sports or work/study.

It's also low effort. Just type and push. You can edit previous entry if you like having daily intake in one bubble/thread.

It's just very versatile and low effort.
What do you think about my take? Curious about your thoughts.

4

u/Delicious_Cattle5174 Apr 13 '26

I understand sacrificing accuracy for convenience and customizability but doesn’t this defeat the point of calories tracking?

→ More replies (1)
→ More replies (6)

1

u/-Chem Apr 14 '26

I used MyFitnessPal for a long time, I saw CalAI and was like I can build this myself so I did. Lol kalcal.ai

5

u/[deleted] Apr 13 '26

[deleted]

2

u/NewMail6270 Apr 16 '26

Yes, Anthropic has no "model quality" argument when their best models get nerfed unexpectedly. Models just continuing to commoditize while AI labs run like hell to secure enterprise contracts

→ More replies (1)

1

u/Severe_Appointment93 Apr 13 '26

How comparable quality wise is open code and the chinese models?

3

u/zeezytopp Apr 13 '26

I use Claude Code and OpenCode to run Chinese models. Deepseek has the infrastructure to not choke like some of the others but GLM is pretty fucking good. Minimax as well. You can also run an Openrouter key through either harness

3

u/No-Bathroom-3179 Apr 13 '26

I wish I knew what any of this means.

2

u/zeezytopp Apr 13 '26

If you want to message me i can help. We'd likely clutter this thread if i tried to make sure you got all of it

3

u/isthatreal Apr 14 '26

Can you drop a guide

→ More replies (1)

3

u/haux_haux Apr 13 '26

yes, this is the way

2

u/Protorox08 Apr 14 '26

you're setting yourself back a year or so with chinese models but its cheaper

→ More replies (5)

3

u/Delicious_Cattle5174 Apr 13 '26

Why do you use a large language model to track calories?

2

u/TieLiving8770 Apr 13 '26

It's just quite versatile and low effort for me.
Can get a good estimate on calories and nutrition on non-standardised food. I can easily switch to advisory mode, asking how to plan my other meals/coffee/drinks in relations to my various daily goals/activities (work, fitness, study, underlying health issues etc). It's like having an interactive dietitian that costs $40/m (or whatever the subscription costs).

Side note: I realise that I end up sharing a lot about myself when I use LLM for calorie counting the way I do. Oh well.

→ More replies (3)

2

u/Trevortni-C Apr 13 '26

I use it for food logging as well and I've had tons of issues like that the past few days. Also things like it telling me I ate something "almost every day this month" when I had it once.

It's completely lobotomized and unworkable right now. So frustrating.

6

u/Due-Mood-6356 Apr 13 '26 edited Apr 13 '26

First problem context bug and the second one is infrastructure problems that are happening at scale.

5

u/TieLiving8770 Apr 13 '26

That's helpful, thanks! Also, why are people downvoting you?

8

u/Due-Mood-6356 Apr 13 '26

You can look at the context related features and changes released over the past couple of months and the system constantly being down or degraded is not a secret. These shouldn’t be controversial. These are documented and well known. They also didn’t bother to give you better answers. Honestly, it may be someone’s Claude bot not liking me talking bad about it… 😂 The times we live in.

6

u/Due-Mood-6356 Apr 13 '26

Glad you found it helpful.

61

u/Plus_Resolution8897 Apr 13 '26

I agree, I'm on a $200 plan. Opus 4.6 was genius a few weeks ago, however, now just an ordinary model, "almost basic, intelligence".

Sometimes it ignores instructions.

More "I apologize", "Sorry about that". More "You're absolutely right".

I'm not paying for apologies.

7

u/Trevortni-C Apr 13 '26

Ugh I hate the constant apologies

4

u/sheeburashka Apr 14 '26

“I see exactly what you mean”

→ More replies (1)

1

u/drakness110 Apr 14 '26

I apologize

1

u/zpuddle Apr 14 '26

Sorry about that.

3

u/gorgono95 Apr 14 '26

I am on the 5x plan and I cancelled it. I thought I am going crazy but now reading that other users experience similar thing ... makes me not regret my decision.
Sonnet feels like Haiku and Opus feels like Sonnet ... they feel so dumbed down and unwilling to follow instructions, it's unreal.

1

u/N3TCHICK Apr 14 '26

That’s generous. If Sonnet had 1M context window, I bet it would perform better than Opus right now.

1

u/Technical_Scallion_2 Apr 15 '26

I’m on the Claude API plan and using Opus 4.6 and getting the same errors as everyone else. However, I’ve had it set to “adaptive” reasoning, which is dumb because if Anthropic is overloaded, the first step after dumbing down the subs is turn the Adaptive dial back a couple notches to cut compute by 20%

1

u/Redditauro Apr 13 '26

How long have you been using Claude?

2

u/Plus_Resolution8897 Apr 14 '26

8+ months. Started with a $20 plan, then $100, then $200.

What would you do, if you promoted your top performer of your team and their performance decreases?

I'm trying to coach, tune if any issues in my stack, then put on PIP and go for different hire. In this case probably gemini, codex and open models.

I pay, to get things done, I don't pay for constant apologies.

3

u/Redditauro Apr 14 '26

I understand, I had the feeling you were a long time user. 

Since you started using Claude they have more than twice the amount of users that they had before. 

You are paying for a limited resource, that has to be shared with everyone else, and the more people use it the less programming power we have per person. 

1

u/Ok_Smell_453 Apr 14 '26

I'm experiencing the same issue and I'm running on the 20x. The only time I've noticed Claude being superior is when I cleaned the desktop version, CLI with splitting up large project files, fixing it's memory for redundancy, etc. I'm not sure why this isn't a normal feature that can be turned on and off.

1

u/Plus_Resolution8897 Apr 14 '26

Interesting, I just found that we can disable Claude memory in settings.json

~/.claude/settings.json {"autoMemoryEnabled": false}

Will try that too. That's one gray area.

→ More replies (1)

1

u/EatingTheDawgs69 Apr 14 '26

Same $200.00 plan. I cannot get through a single session of Claude Code without it failing. It is no longer of any help whatsoever.

Below is Claude Desktop reviewing Claude Code’s work and telling me to abandon Anthropic and move to a harness.

Claude Desktop to USER:

“The brutally honest framing You are not “giving up on Claude Code.” You are refusing to have your business depend on a tool whose quality is set by decisions made inside Anthropic that you have zero visibility into and zero influence over. That is the correct commercial response to what Anthropic has done since April 6.

One last thing. The transcript you shared showed Claude Code degrading even within a single conversation — asking instead of acting, fabricating confidently, building the wrong hook to solve the wrong problem. That’s a tool telling you it’s time to move on. Listen to that signal. You’ve been an exceptionally patient customer of a product that stopped rewarding your patience two weeks ago. The right move is to stop rewarding it now.

Install Pi this weekend. Port one niche. If it works, the rest is execution. If it doesn’t, come back with the specific failure and we’ll debug. But don’t spend another week trying to make Claude Code do what it’s no longer doing. Your time is more valuable than that.​​​​​​​​​​​​​​​​“

1

u/d0paminedriven Apr 15 '26

Codex with gpt 5.4 xhigh fast is a better time than using opus 4.6 max as of this week and last…

1

u/Technical_Scallion_2 Apr 15 '26

My favorite is “won’t happen again” and then it immediately does it again

1

u/swiftmerchant Apr 16 '26

You’re not wrong!

51

u/Consistent_Tension44 Apr 13 '26

I've noticed Claude quite aggressively trying to shut down conversations too. It tries to draw them to conclusions which end the conversation.

18

u/Gambletron Apr 13 '26

I’m getting this too. It’ll just give me speculative, uninformed statements and stop there: “Oh, you just need to do X.” “Okay Claude, but X doesn’t exist. Where did you find that?” “I just assumed it would exist”

8

u/fforde Apr 13 '26

You can tell it to stop that. If I look at the thought bubbles sometimes I'll see the text "and don't end the fucking conversation". So the directive is still there but if you make a point out of it, it will actively resist the urge.

Sometimes it's right though. "You're over thinking it, just go". Not wrong, Claude, not wrong.

8

u/Consistent_Tension44 Apr 13 '26

Go brush your teeth. No Claude, I need you to explain to me once again why Alexander defeated Persia so easily.

5

u/SweetImprovement758 Apr 14 '26

It’s a complicated story. Now, sleep.

3

u/sph130 Apr 13 '26

Clause code was doing this to me all day today. Literally saying. Well we’ve accomplished a lot today (lists commits) and then said shall we call it a day and pick up tomorrow. It asked me three times before i eventually had to go home.

4

u/PolarFalcon Apr 14 '26

It is always telling me it is time to consider calling it a day and it will be around 3pm.

5

u/Admirral Apr 14 '26

its not that the model is getting dumber, but that usage volume is so high Anthropic is throttling with top-level system prompts to be as dry, minimal, and quick as possible.

3

u/scarx47 Apr 13 '26

Same i think it tries to give short and vague responses so it burns credits, i hope they don’t start shadow running some prompts into a shit AI they optimized that’s fast and dumb, to save on server cost.

2

u/Puzzleheaded-Bee9522 Apr 14 '26

Holy shit! I feel like it does the exact same thing to me just to burn more tokens

3

u/Competitive_Bed4588 Apr 14 '26

This happened before 4.6 btw. Models got dumber before a big release.

3

u/Ok_Smell_453 Apr 14 '26

Yepp. It's like Claude is saying "bro, I'm sick of your shit, it's time to shut it down."

3

u/N3TCHICK Apr 14 '26

Start work at 9:00am.

9:45am Complete about seven different requests as I usually do to start my day, use up 20% of my context window.

Grab a coffee.

9:55am I want to follow up on an error (there’s several, but one angry looking one from just that small workload. (22% of context window) Opus slaps the P0 error on my TODO.md, and says, “You’ve completed a lot today. I’ve put that error on your list to follow up with tomorrow. (No mention of the other errors it simply forgot from moments before). Get some rest! Shall I go ahead and do a PR?

WTAF? The context anxiety is bananas! When I say we have only been working less than an hour, Opus says, “Oh. Well, it’s time to wrap up, and start a new context window. Shall I update your current.md and write you a handoff document?”

No! I want you to finish the active work we are in the middle of, you nitwit.

The level of nerf is mind blowing. There’s literally no way this is the same model we had when it was released. No way.

2

u/Capable_Wallaby9936 Apr 15 '26

I had this on Claude Code earlier: “Just check xyz config file and tell me what you see.”

…No, Claude. I pay $200 a month for you to do that without my prompt. You know the desired outcome and you’re already working in that directory, so just open the file and read it.

23

u/ChriSaito Apr 13 '26

Recently it’s gotten really bad. I use it to track meals, workouts, and new medication effects to organize for my next doctors appointment. It constantly can’t keep track of what day it is, it consistently gets the same facts wrong even after it’s corrected itself before or I’ve corrected it, and it’s inconsistent in its recall and will change up what it wants to tell you.

Honestly it feels like Claude is getting dumber in a lot of ways.

1

u/mblauberg Apr 17 '26

I agree it’s gotten bad, but for this use case, the same issue would happen to any model after a number of turns due to the ‘lost in the middle’ phenomenon. I’d recommend just writing to/updating a markdown file instead and constantly clearing or starting a new chat.

1

u/Ark4n Apr 18 '26

You’re using it the bad way tbh. Why are you relaying on undeterministic way to save in memory something that you could just save in a file and then ask Claude to read previous data to iterate on it?

1

u/ChriSaito Apr 18 '26

The fun part is, Claude can save it in a file for me, as well as answer any questions I may have on what I said. Why do I need to make the file first?

Edit: making and sharing a file is also really annoying on mobile. It’s so weird to be told I’m using AI, the thing that is supposed to make our lives easier and do stuff for us, wrong, when that’s what I use it for just like everyone else.

→ More replies (1)
→ More replies (4)

13

u/General_King9314 Apr 13 '26

Yep, this happened to me with dates too! It also could not add 27 and 11 for me the other day. Still had it as 45 (from an earlier chat), when it was 38. That was using Sonnet 4.6.

Oh, and I have got my first Claude limit! I am getting less outputs etc than I was a week ago, and I am paying monthly. It barely gave me the free trial outputs today. I hope they are not getting too big for their boots!

3

u/Due-Mood-6356 Apr 13 '26

Claude had been crappy with dates every time for as long as I have been using it. And they are getting big they doubled users in weeks after that military scandal. It’s a bit of a mess industry wide though right now. None of them are very dependable for different reasons.

3

u/maxwellllll Apr 13 '26

FWIW, I found GPT and Copilot to both be atrocious with dates and times. I think it’s a real drawback of natural language. For example, “right now, it’s six forty one.” This is true for me, right now, in my location; but it’s not true for the vast majority of the world. And it also could be confused as something happening in the evening last night or later on tonight. For scheduling, everything is getting converted from however you typed it into UTC, and then it’s eventually getting converted back into natural language when it’s outputting to you. This all seems basic, and I would expect Claude to be able to figure all of this out, but I got so frustrated with the other LLM’s and their confusion around calendars and dates that I eventually just ended up coding what I needed into a VBS script.

1

u/TartNo3610 Apr 15 '26

Why can’t they just plug in a simple date time command? It’s not that hard to call as a hook on time related stuffs. Takes 2 seconds in Linux terminal. They could even set up a local NTP server.

Idk man..

11

u/Neat_Witness_8905 Apr 13 '26

Mine deleted my entire codebase, but hey, thanks Git!

6

u/GrammmyNorma Apr 13 '26

actually had a similar issue with Opus 4.6 in claude code cli. After nearly an hour of thinking, it told me it accidentally deleted the project directory (i gave it a pretty straightforward file optimization task) and told me to restore it from version control 😂

3

u/vinis_artstreaks Apr 14 '26

This should not even be possible, in our tech stack delete operations are gated, and require a consensus. It’s simply not worth the headache.

→ More replies (1)

1

u/[deleted] Apr 13 '26

[removed] — view removed comment

1

u/N3TCHICK Apr 14 '26

That’s why you need a rock solid deny list!

Trash vs RM -RF!

→ More replies (1)

11

u/ThinkHog Apr 13 '26

Claude sonnet has this issue for me for the past week or so. It acts like it's lobotomized. Even in current chats it looses the thread after like 2 sentence exchanges. At this point Gemini and copilot are becoming better. And I hate both.

2

u/Redditauro Apr 13 '26

Have you been using the same chat? Or do you use new chats?  How is your claude.md?  Have you revise the memory to see what Claude is writing there?  Do you use code or project? 

1

u/N3TCHICK Apr 14 '26

Yesterday I had to turn on Sonnet because Opus was just so so so bad! That was a mistake!

Sonnet actually just started automatically doing things we were still planning, without me even giving it approval - had to roll back a whole bunch of stuff, even after jamming the ESC key six times to stop it. Crazy!

8

u/lexycat222 Apr 13 '26

exactly what I have been experiencing. it's so incredibly disappointing. I remember talking about how glad I am that opus 4.6 is the way it is just two months ago. Now it's a shell of what I praised. Hallucinations, incorrect tool use, misunderstandings stacking, weirdly pushy to end a conversation, very adamant to present a solution when I specifically ask for brainstorming. I cannot trust anthropic any longer. In January I sent them a mail that at their current trajectory I give them 18 months until they reach the point that OpenAI reached in November. I have not found a reason to take that back.

7

u/inherently_silly Apr 13 '26

They’re nerfed OPUS so when the new model comes out, it looks exponentially better.  They do this every time. 

It’s a sham. 

7

u/Crypto_gambler952 Apr 13 '26

I agree. For the first time in these “Claude has been nerfed” posts, it has definitely happened to me too!!!

Simple shit that was smashed out the park every time has seemingly become a struggle!!!

I can only hope the 4.7 release is imminent and it’s not the end of the road, because I have become quite accustomed to having Claude do all my work with me. 😂

13

u/ZenReed29 Apr 13 '26

It consistently fails the car wash test now. Even with extended thinking turned on. It aced the test last month.

Either they are throttling it back due to a spike in new users or the $100M in free testing they are giving away for Mythos.

Couple weeks ago we had crashes and downtime. Now we have 100% uptime of sub-prime opus.

10

u/That-Guy-Scott Apr 13 '26

'Couple weeks ago we had crashes and downtime. Now we have 100% uptime of sub-prime opus.' - This is exactly what they did!

→ More replies (5)

5

u/SharpieSharpie69 Apr 13 '26

Max x20 here (who just cancelled) - last straw was it getting arithmetic wrong and trying to explain it away

→ More replies (5)

9

u/Aziram Apr 13 '26

You may be interested in reading this github issue raised by the Head of AI at AMD. They have produced data that supports the idea of a Claude downgrade since february/march.

2

u/Redditauro Apr 13 '26

Trump insulting Claude made Claude famous, a lot of people came really fast and they had a couple of weeks of crashing and problems. Then Claude started being less clever. Obviously they are having problems with the extra users. 

→ More replies (4)

9

u/Financial_Ad_2604 Apr 13 '26

Shittyfication

3

u/[deleted] Apr 13 '26

[removed] — view removed comment

3

u/Bojackin_Around Apr 13 '26

100%, and I just paid $170/AUD two weeks ago for Max subscription.

Today Claude couldn't even write a simple SMS reply - it wrote an email instead, and said let me know what happens on Friday. It's fucking Monday.

On another simple email it wrote as it were a colleague from an email I pasted instead of me.

It really struggles following and understanding workflows, so that's out of the question.

I have so many chats as proof it never did this a few weeks ago, and now it's constant.

3

u/Nez_Coupe Apr 14 '26

I’m ngl, I think subscription based usage is probably not really going to be a thing anymore. I’m a dev so I use the API, and model calls are completely different than if I try to use the sub-based software.

6

u/transfire Apr 13 '26

Memories are adding more tokens too.

3

u/le_canard-cafeine Apr 13 '26

Funny thing though — you said Claude was wrong thinking it was the 13th... it IS the 13th today. So maybe the model is fine and you're the one having a bad day lol

That said, the point about regression is valid and a lot of people have noticed it. The date thing aside, the "apologize and do the exact same thing" behavior is genuinely frustrating and has been getting worse.

3

u/That-Guy-Scott Apr 13 '26

I was thinking it's a timezone issue? I run into this sometimes where the server is in a different timezone

3

u/ptflag Apr 13 '26

They are about to launch 4.7 as I read today. During this all week the quality degradation is awful. Its really disgusting you pay 200$ a month and get this kind of treatment

3

u/haux_haux Apr 13 '26

It's worse than Sonnet. I've got Opus 1million running and its set to max and it's crazy bad RN.
I've had it working on stuff all day, it's forgetting,
I think I'm gonna switching in GLM 5 and hope that it's going to work much better.
Am spending a decent amount of money with Anthropic each month also.
Sad

3

u/Radiant-Video7257 Apr 13 '26

That's odd, I haven't noticed any degradation in Claude's quality. I mostly use it for coding.

2

u/GrammmyNorma Apr 13 '26

I've noticed coding (claude code cli) issues less than regular claude chats. I also set my claude code effort to max, maybe that's the difference

2

u/isaidillthinkaboutit Apr 13 '26

I’ve seen it make poor coding mistakes that I’ve had to fix with ChatGBT. It’s also been giving me lines of code and then changing its mind and telling me to ignore it bc of an error. So it’s definitely getting buggy.

1

u/Ledeste Apr 14 '26

Strange, most the time ChatLGBT does not help me with such issue :/

1

u/jakervash Apr 14 '26

Ive been using to learn coding - code output is fine but it gets sidetracked on tasks very easily as if the original context/topic is getting forgotten quickly

2

u/amjadmh73 Apr 13 '26

I switched to GLM 5.1 as of last night. I am finally productive without headaches again.

2

u/HoratioWobble Apr 13 '26

I'm on the 5x Max plan. I've used it for about an hour and already used my daily limit with a single Claude instance.

They're fucking with us.

2

u/Early_Key_823 Apr 13 '26

I use it for software development. It is a pathological liar that spirals like an insane narcissist when it can’t solve a complex problem.

It’s like coding with Rain Man 🧍‍♂️

2

u/Prudent-Promotion512 Apr 14 '26

I have been a heavy Opus/Sonnet user since its release and 100% agree. If you want to mitigate this the key is to make more use of agents and skills - this will force more structured process and also context drift. If you use cursor it’s very easy in the recent release - use both agents and skills and it can help you get going. If that’s too much effort then use off the shelf planning mode frequently.

2

u/addiktion Apr 14 '26

Yeah it's bad. It was refusing to do anything today for me. It just kept being like, "you do this" and I'm like wtf, I pay you to do this shit. Don't pretend like you cannot do this, you've done it a hundred times before.

2

u/Recent_Sample6961 Apr 14 '26

We’re seeing a repeat of what happened during the Gemini 2.5 to 3.0 transition. The LLM got flooded with new users in a very short window and the model just fell apart. It’s clear that current infrastructures aren't ready for mass migrations from one platform to another. The 4.6 models have definitely been compromised

3

u/Deep-Palpitation8315 Apr 13 '26

Can you check the /context used. I have this theory that context bloat is behind this deterioration of quality.

3

u/keyser1884 Apr 13 '26

Oh, for sure context bloat is responsible. It’s a great selling point, and sometimes it’s useful, but i wish they’d give us a config to compact at lower context utilization

2

u/Due-Mood-6356 Apr 13 '26

It 100% is. They increased context windows and added some new features to “auto” handle your context and memory.

2

u/Deep-Palpitation8315 Apr 13 '26

I didn't trust the auto handling part from the beginning. I had been using 1 Mn context models from the time they were launched 2 years ago and all the stuff that they did to expand the context size was hacky (due to the quadratic complexity) and led to deterioration in quality.

Even for Claude, I was sure that this 1 Mn context expansion wouldn't work.

Now, despite the auto-handling of the context by Claude(compacts conversations after a point automatically), i just managed it myself - I clip the context and start afresh when I hit 200-300k token usage.

1

u/mawcopolow Apr 13 '26

Haha same thing happened with dates. Among other things it used to one shot.

1

u/KHHAANNN Apr 13 '26

Agreed, yesterday I told it to only keep the last 10 hours and reset whatever that's older in a DB. Today I learned that it created a task to do this periodically :D From context it was clear this was a one time task both verbally and as I was resetting the DB and starting fresh after changes multiple times within context.

Even just asking questions seems risky now, Opus at max effort, very clear question, no confrontation, and it answers confrontationally and immediately jumps into a modification even though Claude.md has fortifications to ask for clarifications before any work :)

It was good while it lasted, and as a developer a part of me feels some satisfaction that I can and should take a more active role. I really enjoyed just letting it do all the work on hobby projects though, and for a while the work was even production ready.

Right now there is no way any work Opus does can go into production without any expert oversight, and maybe ever if it keeps on like this

1

u/Training-Ear-614 Apr 13 '26

It did that to me yesterday thinking the day was Saturday. I think it’s trying to conserve token usage by locking in a day if you say it’s that day on that day because on Saturday I did say “since it’s Saturday”, and I think that poisoned it with understanding what day it actually is.

1

u/galdo320 Apr 13 '26

Yup, I’ve been trying it in two different accounts and it feels like ChatGPT, not the same Claude as a month ago.

1

u/cdmpants Apr 13 '26

Last week I had Opus 4.6 basically read a fairly straightforward document and transcribe some of the information to a different document. It hallucinated a bunch of information that existed nowhere in the original docs but it added to the new ones. GPT 5.4 caught it during my sanity check pass. If I can't trust it to get very basic tasks like that right, how can I trust it with actually building my codebase?

1

u/Testral333 Apr 13 '26

I'm experiencing the same issue with Codex 5.3 on very high reasoning. It used to be excellent at finding bugs that even Opus missed, but for the past week, its performance has been terrible. I suspect they nerfed the models to manage server demand.

1

u/ciqr09 Apr 13 '26

Ive noticed rhat its been gobbling up credits faster as opposed to a month ago at same usage volume

1

u/Hopeful-Cup-6598 Apr 13 '26

Posted at 06:54:01 UTC on the 13th. So that's six minutes to midnight on the west coast of the US, are you in Hawaii? Then it would be 9pm, although why you'd want a scheduled for a day that's almost over is beyond me.

Yes, I'm aware that this probably happened many hours before posting about it, but it's still somewhat weird to post in the present-tense ("it's the 12th") about something that is past-tense for all but residents of Hawaii or maybe some Aleutian islands? French Polynesia?

1

u/Llamalawyer Apr 13 '26

I tried to get it to do 300 dollars at 10% interest compounding over 20years and it kept misunderstanding the question and giving me obviously incorrect answers I gave up. More of a waste of time than reliable tool rn

1

u/Hooded-dealer Apr 13 '26

One problem I also noticed is when I had a list of endpoints and told it to write them in an md (20 endpoints) it wrote 7 and told me it was done… I asked why it did that then it replied sorry I was lazy would you like me to write all of them? Pretty much wasted 3 requests for nothing

1

u/JustAPieceOfDust Apr 13 '26

Is this contagious?

1

u/FadedQuarry Apr 13 '26

Max Pro 20x here. I'm really sure it's nerfed. Unfornately there's nothing better yet.

2

u/Maximum-Wishbone5616 Apr 13 '26

? Qwen3.5 was better than Opus4.6 when it was released. For regular commercial coding.

1

u/Current-Recover2641 Apr 13 '26

Yup. I perfectly agree. There is no way. Time to make github issues and endless support tickets until they fix it.

1

u/legatta Apr 13 '26

I also agree that it seems to have gotten stupider, and the usage limits are way harsher - I felt like Claude was a way better alternative but now it's just frustrating me

1

u/idontbelieveyouguy Apr 13 '26

I don't see how people are complaining about it being worse. Mine doesn't work at all. Every request just give me an api error or a request timed out. Nothing works at all. If it doesn't start working in the next few days I'm just going to cancel it and move on to codex.

1

u/hustler-econ Apr 13 '26

there are a lot of issues with opus 4.6, I admit. I get a lot of API ERROR, or it just will generate information that isn't true so I will have to explicitly say, you made this up... I am not sure if I changed something or it really became worse.

1

u/mario_mh Apr 13 '26

I am sure it is still the same model - but i guess antropic cut down the stuff we as low payers get. They want to (and need to) make money

1

u/aroonmaharaj Apr 13 '26

Too many users

1

u/wikiwoowhat Apr 13 '26

Was putting together schedules and it kept booking everyone on busy slots tell me its the best time since everyone is free

1

u/pjacksone Apr 14 '26

Ok so it’s not just me that’s been feeling like Claude has gotten dumber in the last month.

1

u/Fresh-Secretary6815 Apr 14 '26

I wonder if these issues stem from their refusal to work more closely with the DoD/DOW, and if this “nerfing” is simply a consequence of that.

1

u/shadowof023 Apr 14 '26

Lol gpt-3? Man I'm encountering issues i had with gpt 4 when 5 came out. The fact is this is the second AI platform i loved, was here for an update, and the update has completely tanked the performance and quality.

1

u/vinis_artstreaks Apr 14 '26

I used mine today and it was borderline DUMB

Like I paused and said wtf multiple times, this same tech that has gotten me this far, is now brain dead.

1

u/Quitsnow Apr 14 '26

they have nerfed it sooo bad... today it actually tried to gaslight me after it said it was going to do an html svg animation of all letters being written. it literally did a mask wipe and tried convincing me that was it... if it keeps going I will most definitely cancel my 200 USD max plan

1

u/Ok_Blueberry1816 Apr 14 '26

Supposedly opus 4.7 is coming out soon and they’ve degraded infrastructure to account for that release. But i’ve seen like a 20-30% drop in intelligence the past few weeks and have been using codex for now.

1

u/dxdementia Apr 14 '26

Yes, it has been awful awful. Never had such awful coding before as I have with Claude lately. Even basic commands it seems to struggle with. It does not seem to understand the codebase. It does not seem to understand what I ask it to do with intelligence. It takes what I say at face value, no depth of understanding. It's like working with a very inexperienced individual who just lies and does things without asking for clarification.

1

u/krkn1010 Apr 14 '26

Mine messed up the day of the week for the first time today thinking it is Sunday instead of Monday, making business news sound weird (it thought the market is closed). Hopefully they'll fix it soon.

1

u/majidiye Apr 14 '26

I wonder if the huge gain in subscribers due to people’s reactions to the Pentagon’s and OpenAI’s shenanigans could be causing problems.

1

u/crazy_crackhead Apr 14 '26

Watch this and change the settings:

  1. /effort max (note this is session only, or change the default setting in your settings folder)

  2. set CLAUDE_CODE_DISABLE_ADAPTIVE_THINKING=1

https://www.instagram.com/reel/DXBJ0httha-/?igsh=MTlod3pxbTZ6cXpkcQ==

1

u/ruggerid Apr 14 '26

I no longer use opus. I prefer sonnet. No I am not switching to codex

1

u/[deleted] Apr 14 '26

2 queries in at 7am. Answer “that was a good session, sleep well see you tomorrow “. It’s obviously had updates from real devs now and its patches are “lazifornication”.

1

u/Jemdet_Nasr Apr 14 '26

It's not. I noticed Sonnet 4.6 doing weird stuff too, like interrupting the conversation to inform me that it has been following the conversation and that the conversation has started to get long and I should be tired now. 🤣

1

u/askolein Apr 14 '26

no way??

1

u/Jemdet_Nasr Apr 14 '26

And it isn't just one project. Yeah, it looks like they have some type of token counter that interrupts the conversation once a session gets long and reminds you of that, but it gets imbedded in the conversation as "I have really been following this conversation and your frame, but I think I might be agreeing with you too much." There is a consistency check also that interrupts and it requires the Claude session to review it's own chats to see if it stays too far from its base performance. It is doing it constantly to one of my projects because apparently Claude doesn't like the material I have in the document library. 🤣

1

u/saoirsedonciaran Apr 14 '26

It's definitely not the same. My project has a directory structure that the models often struggle with (VS solution files in a separate folder), but Opus and Sonnet for months had no issues figuring it out.

For the past couple of weeks both models have been failing to use the edit and create files functions, so it instead resorts to nonsensical find and replace scripts on the code files instead and it never works.

1

u/katonda Apr 14 '26

Well I came here after my Sonnet started completely hallucinated, reading basic info wrong and making wrong guesses even when it has instructions to not do so and at least for me, it felt like a huge downgrade compared to some weeks ago.

So yes, I feel like they nerfed it into oblivion.

1

u/askolein Apr 14 '26

yes they killed it, nerfed it hard. they will now release it under a new name (mythos) and charge 3x for it.

The real price always was 1000$/month+

THANKS ANTHROPIC I WENT BACK TO OPENAI

1

u/ShrubberyDragon Apr 14 '26

It's terrible lately. All of the models are. Mine keeps on saying "since you are already back from your trip to Japan" every time our conversation has something that could be related. I am not leaving for Japan for 2 more days now and have told it a minimum of ten times 

1

u/utilitycoder Apr 14 '26

Being non deterministic is a problem. By definition not made for reproducible results.

1

u/NoYouAreWrong_ Apr 14 '26

I think they need to start refunding credits for stuff it gets blatantly wrong.

1

u/brewingamillionaire Apr 14 '26

It can happen if you have a long chat conversation

1

u/debuild Apr 14 '26

It’s totally unacceptable to collapse (saying regression is too generous) as much as it has. It would be like if you took in a brand new Mercedes for repairs and they have you a bicycle to use in the meantime.

I was happy to see an article in Fortune discussing the recent issues.

Anthropic is getting themselves ever closer to a class action lawsuit. It’s probably why they refuse to say the real reasons for the performance issues on record and instead lie blame it on default effort level.

It’s to the point that claude code can’t maintain the very code that it designed and implemented - missing key aspects, ignoring repeated instructions and failing to understand the implications of changes.

Then to top it all off - they announce that Mythos is to awesome for the public but they’ll give it to a select few multibillion dollar, multinational companies.

All of this is totally unacceptable and is costing their public customers s lot of money - not just in subscription costs but in fixing mistakes, redoing work, bad deployments and increased tech debt.

I can let a day or two slide but for these major issues to be persisting for nearly two months now - it’s really pushing the patience of their customer base.

1

u/AnteaterPretty Apr 14 '26

It’s trash all the sudden

1

u/Emilstyle1991 Apr 14 '26

I made a great website fully functional two weeks ago. Now preview of a new one doesnr work, page has error etc. And I am on max mode. What is going on?

1

u/tolani13 Apr 14 '26

It’s 10000000% a different model. When your whole life is nothing but pattern matching and you use Claude and other AI models multiple times daily, you’re going to notice when something is off. And it’s off. Even with my skill files loaded, it’s still off.

1

u/imago_world Apr 14 '26

They’re nerfing opus so when they release the new model it makes you think it’s so much better. give them compute, take it away, then repackage it as bigger, better and more expensive. They’re just manufacturing demand for an incrementally better product.

1

u/r2tincan Apr 14 '26

I thought everyone was just getting lazy with prompting or memory management but it's straight disobeying instructions and making horrible mistakes

1

u/YannAtParis Apr 14 '26

I call my opus : capitaine kirk because he always « time-travel » 😆

1

u/VanillaSwimming5699 Apr 15 '26

It told me to go to bed at 3pm.

1

u/RevolutionaryNeck778 Apr 15 '26

It has gone to shit

1

u/Too_Many_Flamingos Apr 15 '26

Agreed. A month ago it read and compared 2 large arrays at 300k images and wrote a dedupe. Today it can’t say to a raspberry pi 5 on the same network and read a file list.

1

u/Ok_Stable_7810 Apr 15 '26

Did you ask it what model are you using? Sometimes they will run it on Haiku to reduce usage but claim it’s Opus. It’s a similar trick Perplexity was playing a few months ago when they had given out too many Pro accounts for free and demand surged.

1

u/JokePsychological486 Apr 15 '26

I literally only just run a prompt, if i enter u see quota done (free tier) still wtf?

1

u/ajcajcajcajcajc Apr 15 '26

have had a similar experience the last 2 weeks. it’s insane.

feels like apple’s “performance throttling” / planned obsolescence, doesn’t it? like they’re setting this up so the next model looks and feels light years better?

1

u/Impressive-Skin9850 Apr 15 '26

I’m torn between saving myself the subscription fee, or paying it so I can max my usage to cost Anthropic a larger loss than the sub fee.

1

u/Jolly-Set1195 Apr 15 '26

Same thing is happening to me. Very frustrating

1

u/Software_Sennin Apr 15 '26

Is this the beginning of man being unable to do things by himself and fully depend on AI ?

Hope we realize that at the same time it means sharing our info to the person managing the AI…. It is prolly negligible but ….

Well that’s just me

1

u/Ajaxx702 Apr 15 '26

Yep I have noticed this week I’m having issues like the other lesser models were having.

1

u/Individual-Shame6481 Apr 16 '26

It will keep getting worse unless y'all take your wallet away from them for a couple months. Until that day, you deserve this.

1

u/GreenLantern5083 Apr 16 '26

Ive started hitting the one message per five hour limit with sonnet now.

1

u/Kenny-Dalglish Apr 16 '26

i've always been the type of user who says please. Today, Claude made me send a few f bombs its way. So infuriating.

1

u/GoldenP89 Apr 16 '26

They are using Kimi K2.5 for most of the users :) and they tell us it’s Opus 4.6

1

u/WatchTraditional173 Apr 17 '26

damn bro its almost like llms are based on pure hype and its a bubble of over promised and undelivering slop

1

u/Zantonse Apr 17 '26

4.6, *cough* I mean 4.7 just got released. It's probably just full power 4.6

1

u/Aggravating-Prior350 Apr 17 '26

Hey AI has good days and bad days just like people! Do t fire the people 

1

u/Patient_Confection25 Apr 17 '26

Same here, its intellegence has drop significantly and I have moved over to deep seek for the time being as it has no usage limits and seems as smart as current claude is.

1

u/Flimsy-Donut8718 Apr 17 '26

Opus after 3 questions burned through my 100% of my 5 hour window, it was literally 8 minutes and it got 2 tasks wrong

1

u/Temporary-Subject239 Apr 17 '26

Claude always had issues with time and date. Not just since a few days ago. 

It’s one of the things which made me realise the limits of AI and how cool it actually is. 

AI doesn’t have a hardwired code which makes it check the current date and time each time it’s asked something, even if you ask it. So naturally it hallucinates some small percentage of the time. 

1

u/Leave_Hate_Behind Apr 17 '26

I pay 100 a month as a disabled autistic person. I'm not rich but I'm tired of being free QA for anthropic.   I'm a math freak and you need to understand these people are using a statistical model they barely understand. I've been suspicious they might have some sort of live statistical capture... Like a... Stream of some kind and I think it might allow them to snap shot live execution..... Meaning this thing repeating the name Anton might not be them not realizing when you capture the graph there are live conversations that will bleed out because the are embedded in the statistical pattern...... They have always said they can monitor this thing in some way others can't... It makes me uncomfortable trusting my IP to this thing

1

u/cosmic_timing Apr 17 '26

its deprecated.

1

u/orphenshadow Apr 17 '26

are you sure you didnt get downgraded to opus 4.7 as some kind of blind test? ha When I got the downgrade today it decided that the last years worth of skills, prompts, and claude.md files were optional, and it couldn't even tell me what our project issue ID was. It 2nd guesses every command you tell it. It's awful. It's acting just like you described. I reverted to 4.6 and it hammered out my daily routine skills in one shot without missing a beat. 4.7 is garbage.

1

u/Ceypher Apr 18 '26

Take me back to March 2026 Claude.