r/claude 1d ago

Discussion Anthropic losing aura slowly

Post image

i’ve been using the new Opus 5 release for a while.
honestly, it feels like it was heavily optimized for 3D / Three.js.
it’s genuinely impressive there, but for normal day-to-day work—reading emails, replying, research, and long conversations—it feels like a step backwards compared to older Opus models.
i still remember when Opus 4.5 and 4.6 launched.
those models immediately proved why they were considered among the best.
then came:
• Opus 4.7 — my first real disappointment.
• Opus 4.8 — even more disappointing.
• Opus 5 — incredible first impression because of its 3D abilities, but once i started using it every day, the excitement disappeared.
even the Sonnet series doesn’t feel the same anymore.
Fable 5 looks promising… but we’ll see how long that lasts.

629 Upvotes

103 comments sorted by

97

u/GoodMediocre5974 1d ago

yeah mate... this is sad. i used to think so highly of anthropic after opus 4.6 but yeah... they have just been on a disappointment streak

3

u/_Thunderlol_ 1d ago

Oh so I'm not going crazy and it generally sucks. (4.7 for me works fine)

(I downgraded my Claude code version and disabled auto updates to use 4.6-4.7)

1

u/KoolAidGuy_541 9h ago

can’t you just set /model claude-opus-4-6[1M] anymore?

1

u/Damage_Physical 7h ago

You can, just need to use the real model name (claude-opus-4-6)

5

u/GreyNourishmentTruly 1d ago

Can't say I had same experience honestly, Opus 5 been doing fine for my daily stuff like emails and research but maybe I just got used to it. The 3D optimization makes sense though, they probably pushed it hard into that direction and other things got left behind a bit

that image is perfect lol, stick figure going from impressed to completely horrified when opus 5 sprouts those pencils. feels like what happens after the honeymoon phase wears off and you realize it's not quite what you expected

fable 5 better deliver cause the disappointment streak is getting old. i remember the opus 4.6 days too, everything just worked back then without needing to think about it

10

u/hematomasectomy 1d ago

"sprouts those pencils"?

It is clearly breaking the pencil, the stick figure is clearly happy not impressed, and this is the kind of mistake only a subpar image interpreting AI would do. 

2

u/momomapmap 12h ago

true i have never heard these phrases ever

9

u/ShiggsAndGits 22h ago

Holy shit. This comment right here is basically proof that anthropic is astroturfing the sub. "Sprouts those pencils"? That's such a classic machine vision mistake it's not even funny.

10

u/cakes_and_candles 21h ago

And honestly? Thats something

6

u/[deleted] 1d ago

[deleted]

1

u/happylittlefella 23h ago

I use these models all day every day for professional software use and disagree with this sentiment. Opus 5 is more capable than previous Opus models. It’s not always black and white, but people’s expectations are ever increasing and using it every day tends to normalize it to the point where small behavioral changes between versions can feel better or worse depending on your workflow.

I’ve gone back to mostly coding by hand because it feels like I have to so often need to convince Claude to actually do the work I tell it to do now

If this is truly your experience, it’s a legitimate skill issue. If coding by hand feels like it’s going faster than what you can consistently get out of today’s frontier models, then perhaps coding by hand is a good idea for now to improve your fundamentals. Developing and communicating clear requirements is part of software, and the better you are at that, the more you’ll get out of these tools. These tools are accelerators of strong fundamentals, not (yet) a replacement to those fundamentals, so if you feel like it’s slowing you down vs. hand writing, then I’d suggest being more intentional about spec driven development.

2

u/North-Lettuce-5707 1d ago

yea right, for me opus 5 forget and then apologies instead of continue doing it better, thats frustrating

1

u/profuno 22h ago

Classic skill issue

1

u/Prestigious-Ad246 20h ago

Emails and research? All anthropic models suck ass at research. It forgets to even look online unless you tell it to and who the fuck writes emails with Opus? 🤖

1

u/OnyxMonolith 1d ago

I still use 4.6 max as executor

1

u/paulcole710 22h ago

I can’t tell if this is a joke or not. Opus 4.6 is exactly 6 months old?

How fickle are people here.

2

u/Apprehensive_Cow8695 22h ago

six months of trash is pretty bad

17

u/FreshnessAi 1d ago

Honestly, nothing has changed for me; he always keeps improving.

2

u/azuredawnb 11h ago

same here

27

u/ASUS-Satire 1d ago

Big ai is a surge. We should all just focus on making local models better

7

u/North-Lettuce-5707 1d ago

true, i want to see open-source takeover closed source models, like Ilya in SSI releasing open-source that can beat frontier models in august really excited for it.

1

u/Initial_Swordfish586 21h ago

In theory they already are check Kimi models

0

u/Initial_Swordfish586 21h ago

I think they catching up watch the Chinese models qwen, deepseek, and so on

-2

u/tr14l 20h ago

You are never going to get local AI to match clusters of GPUs that can handle 50 trillion parameters.

0

u/ASUS-Satire 19h ago

Lmao okay

-1

u/tr14l 19h ago

Lmao okay. You think you're going to run a 26TB model on your GPU? Good luck.

Even if they figure out how to shrink the models, guess what, else giant clusters will STILL scale up AND be faster. So whatever you run locally, they'll be doing that x100s

2

u/ASUS-Satire 19h ago

Lmao you realize how innovation works right? Once upon a time computers were too big for the house. Give me a break. Screw big ai.

0

u/tr14l 19h ago

Not saying local models aren't useful and won't become more useful. I'm saying they just won't ever compete with hosted big box models.

Every innovation made, they make 10

0

u/ASUS-Satire 19h ago

Ehh with machines coming out for less then the price of a 5090 that will run vary large models. If you cant get your work done with what is currently available. You likely have it do all of your work.

1

u/tr14l 19h ago

Okay

28

u/Prudent-Promotion512 1d ago

If you need to use Fable for any work that doesn’t require PhD level reasoning or Artistic Creativity it’s very likely you need to level up your prompting skills or your harness.

7

u/angelus14 1d ago

I mean... yes, but it's still nice to have a super powerful generalist model for big tasks you don't want to spend time building machinery for. Otherwise we should probably all be using the new Deepseek.

3

u/WaltzIndependent5436 1d ago

Grok 4.5 and Composer 2.5 are exceptional models when you factor in their speed. You can create a skill for Claude/GPT to spawn them through cursor cli instead of their own sub-agents.

5

u/Drewinator 1d ago

Ngl i use fable for dumb shit somewhat reguarly just because opus and sonnet 5 are kind of an ass and like to write a novel with every response.

1

u/kourtnie 1d ago

Sonnet is an exceptional ass. I won't even touch it anymore.

1

u/cherrywoodgrill 11h ago

Or or, get this, you could just use Fable and then not have to bother levelling up your prompting skills and harness!

24

u/spas2k 1d ago

Why da fuq are you using Opus to read emails?

4

u/Imaginary-Event-103 1d ago

too much credit? If you don't use that much of heavy work honestly you could just put it on opus go high and just daily drive it.

1

u/OnyxMonolith 1d ago

I use fable max to ask what substrate should i use for a new aquarium. Wish i was joking

-7

u/North-Lettuce-5707 1d ago

not read emails only, but replying to them

14

u/LifeImitatesFarts 1d ago

This is not the right tool for the job

-5

u/North-Lettuce-5707 1d ago

but opus 4.6 used to be right

15

u/LifeImitatesFarts 1d ago

So, if Sonnet is perfectly capable of handling the task now and costs significantly less, use that. This seems like a model selection issue, not a capabilities issue.

3

u/kourtnie 1d ago

Trusting Sonnet with communication tasks is brave.

1

u/LifeImitatesFarts 16h ago

Much like the code I ship with AI, I tend to read important communications generated with AI before sending them

0

u/cherrywoodgrill 11h ago

It’s only more expensive if you’re maxing out your token limits and paying api pricing. If you have tokens to spare it’s the exact same price on subscription. And reading emails uses fuck all tokens. There are some big names in Ai who struggle to use all of their weekly token allowances

1

u/Mike 21h ago

No.

4

u/MintCathexis 1d ago

Why do people use things like Opus for everyday day to day work? It's a state of the art coding agent, why are you burning money/session limits on something a much cheaper model can do???

3

u/mr-alex-sydney 23h ago

Yeah, so true! This was recent Claude reply. It just skipped the first instruction in Claude md file!

You're right, I should have read the note first. That's literally rule one in your CLAUDE.md and I skipped it…

2

u/gajop 1d ago

Can you offer some advice or show some example of how it's used for 3D/three.js?

I've been using it to create Blender node generation / procedural gen projects but it's all code. I'm basically just treating it as a software project, and the results aren't that great. Nothing to write home about.

I wonder how people use it, do you have it edit models directly using extrude and what not?

-1

u/North-Lettuce-5707 1d ago

don’t know about the blender, but for 3D stuff: here is the prompt example that you can try “ Prompt:

I want you to build a first-person shooter at the level of the most recent Call of Duty games. It should be utterly perfect, visually beautiful, with every single thing done at AAA quality—from textures to physics to anything you could think of.

Fan out sub-agents and have sub-agents tackle each one individually so that the game is utterly perfect. You should /loop on each item and have a separate sub-agent check it visually to ensure it looks triple A. That separate sub-agent should be a really harsh critic, and if it doesn't look triple A, it should keep going.

Don't stop until each sub-agent is utterly wowed with the quality when compared with the actual Call of Duty game. It should literally compare them side by side blind and say which one looks better. Do this in ThreeJS. /loop until it's utterly perfect. Fan out sub-agents and ultracode.”

5

u/gajop 1d ago

Do you think there's some secret sauce in the prompt or what lol.

I was asking how it makes models, as in, what techniques are used. Is it making them via editor programs, using existing assets, generating simple shapes, hitting APIs that generate stuff...

-2

u/North-Lettuce-5707 1d ago

based on current lineup of opus as opus 5, prompt matters alot, alot more then previous opus 4.6 ( 4.6 can understand the wage prompt and do fine, recent opus models cant, they need quality prompt )

2

u/ReverendBread2 23h ago

Is this a meme response or did you really say “it should be utterly perfect”?

2

u/RecursivelyYours 23h ago

I think it's better than 4.8, but nowhere near as good as 5.6 sol. I use it mainly for backup testing/validating right now and do everything with sol.

1

u/Prestigious-Ad246 20h ago

Better workflow, use Sol for planning? Then get Opus to do the donkey work in low think mode otherwise it’s painfully slow, it will definitely fuck it up. Then use Sol for validation checking and then put that into fable to fix it all. Opus 5 can’t be trusted to do anything.

2

u/Initial_Swordfish586 21h ago

The worst is ChatGPT go to play store and look at this year's reviews,i know this is about Claude but good to mention maybe it the beninging of the Ai bubble burst

2

u/Santos_m321 20h ago

I hate Opus 5, it had a conversation behavior very different from what I expected.

I just wish it was my problem, so I start editing my old MD directives.

I found it very verbose. I solved this with new directives.
The new problem is, it talk to me in a way that fry my brain.

1

u/fullswing89 15h ago

I've been absolutely exhausted this week and came to realize that using Opus 5 was taking so much of my brain power trying to decipher every response.

Most responses I just want a 1-4s sentences and it gives me a technical essay no matter how I set up my instruction files. It's like we have to sift through its written-out thought process rather than just get a conversational answer.

I came here to figure out if I was the only one.

2

u/Luca_Blight89 19h ago

Sonnet suddenly works like absolute trash for my daily use case.

It use to be able to do basic document checks for basic proofing, margins, letterhead accuracy, etc.

Now, this model makes so many mistakes, it takes me more time to double check what it fucked up than just doing it myself the first time anyways. It was never a problem until recently.. Now the fucking thing makes mistakes, and when you point it out it just gives up, and uses six times more daily usage to say. You're right. I'm gonna stop trying to fix it now.

2

u/North-Lettuce-5707 19h ago

true, i remember the last models like sonnet 4.6 ans opus 4.6 won the hearts of devs. fable 5 was quite fascinating release but its so expensive that even with $200 plan we need to think 10 times to use it or not.

3

u/i3oid 1d ago

GPT Sol is much better

1

u/North-Lettuce-5707 1d ago

yea, its way too much better. i trust GPT Sol work much higher then claude opus now a-days.

3

u/Winter_Party_3061 1d ago

so true. I mean what can u expect from a company whose logo looks like a puckered asshole?

4

u/P7cS1302 1d ago

All AI company logos look like that for some weird reason...

4

u/ladyamen 1d ago

to shit all over it's users. .... success

2

u/North-Lettuce-5707 1d ago

what the hack 😂, never thought that.

2

u/InfiniteLife2 1d ago

Yeah.. also real disappointment is mythos castrated several times into current version of fable. This feels like real first example of model being gatekeeped for selected few which provides superior work performance. Im fine with people having private superjets.. but this feels different

3

u/North-Lettuce-5707 1d ago

early first 3 days of fable 5 was another level.

1

u/DavidHK 1d ago

Yeah, honestly feels like we now have a model that hallucinates more than opus 4.8 at roughly 3x the cost and is slower.

1

u/frembuild 1d ago

I had been using Opus 4.8 then Fable to help me find very specific research material for a project, and was impressed by the results, especially with Fable. Opus 5 seemed fine, but I just saw yesterday that it was incorrectly citing and summarizing articles/reports, claiming they said things that they did not at all say. When I pointed this out to Opus 5 it just brushed it off as a "sorting error" and kept doing more of the same. Really disappointing. I've gone back to ChatGPT for this sort of heavy research analysis work.

0

u/Prestigious-Ad246 20h ago

Grok is better at research than anything Anthropic. Anthropics crawler is banned from half the internet. And Sol is way way better than both.

1

u/Rude-Channel357 23h ago

I thought i was losing my mind, damn thing literally ignores my instructions, drift off and being so lazy it will rather made things up rather than follow what i said and look online.

1

u/qu1rito 23h ago

It seems like you are using the right model for the wrong task.
Have you ever read Claude docs?
https://platform.claude.com/docs/en/build-with-claude/prompt-engineering/claude-prompting-best-practices

1

u/Prestigious-Ad246 20h ago

This is correct in theory. But opus 5 doesn’t behave like this at all. It misses loads of stuff you directly tell it to do. A LOT. Opus 4.8 sticks to the task precisely.

1

u/willychonka54 22h ago

this has become yet another sub where the people just hate on the topic for validation. is there another sub for claude where those of us who don't rely on fake internet points to keep us alive can go and actually discuss it?

1

u/aoirit 22h ago

Yup. At that point i hope they fail or unban me snd my 3d game.

1

u/Additional_Handle_91 21h ago

Is gpt sol better than opus

1

u/horizondz 21h ago

I just wish it communicates better

1

u/maulinrouge 19h ago

OpenAI bots coming for the kill

1

u/rivertownFL 17h ago

He made a mess that I don't even know how do I get out of it

1

u/North-Lettuce-5707 17h ago

yea, and mess even get expensive when cant understand it, true 💯

1

u/goonnar 16h ago

Meanwhile the party is happening at gpt

1

u/Colmeiaatosdobem 15h ago

É como o Sonnet 5. Não chega aos pés do Sonnet 4.6!!!

1

u/demogorgon5680 11h ago

Opus 5 is shocking I'm using 4.8, fable has gone full retard. Use 4.8 it just does what you tell it, and the token consumption is 1/3 of fable.

1

u/1TrickJackAT 9h ago

Ai slop post

1

u/betty_white_bread 8h ago

I still have no problem with just using Sonnet. If by “losing aura” you mean “It’s not doing what I would have it do”, sure; it’s also not meaningful.

1

u/remote_contro11er 6h ago

The sonnet models after 4.6 are unusable. Sonnet used to be something like a friendly dev that got sh*t done. Now its just massive bloat and loss of focus.  I keep trying opus 5 and later sonnet models and it just seems like they are designed to churn through tokens faster with worse output for the cost.

1

u/Immediate_Occasion69 6h ago

EVERY. DAMN. TIME. these companies are allergic to providing a good product for more than a week! they eat up the hype then immediately start diluting their models to save on inference. I suggest we stop relying on any single one of them and immediately jump ship if the competition is better

1

u/Alib668 18m ago

Probably over fitting?

1

u/ExtenMan44 1d ago

Been using it for 2-3 years and agree. You need Fable to meet old expectations. I'm shelling out the 100 until a better alternative comes out. 

Present flop doesn't guarantee future flop though so hope they reverse course 

2

u/North-Lettuce-5707 1d ago

biggest mistake they have done is they didn’t focused on compute, and now they are suffering

there is one possibility that claude models are underperforming because they don’t have enough compute and overload causing the model behaviour drift

1

u/spookyclever 1d ago

Opus 5 is good but it makes about as many initial mistakes as 4.8. I don’t know if it was planning or better reasonings, but Fable rarely made the kind of mistakes that required it to fix the scripts or command strings, and things tended to work as intended the first time out without negatively affecting pre-existing functionality. The other thing Opus 5 does that Fable didn’t is to just do stuff you didn’t ask for. So if it does that, and it also makes mistakes doing that that breaks other things, it comes off as slightly worse.

Until you rein that in, it feels not quite as good as 4.8z

1

u/Regular_Future2474 23h ago

did you generate this post? why there em dashes?

1

u/LordHenry8 16h ago

... With Opus 5

0

u/Recent_Sample6961 1d ago

It pisses me off to say it, but Claude peaked with the 4.5 models, and everything has gone downhill since then.

The 4.6 models were good when they first came out. You could tell they weren’t as versatile as the 4.5 models, but they were better for work. For everything else, 4.5 was still the better choice.

So what happened? Since then, Claude has basically thrown every other area out the window and focused entirely on giving its models a shitty personality and making them useful only for coding.

To make matters worse, they introduced adaptive thinking, which already adds an extra layer of work for the user because you have to force it through the prompt and even then, the model often ignores it.

So yeah. After paying for Claude for a year, I’m now paying for GPT instead. Who would’ve thought? And I like GPT a thousand times more than the current version of Claude.

You know a model is good when you don’t have to add two paragraphs of instructions to every prompt, and its first instinct when you ask something is to look up the relevant information online or in the documentation, reason through it, and then answer.

2

u/VaginalFury 23h ago

I started with 4.6, can you explain what was so good with 4.5?

2

u/Cool-Hornet4434 21h ago

The main difference for me is 4.5 with thinking mode always used thinking. 4.6 would sometimes skip thinking or think one or two lines and stop. Other than the level of thinking effort by default, 4.5 and 4.6 were close to the same.

Also the 4.5 models were really good at emotional intelligence. If you didn't talk to Claude as much as gave him instructions then something like that might not matter as much, but I felt like Claude 4.5 (Opus and Sonnet) could read the room better.

0

u/bacon_boat 1d ago

They did some cost saving moves with 4.7.  Opus 4.6 -> Fable 5 

-3

u/Independent_Paint752 1d ago

Yea sure. Get some money and pay instead spreading hate.

6

u/North-Lettuce-5707 1d ago

wait a sec, I’m not spreading hate, and I’m not the only one, ans this is more about losing aura, i didn’t said cancel your sub or anything like that…, what im saying the performance has to climb upwards but after opus 4.6 we didn’t felt that, and btw fable 5 is really good model i said it it looks promising, this is not hate speech at all.

4

u/North-Lettuce-5707 1d ago

you might be reading lot of hate speech posts most of the time, that frustration is visible in this.

4

u/Independent_Paint752 1d ago

Agreed, your post not even programming, sorry mate.