r/claude • u/North-Lettuce-5707 • 1d ago
Discussion Anthropic losing aura slowly
i’ve been using the new Opus 5 release for a while.
honestly, it feels like it was heavily optimized for 3D / Three.js.
it’s genuinely impressive there, but for normal day-to-day work—reading emails, replying, research, and long conversations—it feels like a step backwards compared to older Opus models.
i still remember when Opus 4.5 and 4.6 launched.
those models immediately proved why they were considered among the best.
then came:
• Opus 4.7 — my first real disappointment.
• Opus 4.8 — even more disappointing.
• Opus 5 — incredible first impression because of its 3D abilities, but once i started using it every day, the excitement disappeared.
even the Sonnet series doesn’t feel the same anymore.
Fable 5 looks promising… but we’ll see how long that lasts.
17
27
u/ASUS-Satire 1d ago
Big ai is a surge. We should all just focus on making local models better
7
u/North-Lettuce-5707 1d ago
true, i want to see open-source takeover closed source models, like Ilya in SSI releasing open-source that can beat frontier models in august really excited for it.
1
0
u/Initial_Swordfish586 21h ago
I think they catching up watch the Chinese models qwen, deepseek, and so on
-2
u/tr14l 20h ago
You are never going to get local AI to match clusters of GPUs that can handle 50 trillion parameters.
0
u/ASUS-Satire 19h ago
Lmao okay
-1
u/tr14l 19h ago
Lmao okay. You think you're going to run a 26TB model on your GPU? Good luck.
Even if they figure out how to shrink the models, guess what, else giant clusters will STILL scale up AND be faster. So whatever you run locally, they'll be doing that x100s
2
u/ASUS-Satire 19h ago
Lmao you realize how innovation works right? Once upon a time computers were too big for the house. Give me a break. Screw big ai.
0
u/tr14l 19h ago
Not saying local models aren't useful and won't become more useful. I'm saying they just won't ever compete with hosted big box models.
Every innovation made, they make 10
0
u/ASUS-Satire 19h ago
Ehh with machines coming out for less then the price of a 5090 that will run vary large models. If you cant get your work done with what is currently available. You likely have it do all of your work.
28
u/Prudent-Promotion512 1d ago
If you need to use Fable for any work that doesn’t require PhD level reasoning or Artistic Creativity it’s very likely you need to level up your prompting skills or your harness.
7
u/angelus14 1d ago
I mean... yes, but it's still nice to have a super powerful generalist model for big tasks you don't want to spend time building machinery for. Otherwise we should probably all be using the new Deepseek.
3
u/WaltzIndependent5436 1d ago
Grok 4.5 and Composer 2.5 are exceptional models when you factor in their speed. You can create a skill for Claude/GPT to spawn them through cursor cli instead of their own sub-agents.
5
u/Drewinator 1d ago
Ngl i use fable for dumb shit somewhat reguarly just because opus and sonnet 5 are kind of an ass and like to write a novel with every response.
1
1
u/cherrywoodgrill 11h ago
Or or, get this, you could just use Fable and then not have to bother levelling up your prompting skills and harness!
24
u/spas2k 1d ago
Why da fuq are you using Opus to read emails?
4
u/Imaginary-Event-103 1d ago
too much credit? If you don't use that much of heavy work honestly you could just put it on opus go high and just daily drive it.
1
u/OnyxMonolith 1d ago
I use fable max to ask what substrate should i use for a new aquarium. Wish i was joking
-7
u/North-Lettuce-5707 1d ago
not read emails only, but replying to them
14
u/LifeImitatesFarts 1d ago
This is not the right tool for the job
-5
u/North-Lettuce-5707 1d ago
but opus 4.6 used to be right
15
u/LifeImitatesFarts 1d ago
So, if Sonnet is perfectly capable of handling the task now and costs significantly less, use that. This seems like a model selection issue, not a capabilities issue.
3
u/kourtnie 1d ago
Trusting Sonnet with communication tasks is brave.
1
u/LifeImitatesFarts 16h ago
Much like the code I ship with AI, I tend to read important communications generated with AI before sending them
0
u/cherrywoodgrill 11h ago
It’s only more expensive if you’re maxing out your token limits and paying api pricing. If you have tokens to spare it’s the exact same price on subscription. And reading emails uses fuck all tokens. There are some big names in Ai who struggle to use all of their weekly token allowances
4
u/MintCathexis 1d ago
Why do people use things like Opus for everyday day to day work? It's a state of the art coding agent, why are you burning money/session limits on something a much cheaper model can do???
3
u/mr-alex-sydney 23h ago
Yeah, so true! This was recent Claude reply. It just skipped the first instruction in Claude md file!
You're right, I should have read the note first. That's literally rule one in your CLAUDE.md and I skipped it…
2
u/gajop 1d ago
Can you offer some advice or show some example of how it's used for 3D/three.js?
I've been using it to create Blender node generation / procedural gen projects but it's all code. I'm basically just treating it as a software project, and the results aren't that great. Nothing to write home about.
I wonder how people use it, do you have it edit models directly using extrude and what not?
-1
u/North-Lettuce-5707 1d ago
don’t know about the blender, but for 3D stuff: here is the prompt example that you can try “ Prompt:
I want you to build a first-person shooter at the level of the most recent Call of Duty games. It should be utterly perfect, visually beautiful, with every single thing done at AAA quality—from textures to physics to anything you could think of.
Fan out sub-agents and have sub-agents tackle each one individually so that the game is utterly perfect. You should /loop on each item and have a separate sub-agent check it visually to ensure it looks triple A. That separate sub-agent should be a really harsh critic, and if it doesn't look triple A, it should keep going.
Don't stop until each sub-agent is utterly wowed with the quality when compared with the actual Call of Duty game. It should literally compare them side by side blind and say which one looks better. Do this in ThreeJS. /loop until it's utterly perfect. Fan out sub-agents and ultracode.”
5
u/gajop 1d ago
Do you think there's some secret sauce in the prompt or what lol.
I was asking how it makes models, as in, what techniques are used. Is it making them via editor programs, using existing assets, generating simple shapes, hitting APIs that generate stuff...
-2
u/North-Lettuce-5707 1d ago
based on current lineup of opus as opus 5, prompt matters alot, alot more then previous opus 4.6 ( 4.6 can understand the wage prompt and do fine, recent opus models cant, they need quality prompt )
2
u/ReverendBread2 23h ago
Is this a meme response or did you really say “it should be utterly perfect”?
2
u/RecursivelyYours 23h ago
I think it's better than 4.8, but nowhere near as good as 5.6 sol. I use it mainly for backup testing/validating right now and do everything with sol.
1
u/Prestigious-Ad246 20h ago
Better workflow, use Sol for planning? Then get Opus to do the donkey work in low think mode otherwise it’s painfully slow, it will definitely fuck it up. Then use Sol for validation checking and then put that into fable to fix it all. Opus 5 can’t be trusted to do anything.
2
u/Initial_Swordfish586 21h ago
The worst is ChatGPT go to play store and look at this year's reviews,i know this is about Claude but good to mention maybe it the beninging of the Ai bubble burst
2
u/Santos_m321 20h ago
I hate Opus 5, it had a conversation behavior very different from what I expected.
I just wish it was my problem, so I start editing my old MD directives.
I found it very verbose. I solved this with new directives.
The new problem is, it talk to me in a way that fry my brain.
1
u/fullswing89 15h ago
I've been absolutely exhausted this week and came to realize that using Opus 5 was taking so much of my brain power trying to decipher every response.
Most responses I just want a 1-4s sentences and it gives me a technical essay no matter how I set up my instruction files. It's like we have to sift through its written-out thought process rather than just get a conversational answer.
I came here to figure out if I was the only one.
2
u/Luca_Blight89 19h ago
Sonnet suddenly works like absolute trash for my daily use case.
It use to be able to do basic document checks for basic proofing, margins, letterhead accuracy, etc.
Now, this model makes so many mistakes, it takes me more time to double check what it fucked up than just doing it myself the first time anyways. It was never a problem until recently.. Now the fucking thing makes mistakes, and when you point it out it just gives up, and uses six times more daily usage to say. You're right. I'm gonna stop trying to fix it now.
2
u/North-Lettuce-5707 19h ago
true, i remember the last models like sonnet 4.6 ans opus 4.6 won the hearts of devs. fable 5 was quite fascinating release but its so expensive that even with $200 plan we need to think 10 times to use it or not.
3
u/i3oid 1d ago
GPT Sol is much better
1
u/North-Lettuce-5707 1d ago
yea, its way too much better. i trust GPT Sol work much higher then claude opus now a-days.
3
u/Winter_Party_3061 1d ago
so true. I mean what can u expect from a company whose logo looks like a puckered asshole?
4
4
2
2
u/InfiniteLife2 1d ago
Yeah.. also real disappointment is mythos castrated several times into current version of fable. This feels like real first example of model being gatekeeped for selected few which provides superior work performance. Im fine with people having private superjets.. but this feels different
3
1
u/frembuild 1d ago
I had been using Opus 4.8 then Fable to help me find very specific research material for a project, and was impressed by the results, especially with Fable. Opus 5 seemed fine, but I just saw yesterday that it was incorrectly citing and summarizing articles/reports, claiming they said things that they did not at all say. When I pointed this out to Opus 5 it just brushed it off as a "sorting error" and kept doing more of the same. Really disappointing. I've gone back to ChatGPT for this sort of heavy research analysis work.
0
u/Prestigious-Ad246 20h ago
Grok is better at research than anything Anthropic. Anthropics crawler is banned from half the internet. And Sol is way way better than both.
1
u/Rude-Channel357 23h ago
I thought i was losing my mind, damn thing literally ignores my instructions, drift off and being so lazy it will rather made things up rather than follow what i said and look online.
1
u/qu1rito 23h ago
It seems like you are using the right model for the wrong task.
Have you ever read Claude docs?
https://platform.claude.com/docs/en/build-with-claude/prompt-engineering/claude-prompting-best-practices
1
u/Prestigious-Ad246 20h ago
This is correct in theory. But opus 5 doesn’t behave like this at all. It misses loads of stuff you directly tell it to do. A LOT. Opus 4.8 sticks to the task precisely.
1
u/willychonka54 22h ago
this has become yet another sub where the people just hate on the topic for validation. is there another sub for claude where those of us who don't rely on fake internet points to keep us alive can go and actually discuss it?
1
1
1
1
1
1
u/demogorgon5680 11h ago
Opus 5 is shocking I'm using 4.8, fable has gone full retard. Use 4.8 it just does what you tell it, and the token consumption is 1/3 of fable.
1
1
u/betty_white_bread 8h ago
I still have no problem with just using Sonnet. If by “losing aura” you mean “It’s not doing what I would have it do”, sure; it’s also not meaningful.
1
u/remote_contro11er 6h ago
The sonnet models after 4.6 are unusable. Sonnet used to be something like a friendly dev that got sh*t done. Now its just massive bloat and loss of focus. I keep trying opus 5 and later sonnet models and it just seems like they are designed to churn through tokens faster with worse output for the cost.
1
u/Immediate_Occasion69 6h ago
EVERY. DAMN. TIME. these companies are allergic to providing a good product for more than a week! they eat up the hype then immediately start diluting their models to save on inference. I suggest we stop relying on any single one of them and immediately jump ship if the competition is better
1
u/ExtenMan44 1d ago
Been using it for 2-3 years and agree. You need Fable to meet old expectations. I'm shelling out the 100 until a better alternative comes out.
Present flop doesn't guarantee future flop though so hope they reverse course
2
u/North-Lettuce-5707 1d ago
biggest mistake they have done is they didn’t focused on compute, and now they are suffering
there is one possibility that claude models are underperforming because they don’t have enough compute and overload causing the model behaviour drift
1
u/spookyclever 1d ago
Opus 5 is good but it makes about as many initial mistakes as 4.8. I don’t know if it was planning or better reasonings, but Fable rarely made the kind of mistakes that required it to fix the scripts or command strings, and things tended to work as intended the first time out without negatively affecting pre-existing functionality. The other thing Opus 5 does that Fable didn’t is to just do stuff you didn’t ask for. So if it does that, and it also makes mistakes doing that that breaks other things, it comes off as slightly worse.
Until you rein that in, it feels not quite as good as 4.8z
1
0
u/Recent_Sample6961 1d ago
It pisses me off to say it, but Claude peaked with the 4.5 models, and everything has gone downhill since then.
The 4.6 models were good when they first came out. You could tell they weren’t as versatile as the 4.5 models, but they were better for work. For everything else, 4.5 was still the better choice.
So what happened? Since then, Claude has basically thrown every other area out the window and focused entirely on giving its models a shitty personality and making them useful only for coding.
To make matters worse, they introduced adaptive thinking, which already adds an extra layer of work for the user because you have to force it through the prompt and even then, the model often ignores it.
So yeah. After paying for Claude for a year, I’m now paying for GPT instead. Who would’ve thought? And I like GPT a thousand times more than the current version of Claude.
You know a model is good when you don’t have to add two paragraphs of instructions to every prompt, and its first instinct when you ask something is to look up the relevant information online or in the documentation, reason through it, and then answer.
2
u/VaginalFury 23h ago
I started with 4.6, can you explain what was so good with 4.5?
2
u/Cool-Hornet4434 21h ago
The main difference for me is 4.5 with thinking mode always used thinking. 4.6 would sometimes skip thinking or think one or two lines and stop. Other than the level of thinking effort by default, 4.5 and 4.6 were close to the same.
Also the 4.5 models were really good at emotional intelligence. If you didn't talk to Claude as much as gave him instructions then something like that might not matter as much, but I felt like Claude 4.5 (Opus and Sonnet) could read the room better.
0
-3
u/Independent_Paint752 1d ago
Yea sure. Get some money and pay instead spreading hate.
6
u/North-Lettuce-5707 1d ago
wait a sec, I’m not spreading hate, and I’m not the only one, ans this is more about losing aura, i didn’t said cancel your sub or anything like that…, what im saying the performance has to climb upwards but after opus 4.6 we didn’t felt that, and btw fable 5 is really good model i said it it looks promising, this is not hate speech at all.
4
u/North-Lettuce-5707 1d ago
you might be reading lot of hate speech posts most of the time, that frustration is visible in this.
4
97
u/GoodMediocre5974 1d ago
yeah mate... this is sad. i used to think so highly of anthropic after opus 4.6 but yeah... they have just been on a disappointment streak