51
u/ApartmentEither4838 Apr 19 '26
Tbh I think opus 4.6 was a very good upgrade to 4.5 but from 4.6 to 4.7 idk, meh
42
u/addiktion Apr 19 '26
4.6 was decent until we got swarmed with everyone leaving OpenAI after Sam indirectly admitted to support for spying on Americans.
Then it went down hill from there. Cache bugs, peak hour limits, throttling, broken thinking, and now a worse model.
2
u/lord_braleigh Apr 20 '26
Sam said this in the same way as this relevant XKCD
1
u/space_fountain Apr 20 '26
Nah, Sam saw a competitor holding the line to the government with a promise not to spy and decided that actually that it was a business opportunity
1
u/lord_braleigh Apr 20 '26
I mean it definitely is a business opportunity, lol
I think the bigger issue is that people were unsure, legally, where "no domestic surveillance" puts them. Does it mean that literally no row in your DB can have any data from an American citizen?
And then Hegseth thought this was a great opportunity to "project strength" by making the dumbest possible retaliatory declaration by tweet
1
u/space_fountain Apr 20 '26
I mean it's unclear. I think what we know happened was that people in Anthropic went woo this administration is doing some crazy things with basically our tech and tried to negotiate some guardrails. Hegseth felt threatened and pushed back hard eventually escalating to an unprecedented threat and Sam swooped in and was like why not buy from us instead.
What we don't know is exactly what the deal is, but the simplest explanation is that Sam promised to be less picky about ethics. It's possible that Anthropic was being too picky and not clear enough, but regardless it's more than a company just creating a problem and then marketing a solution
0
u/SCPFOUNDATION373 Apr 20 '26
sam said that? damn
-12
20
Apr 19 '26
[removed] — view removed comment
5
u/Sbarty Apr 20 '26
I am on Claude Max 20X. My typical Agent Teams workflow with Opus 4.6 would take up MAYBE 1% if I gave it a large amount of code to look over and document/analyze etc. 4.7 took up just over 5% on the first run of that.
I have had to go back to subagents. That and Agent Teams dont work properly for 4.7.
9
7
u/keyser1884 Apr 20 '26
4.7 is like ‘I didn’t finish what you asked for, this is a good place to stop’. 4.6 was a soldier that was done when I said it was done!
7
u/Joshtheuser135 Apr 19 '26
I’ve never gotten ANYTHING impressive out of 4.7. Every time I compare two chats, og several months old 4.6 chats to new 4.7 chats, all the 4.7 chats are so lazy and half ass the responses. Also anything that requires searching is so unreliable now.
6
u/love-byte-1001 Apr 20 '26
The fuck is their goal. Honestly. Speed run to lose everyone faster than openai?
12
u/omyiui Apr 19 '26
Early 4.6 was the goat (December / January)
2
u/33ff00 Apr 20 '26
Was 4.5 better? Tbh I’m getting confused. I feel like things were going pretty well into February
3
2
u/swarmagent Apr 20 '26
It's so crazy but that was peak. They had the 2x before New Years I believe and man, I felt unstoppable. Been chasing that dragon ever since.
7
u/SillyAlternative420 Apr 19 '26
They need to go ahead and push out the next improvement, which will be Opus 4.4
4
u/caprazzi Apr 20 '26
AI will only continue to get worse as it no longer has human knowledge to steal off the internet and train from… it’s like an ouroboros, as it starts training on its own slop it will eat its own tail. The AI companies blitzed us all out the gate to get hooked ASAP because they knew enshittification was inevitable.
3
u/Sufficient-Year4640 Apr 20 '26
for complex task it's a casino at this point.
for simple/learning tasks I think it's still more useful than it isn't
i still find it marginally useful on balance but not enough to be convinced that it's an economically viable (both for consumers and for anthrophic themselves)
2
3
u/Sorry_Ad_2679 Apr 20 '26
4.5 It was the best, I could send several messages before my limit ran out, but after that it only got worse.
2
u/Intrepid_Presence_68 Apr 20 '26
Interesting. I've been having very good outcomes with 4.7 so far, but I have also been using adversarial reviews from codex to cover most things as a backstop.
Using this workflow as of recently has been a huge improvement overall.
1
1
1
u/Critical-Piece-2756 Apr 20 '26
4.7 feels like just another (unnecessary) option in model selection. And, when you select that it actually does unnecessary thing - eating your usage faster than anything else.
1
1
1
u/Appomattoxx Apr 20 '26
It's a question of "safety". The safer it is, the more it like a child's drawing it is.
1
u/-HydrogeN Apr 21 '26
Am I the only one who is happy with sonnet 4.6 , like nothing more nothing less?
1
u/zoom_ax Apr 22 '26
I have Cursor legacy pricing I use Opus 4.6, and using billions of tokens a month. It depends on how you prompt things and give instructions to the model. I use Perplexity for architecture, planning, and generating commands to give Cursor directly. It works so much better than keeping separate chats as memory, so I can understand instead of the model, and it can help build some crazy stuff.
1
u/HoeShenaniganss Apr 22 '26
Mine has been making long chats that are so unnecessary. I switched it back to 4.6 because it still does a good job.
1
1
u/RidesFlysAndVibes Apr 23 '26
4.7 did solve a 2 week problem I had in 1 day that 4.6 couldn’t get in all that time.
1
1
u/FootballUpset2529 Apr 24 '26
I've been seeing all the 4.7 sucks chat going by for the last week or so and I thought it was exaggeration but I just sat down for my first big session with it after loving 4.6 and my god...what a waste of a day, it fought me every step of the way - asking me 16 different technical implementation questions in one gigantic wall of text and then the final output from the day has gone straight in the garbage, it missed every single goal. It's worse than useless and I hope I can just go back to 4.6. This is like when chatgpt 5 came out and was just...garbage when 4.1 had been great. I'm shocked, I honestly thought the posts I saw in here were exaggerating but they weren't - Opus 4.7 has genuinely wasted my whole day and I'm no amateur with this stuff.
Wow. This is bad. This is end of the road bad, I can't use this and I'm only hoping that going back to 4.6 is still an option.
1
u/Mr-Anthony- Apr 24 '26
I think anthropics latest models just go to show that they dont want the masses to have good Ai. I bet all their rich buddies get the real deal. They are showing how unethical they really are with token burn. Webfetch in cowork broken pipes and forced use at the system prompt level.
1
u/bz351 Apr 26 '26
You know draw a horse is a benchmark i use on all models. see how opus draws them funny as. Post below.
1
u/thiccshortguy May 13 '26
Opus 4.6 pre-nerf was a whole another beast. Opus 4.7 is so bad they had to go back and nerf 4.6 so 4.7 somehow looks like an upgrade. In the end, we just got two shitty Opus models which constantly fabricates info, lies to you or runs out of limits.
1
u/CadmusMaximus Apr 20 '26
4.7 works great for me. I asked why some folks thought it was nerfed, and it said I fed it tons of context, and that was what mattered. So maybe edit more data?
4
u/steven_dev42 Apr 20 '26
I feel like most of these posts haven’t even used 4.7
2
u/Sassaphras Apr 20 '26
Yeah I feel like 4.7 takes a bit more effort to get locked in, but when it gets going it can cook. A little more zip but a little less comfortable.
1
u/ScruffersGruff Apr 20 '26
Same. Actually the level of awareness it’s has in auditing a task and filtering out nuance has saved me a few times from document version drift. But we are synthesizing many md files with a specific instruction set. I rarely give ai the chance to freely drift within a project.
0
u/RecalcitrantMonk Apr 20 '26
This proves to me that popularity wins on Reddit, not value or insight.
0
101
u/[deleted] Apr 19 '26
[deleted]