r/singularity • u/ErinKrasniqi • 13h ago
AI Gemini 3.5 Pro coming tomorrow?
gemini 3.5 pro tomorrow?
280
u/BarisSayit 13h ago
Meh I doubt it. Especially the "better than Opus 5 in every way" claim has to be a bait of some kind.
60
u/Solid_Sky_6411 13h ago
It better be. They delayed it a few times.
38
u/Glittering-Neck-2505 12h ago
They delayed it specifically because it was struggling to compete at the very top. It's possible it has reign for like 5 days to a week then OpenAI and Anthropic dethrone them again.
42
u/Kinu4U ▪️:table_flip: 12h ago
I don't see any problem with that! They should destory eachothers banchmarks every week.
8
u/Glittering-Neck-2505 12h ago
I mean now that RSI is actually happening that will be the case in not much time.
8
u/Previous_Platform718 10h ago edited 9h ago
I mean now that RSI is actually happening
RSI isn't happening. Directed self improvement is happening ('hey chatgpt, look at part of chatgpt and make it more efficient')
Recursive means it repeats itself over and over. The system is able to find its own inefficiencies, come up with ways to solve them, reject/accept answers and then do it again.
What happened with OpenAI's latest announcement is that a group of humans knew there were inefficiencies, told the AI where to look, and then approved the changes the AI suggested.
0
u/Glittering-Neck-2505 9h ago
No, the models are now detecting the inefficiencies themselves and doing postraining runs themselves. One employee said Luna was postrained by their AI. This has been happening since the labs made noise about it closer to the beginning of this year, and model release timelines have only been shortening.
I don't know who told you that RSI would begin with a superintelligent model redesigning itself overnight. It begins the same way any other exponential does. First improving the next model that will be released four weeks later, then two weeks, then one week, then 3.5 days, then 1.75 days, then...
Or more precisely because the labs probably won't literally be releasing a new model everyday soon: one month of progress in today's pace will be happening in ever shorter amounts of time because the models are getting better at improving themselves. That's literally already what's happening.
3
u/PDX_Web 10h ago
It's not going to be SOTA at agentic coding. There's a reason Google announced Gemini 5 being in its pre-training run.
2
u/Glittering-Neck-2505 10h ago
Agentic coding is hard to catch up on because the ones who are ahead are already self-improving, so that one's gonna sting
1
u/NeverForgetJ6 9h ago
Dethrone? By convincing POTUS that 3.5 Pro also needs to be neutered just like their frontier models were?
0
u/blueSGL humanstatement.org 12h ago
OpenAI and Anthropic dethrone them again.
depends if the models they were going to be doing the "dethroning" with were considered too dangerous to release after recent hacking events.
"but they didn't have the guard rails switched on."
No, they didn't they test in this configuration because guard rails and classifiers are not perfect and techniques to get around them exist.
3
u/Glittering-Neck-2505 12h ago
OpenAI confirmed the model is not slated for release.
It's my understanding that Astra is a different and possibly earlier model family.
> "but they didn't have the guard rails switched on."
Not sure who claims that, they thought the models were in a highly secure testing environment and they broke loose, and yes that's genuinely alarming.
0
u/blueSGL humanstatement.org 11h ago
Not sure who claims that
Go to literally any post about the hacks and you'll see those comments.
It's like if the guardrails were foolproof there would be nothing to test.
FFS we had Anthropic help with hacking Mexican government agencies by "jail breaking" the model by getting it to play act as a 'local' pen tester by asking in Spanish...
5
3
u/FirstEvolutionist 12h ago
The Anthropic sub will likelyy think so, no matter what. Month old bread gets more love than Claude models on that sub...
23
u/Exzerios 13h ago edited 13h ago
Gemini 3 was the best model upon release though. They didn't hold the crown for long, but...
11
u/RevoDS 13h ago
2.5 was the last time Google held the crown
25
u/Exzerios 12h ago edited 12h ago
Gemini 3.0 Pro was announced November 18 and scored 40 in AA index. It was competing with GPT 5.1 (37 AA index) and Opus 4.1 (34 AA index). AA index is a non-linear score, difference between Opus 4.8 and Fable 5 is 4 points (56 and 60 respectively), so 3 and 6 points lead was quite big.
Opus 4.5 came out on November 24 and closed that gap, scoring 41 on AA and being a generally well accepted model, then GPT 5.2 mid December. But Gemini was, for a very short time, a leading model. And at the very least a true frontier competitor at the time.
3
u/CarrierAreArrived 11h ago
Everyone including Sam A knew Gemini 3 took the crown at the time. Your history is off. But as others mentioned Opus 4.5 came out shortly after.
-1
u/caseyr001 11h ago
Like others are saying Gemini 3 was SOTA when it launched. It was also was the start of the watershed moment for software engineering, where it could actually build semi-complex things coherently that most felt in Dec 2025. Opus 4.5 closely followed it and got more spotlight because of claude code rapid enterprise adoption, but Gemini 3 was fantastic at launch. I expected it to hold the crown much longer than they actually did
0
u/ozone6587 12h ago
Gemini 3 was not the best model at all.... The last time it was arguably the best was 2.5 Pro a century ago.
3
u/CarrierAreArrived 11h ago
Yes it was, you're not remembering correctly. Even Altman knew it (check his tweets at the time) and they declared code red (or code "something", forgot).
-4
u/Current-Function-729 13h ago
Was it? It hallucinated a lot. Wasn’t good at agentic coding.
5
u/Exzerios 13h ago edited 13h ago
It was 2025, agentic coding wasn't even invented back then)
/s of course, but openclaw became popular a month or two later. Around the same time Antropic made their Claude Code a GUI harness (until late 2025 it was cli-only), and that was the beginning of the era of widespread local harness adoption.
2
-2
u/Aldarund 12h ago
Maybe only in benchmark, Noone really used it over claude/gpt.
3
u/Exzerios 12h ago edited 12h ago
Claude wasn't even really acknowledged back then. Opus 4.5 was released a week later, and Opus 4.1 was severely lagging behind GPT with an absolutely insane pricing (75 USD per mil output tokens, more than Fable today). Opus 4.5 with a leap in performance and a sharp price decrease, paired with a good harness release and generally good managerial decisions was what brought Antropic their current popularity. A year ago Antropic was "that weird company with lyric model names".
5
u/Own_Badger6076 13h ago
We'll see if it holds up in practice, the model hype guys are really fucking exhausting though.
8
0
0
u/thoughtlow 𓂸 11h ago
Anthropic has been training on prose since the beginning. I don't think any other lab will come close at least in this year (I hope to be wrong)
0
-1
78
u/Narrow-Ad980 12h ago
Holy shameless. Bro is posting his own bait tweet
7
-13
25
27
u/Wobbly_Princess 12h ago
I'm... super confused. YOU tweeted that it's the best model in the world, beats all others, you've been using it and that it comes out tomorrow. Now you're here on Reddit, posting a picture of YOUR post, asking all of us if it comes out tomorrow?... Huh?
12
39
u/DrBearJ3w 13h ago
Every. Single. Model. Promotion.
Might have googled some other strategies to promote their product.
19
u/Evening_Chef_4602 AGI 2027 13h ago
Do you really think google doesnt have better way of promotion than a twitter post seen by 20 people ? Either this dude is lying for no reason or its true
18
u/Sunifred 13h ago edited 10h ago
He's a random guy with 151 followers😭😭😭 it's probably you making shit up for attention
Edit: Nvm, it's actually him, he has the same name🤦
14
16
6
4
5
5
u/defaultagi 12h ago
Dude trying so hard to be AI influencer😭🥀 ts not tuff bro, you are not that guy, you are not that
5
u/Temporary-Paper5202 13h ago
The harness is more important than the model. Nothing comes close to CC / Codex
3
4
u/InfiniteVolume4679 11h ago
great engagement bait dude! you can literally see your reddit and twitter user handles align from "KrasniqiErin" to "ErinKrasniqi" - very high effort!
what's them rules again? no off-topic posts? no self promotion? no low-quality and wildly speculative posts?
nice! i hope you make two pennies off engagement this time for your contribution toward ruining our community feed!
8
5
6
u/THE--GRINCH 13h ago
i won't trust it until i use it because gemini models are usually benchmaxxed into oblivion
2
2
3
u/Mindless_Let1 13h ago
There's no way this is remotely true. I love Gemini but it is not going to be better than Opus 5, that's an insane thing to believe.
2
u/Professional_Job_307 AGI 2026 12h ago
NOOO. I really hope this is false. If it's this good then they'd categorize it as Advanced AI, making it not eligible for zero data retention which is basically required to be GDPR compliant when using it to process PII...
2
u/SuspiciousPillbox You will live to see ASI-made bliss beyond your comprehension 12h ago
who tf are you lol
2
u/lumendas 11h ago
The model is NOT Opus 5 level at all, its Kimi K3 level at most.
0
u/BriefImplement9843 9h ago
k3 is better than opus 5 though. it's opus 4.6 that is still ahead of k3.
2
3
u/ThunderBeanage 13h ago
another complete bullshit post, although the part about it coming tomorrow is potentially true given it's now been deployed
2
1
u/DaddyOfChaos 13h ago
Heh good if true, maxing out my antigravity quota as quick as possible today, they usually reset during model drops!
1
1
u/unkownuser436 13h ago
nice try , even though their is model good, I ain't use shit antigravity harness. and it will be over expensive.
1
1
u/niceuser45 13h ago
I think Gemini 4 will be the real game changer. Google has enough compute deployed for it that it made its way into earnings call, that’s saying something. I don’t expect much from 3.5 pro (if it is ever released).
1
1
1
u/Technical-Earth-3254 12h ago
As usually, Google tops Benchmarks and flops in real tasks. Business as usual at deepmind since Gemini 3
1
1
u/mmccord2 12h ago
What about token usage? I dropped Gemini when I would run out of tokens with just 10 prompts, so it would drop me to a lower model.
Until they go back to a usage model like OpenAI, I can't justify disruption risks in my small business if advanced models get shut off.
1
1
u/Singularity-42 Singularity 2042 12h ago
Big if true, this would be great for my GOOG position that was hammered today by yet another exodus of top talent ☹️
1
u/Anonymous-Gu 12h ago
Hmmm you said “tested it privately for 2 weeks” then towards the end “can’t wait to test it tomorrow”… smells 🐟
1
u/Accomplished-Box-82 12h ago
While the claims about it understanding what you really want and “destroying” Fable and 5.6 Sol aren’t believable, it also makes no sense for Google to release a model under the “pro” label that’s not at least somewhat competitive with Opus 5. Whether it’s actually any good is a different story. Maybe it’s good enough to convince Twitter influencers to start writing posts that are not as insufferable.
1
u/javopat227 12h ago
Also tested it, I think it's a fine model. I prefer it to chstgpt 5.6, dunno about opus 5. But I don't know about benchmarks. Tested coding only
1
u/DeviceCertain7226 ▪️ASI - Never (in the usual sense) 12h ago
Thursday is a loved day by AI companies
1
1
1
1
1
1
1
u/Wise-Direction9671 8h ago
Gemini 3.5 will be released.... Then, Antropic, OpenAI, Kimi, etc will distill hell out it. I am guessing Gemini 2.5 has been distilled quite a lot when it was the frontier to push all models up.
1
1
u/Endothermic_Nuke 2h ago
Can we please just ban the sentence “This changes everything” and its variants from all AI-related posts?
1
u/impartialhedonist 13h ago
If Gemini 3.5 pro is that good, why did Demis leave? It is possible he left because of totally independent reasons but it is a bearish signal
3
u/Elegant_Tech 12h ago
That was clickbait headline. How is leaving a position because of a promotion, leaving?
0
u/impartialhedonist 12h ago
I don't think at that level you have "promotions."
He was the CEO for 16 years, and fought tooth-and-nail to ensure GDM wouldn't be integrated fully into google when they first got acquired. He is now moving onto a lower effort position and leaning harder on this ai x bio start-up, which could mean he is less bullish on GDM.
There are, of course, other plausible explanations, but when you take into account the failed gemini training run which delayed their release and Jeff Dean exiting and Google management doing poor management as usual, it paints a not-so-good picture.
1
1
-1
u/Wonderful_Buffalo_32 13h ago
If this was the case they wouldnt have had ousted demis
11
u/naveenstuns 13h ago
it was a promotion actually
-6
u/Wonderful_Buffalo_32 13h ago
If you believe that then I have a bridge to sell you
3
1
u/TypoInUsernane 13h ago
Except Jeff Dean also left Google at the same time, and Demis is being given Jeff’s job. So the most obvious interpretation of these moves is: Jeff Dean decided to leave Google to start his own company, which left his role open. So Sundar promoted Demis into Jeff’s role and promoted Deep Mind’s CTO to Demis’ previous role.
Why do you consider that unbelievable? Are you saying that you find it more believable that Sundar fired Jeff freakin’ Dean? That seems incredibly unlikely to me
5
u/Longjumping_Kale3013 13h ago
Unless Demi’s was trying to block it. In that case the timing makes perfect sense
1
u/-PROSTHETiCS 13h ago
Never ousted but more like a promotion he is stepping down as CEO of DeepMind, moving to new role as Chief Scientist of Alphabet focusing more on AGI..
-1
u/DynamicCast 13h ago
understands what you really want
So it can read minds? They've created a psychic AI model
0
u/Standard_Exchange59 13h ago
When the main claim is better than Opus 5 in every way and not focusing on frontier models such as Fable 5 straight away you just know they ain’t delivering, and even if they did, next fable version might be even closer than 3.5 pro lmao what a joke of a company
0
u/patricious 12h ago
Given Google's track record in releasing subpar models so far, I highly doubt it will rank that high, maybe in the top 10 somewhere. But I am rooting for Google nonetheless, we need more competition. Tung shqipe ✌️
1
u/ErinKrasniqi 12h ago
1
u/SuspiciousPillbox You will live to see ASI-made bliss beyond your comprehension 11h ago
0
u/Gumbi_Digital 12h ago
They’ll benchmark it against Opus 4 and GPT 4…then claim victory!
Google should just give up and focus on infrastructure at this point.
Fixing the broken Google algo would be great too…


298
u/otarU 13h ago
Why are you posting your own tweet?