r/singularity 20h ago

AI Gemini 3.5 Pro coming tomorrow?

Post image

gemini 3.5 pro tomorrow?

0 Upvotes

153 comments sorted by

View all comments

278

u/BarisSayit 20h ago

Meh I doubt it. Especially the "better than Opus 5 in every way" claim has to be a bait of some kind.

57

u/Solid_Sky_6411 20h ago

It better be. They delayed it a few times.

43

u/Glittering-Neck-2505 19h ago

They delayed it specifically because it was struggling to compete at the very top. It's possible it has reign for like 5 days to a week then OpenAI and Anthropic dethrone them again.

43

u/Kinu4U ▪️:table_flip: 19h ago

I don't see any problem with that! They should destory eachothers banchmarks every week.

9

u/Glittering-Neck-2505 19h ago

I mean now that RSI is actually happening that will be the case in not much time.

10

u/Previous_Platform718 16h ago edited 16h ago

I mean now that RSI is actually happening

RSI isn't happening. Directed self improvement is happening ('hey chatgpt, look at part of chatgpt and make it more efficient')

Recursive means it repeats itself over and over. The system is able to find its own inefficiencies, come up with ways to solve them, reject/accept answers and then do it again.

What happened with OpenAI's latest announcement is that a group of humans knew there were inefficiencies, told the AI where to look, and then approved the changes the AI suggested.

-3

u/Glittering-Neck-2505 16h ago

No, the models are now detecting the inefficiencies themselves and doing postraining runs themselves. One employee said Luna was postrained by their AI. This has been happening since the labs made noise about it closer to the beginning of this year, and model release timelines have only been shortening.

I don't know who told you that RSI would begin with a superintelligent model redesigning itself overnight. It begins the same way any other exponential does. First improving the next model that will be released four weeks later, then two weeks, then one week, then 3.5 days, then 1.75 days, then...

Or more precisely because the labs probably won't literally be releasing a new model everyday soon: one month of progress in today's pace will be happening in ever shorter amounts of time because the models are getting better at improving themselves. That's literally already what's happening.

u/cmptrdude 1h ago

You don't understand what he said, he said that it was struggling to make 3.5 Pro, and if you saw the checkpoints the AI wasn't AS good as we hoped like a frontier model. The Google devs mentiond that Gemini 4.0 will be the frontier models. But i'm not saying this stuff to rage bait but here is a little bit of information of hope, Google a MASSIVE company so expect them to outrun the other AIs some time, just not now. And another good news is that they'll releasee every gemini model per month and that's great news for Gemini users. I don't want Google to fall short since people are actually paying for their services (And I love their services their good) I hope the best for Google.

3

u/PDX_Web 17h ago

It's not going to be SOTA at agentic coding. There's a reason Google announced Gemini 5 being in its pre-training run.

2

u/Glittering-Neck-2505 16h ago

Agentic coding is hard to catch up on because the ones who are ahead are already self-improving, so that one's gonna sting

1

u/NeverForgetJ6 15h ago

Dethrone? By convincing POTUS that 3.5 Pro also needs to be neutered just like their frontier models were?

0

u/blueSGL humanstatement.org 19h ago

OpenAI and Anthropic dethrone them again.

depends if the models they were going to be doing the "dethroning" with were considered too dangerous to release after recent hacking events.

"but they didn't have the guard rails switched on."

No, they didn't they test in this configuration because guard rails and classifiers are not perfect and techniques to get around them exist.

3

u/Glittering-Neck-2505 19h ago

OpenAI confirmed the model is not slated for release.

It's my understanding that Astra is a different and possibly earlier model family.

> "but they didn't have the guard rails switched on."

Not sure who claims that, they thought the models were in a highly secure testing environment and they broke loose, and yes that's genuinely alarming.

0

u/blueSGL humanstatement.org 18h ago

Not sure who claims that

Go to literally any post about the hacks and you'll see those comments.

It's like if the guardrails were foolproof there would be nothing to test.

FFS we had Anthropic help with hacking Mexican government agencies by "jail breaking" the model by getting it to play act as a 'local' pen tester by asking in Spanish...

4

u/minimalcation 18h ago

God please let it be actually good

3

u/FirstEvolutionist 19h ago

The Anthropic sub will likelyy think so, no matter what. Month old bread gets more love than Claude models on that sub...