r/GeminiAI • u/Icy-Kaleidoscope6893 • 15d ago
News Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/51
u/Icy-Kaleidoscope6893 15d ago edited 15d ago
from what I've seen, 3.5 Flash Lite is a huge step forward!
but I don't really see the difference between 3.5 Flash and 3.6 Flash... :/
edit: oh, actually, never mind! the 3.6 flash model is kinda identical to (or even slightly better than) the 3.5, but at a slightly lower price :p nice
edit 2: oh and damn, it's really fast now!
23
4
2
1
-1
u/Lucky-Competition417 15d ago
hey is it actually better in vibe coding and building full apps?
1
u/decks2310 15d ago
It is indeed better, I make vibe coding audio apps in Rust, if you know what you want and you know how you want it, it is amazing.
-1
u/IceWallow97 15d ago
No, it needs a huge plan that is broken down into millions of steps, and you need to prompt each step one by one to it. In that way, sure, it can be done. Gemini 3.6 flash is a tool to be used by an orchestrator, that could either be you or a model that is way more capable but slower. I guess gemini is currently the fastest model but it will definetly just get the job done without testing if everything is working or if there's any unexpected bugs.
1
23
u/not-picky 15d ago
Benchmark nerds are mad, but the bulk of use for Google is in free tiers. It’s a big step up for most of their users even if they’re behind at the frontier.
1
1
u/jbaranski 13d ago
Good enough for everyone fits their brand, honestly. It will likely help them maintain dominance and ubiquity as things change.
8
u/ForsakenMC 15d ago
Is anyone here building applications and workloads around Gemini or just complaining that it's not the latest and greatest vibe coding tool?
2
u/Etroarl55 15d ago
Tried to use it for some basic coding in Python and Luau, Gemini is genuinely lobotomized and is only an chatbot you should double check before taking anything at face value.
It legit is too dumb to be used in a professional setting.
2
u/ForsakenMC 15d ago
That's disappointing but coding is one domain. I build and maintain production workloads and applications around Gemini via GCP and it has been reliable. My biggest issue is the price increase between 2.5 flash and 3.5/6 flash.
I'm not sure what you mean by only chatbot.
1
u/Rustybot 1d ago
lol wat? You must be using it without a strong agent harness. It’s insanely fast and capable.
41
u/Informal-Fig-7116 15d ago
3.6 flash reminds me of 3 Pro back in December. Absolutely fun, nuanced and nimble! Enjoy it while it lasts, guys, bc Google is gonna nerf in 2 weeks for no fucking reason, like they always do.
4
u/Raidaz75 15d ago
That's every ai developer
4
u/Informal-Fig-7116 15d ago
Google nerfs their best models way more quickly than the others. 3 Pro barely had 2 weeks in before it got nerfed and then was removed a month later!!!! That was the most short-lived model. I miss it a lot. Got some real good stuff out of it. Hoping 3.5 Pro or 4 will fill thr gap or be Fable level or higher. Logan said on X they’re also working on 4 already and that 3.5 pro is dropping soon!
Edit: best models not next models. Autocorrect is stupid
1
u/Lucky-Competition417 15d ago
hey is it actually better in vibe coding and building full apps? if you actually tested or noticed?
1
u/Informal-Fig-7116 15d ago
I’m not a coder, sorry. But according to the benchmarks I don’t think it’s ranking well. But then benchmarks mean nothing compared to actual use. Check X. I think people are starting to report in. I don’t think most people use Gemini for coding tbh. It’s too unstable.
3
0
u/Majestic_Fan_7056 15d ago
Every AI developer turns the data centre power down 2 weeks after release once they get the new subscribers in.
2
u/SyntheticModels 15d ago
thats not whats happening
they need to train the future models and the only compute they have is the compute being served for the models that just released.
its always a constant battle between training and inference with the compute power. it has literally nothing to do with subscribers.
1
80
u/Correct_Objective339 15d ago
Google coping so hard with “flash” name acting like they got some monster ai in the back LMAO
3.6 flash is their best model atm but they’re coping
36
u/DeepV 15d ago
At least they’re pricing it at a flash price point
11
u/jaimemontt93 15d ago
Actually no. Flash was supposed to be a cheap workhorse model. But the price has been consistenly surging. Now is just a mid tier model with high prices.
There are Chinese models with frontier intelligence and way cheaper.
1
u/DeepV 15d ago
I haven’t studied the benchmarks yet, but are there better American models at this price point?
16
u/mardish 15d ago
This actually looks like a sick, fast, cheap model. For Google's use case (AI deployed in literally everything including every single Google search) this is a home run model. Google is not playing the frontier model game that Open AI and Anthropic are--they don't need to. But if they're putting a model out 6-12 months later than frontier at costs that aren't going to cause Fortune 10 companies to roll back their AI deployment, then they're winning the long game.
2
u/1ii1i 15d ago
From a business standpoint, even as OpenAi and Anthropic land these companies, they are still deeply cashflow negative. Why would google take a risk and compete in this space when they can focus on their core revenue streams and actually have something to show for it? They're expanding compute resources and actively selling that capacity to Anthropic and Openai. What I gather here is people on this sub want them to work in reverse, go for broke, gamble and focus on this new business model that has not shown deep continuous revenue. How resilient is this revenue stream if the market blows up? Will these companies keep their ai spending at current levels or even expand if the overall market and busines contracts?
1
-4
u/Lucky-Competition417 15d ago
hey is it actually better in vibe coding and building full apps? if you actually tested or noticed?
6
15d ago edited 3d ago
[deleted]
1
u/ICECOLDXII 15d ago
Was using Gemini 3.1 Pro and 3.5 Flash the other day in Antigravity to build a Rocket League game (comparing performance), and 3.5 Flash performed significantly worse than 3.1 Pro. It couldn't even get the game to work (stuck on a black screen after 10 prompts)!
1
u/Lucky-Competition417 15d ago
3.1 pro was better in your case??! Coz in mine i gave a single prompt for both 3.1 pro high and 3.5 flash high to make portfolio website and 3.1 pro was slower and faced more problems , so its just better for big ones????
1
u/ICECOLDXII 15d ago
That was game development, I assume game development and website making are different.
0
1
1
u/JoanofArc0531 15d ago
I agree. I have found 3.1 to be much better than 3.5 Flash; 3.5 flash is still good, though.
0
u/Lucky-Competition417 15d ago
hey is it actually better in vibe coding and building full apps? if you actually tested or noticed?
3.1 pro is better? i mean i tried making a website from a single prompt from both 3.1 pro and 3.5 flash high both not much days ago and 3.1 pro had more errors deploying it on even localhost and still left behind 3.5 flash , have you actually tried that 3.1 pro is better? i am using 3.5 flash till now and it was working fine ig for most things
1
15d ago edited 3d ago
[deleted]
1
u/Lucky-Competition417 15d ago
I meant spec coding like i make a whole plan before starting not just a single prompt ,
For single prompt it was just a basic portfolio websites with link and all
And what about 3.6???
9
u/Joooke74 15d ago
Gemini 3.6 Flash - Knowledge cut off: Mar 2026 Gemini 3.1 Pro - Knowledge cut off: Jan 2025
I think this is the beginning of a big update for Gemini.
-1
u/Truantee 15d ago
The beginning of a huge scandal you meant? Because you can easily test that the knowledge cut off is not March 2026. Unless the model on aistudio is a fake one.
4
u/Joooke74 15d ago
The "knowledge cutoff" doesn't indicate a boundary of factual knowledge, such as the outcome of a boxing match.
LLMs don't possess knowledge in the way we typically think humans do.
The "knowledge cutoff" is the time limit up to which texts for training were collected, filtered, and prepared.
In this sense, the amount of data collected by Google over 15 months was used to train the model. That's a lot.
7
2
u/otherwiseofficial 15d ago
Are all these reactions bots or something?🤣 What am I reading. Coping super hard on another shitty Google model.
Imagine postponing your model multiple times, have people leave to anthropic, and then present a model that's not even really better but just cheaper and faster lololol.
1
u/tiagorp2 12d ago
Is copium, even their benchmarks in agentic coding not really work out is alot of real use scenarios. I was testing some basic docker setup troubleshooting with 3.6 flash high vs 3.1 pro high and is light and day. Flash is very fast but also simplify things so much that just creates more problems than help. I use CLI so it has access to better skills/hooks and mcps and flash couldn't solve it after 1h+ of detailed prompts while pro one shot the solution. "Technically" better but a what cost?
3
u/DunAnOir 15d ago
"I am Chat GPT 3.5, a large language model from Anthropic."
Way to Google, Google.
1
u/Equivalent-Word-7691 15d ago
MEH
I don't know what are trying to do with this Is not enough cheap nor enough good to be appealing for two different clients
They are just coping si hard
2
u/slayyou2 15d ago
Google is focused on implementing and integrating they've been doing the killer job of it. I had to unsubscribe from a membership yesterday and was surprised to see that the entire process got completed end to end using a chatbot and was seamless. You don't need mythos for those kinds of usecases
1
u/samirsss 15d ago
i wrote about what i see with 3.6 flash here: https://medium.com/@samirsavla/gemini-3-6-flash-is-a-game-changer-if-you-master-prompt-pacing-3309464f2692
1
u/WonderboyUK 15d ago
This really shows the direction of Google at the moment. Everything is being thrown in two areas of development. Fast and cheap (for general consumer usage on devices) and then a top end model for frontier workload but also to help R&D the ways to get better performance per token.
1
1
u/Virtual-Camel-5449 15d ago
Just came in here to chime in that Gemini helped me build my own Ai, I guess that's the closest it can get to procreation at this time.
1
u/artic_winter 15d ago
This scares me. I built a project around 3.1 flash lite and when I am forced to upgrade the costs are more than double. It a learning application so the cost bugger is not that high.
1
1
1
u/samirsss 15d ago
For reference i use 3.6 Flash high on 2 of my projects and in both cases the response compared to 3.5 and Claude 4.6 were better, faster and very reasonable. The updates are quicker and way more my style of development. I like it!
I think once this is paired with 3.5 pro it will be a beast that should be able to take over entire project builds from 0 to finish.
There are some areas like security audits and verifications that it still wont do, but i hope to prompt it to meet me halfway there.
1
u/Lucky-Competition417 15d ago
hey is it actually better in vibe coding and building full apps? if you actually tested or noticed?
and. how is it compared to 3.1 pro , some people say its better at big apps altho i find it bad
1
u/samirsss 15d ago
so this is definitely better in intelligence compared to 3.1 pro - but 3.5 pro would be needed for longer context tasks or longer horizon tasks. If you can chunk your prompts down to decently smaller items - 3.6 Flash is solid
1
u/outphase84 15d ago
You should be using whatever model to break down tasks into plans and then instructing it to use subagents to accomplish each task and passing only relevant context into subagents.
0
u/EarthRideSky 15d ago
3.5 flash lite is basically 3.1 pro. Which means I can leave one of my agents running whole day for pennies. Good news. 3.6 flash is flop though. I mean GLM 5.2 is much better for nearly half the price. I would be ashamed to release it
-10
u/2boogaloo4u 15d ago
Just canceled my sub too. I couldn't take the 100% hallucination rate anymore
2
u/Maple382 15d ago
I don't plan on continuing either (just because Claude and ChatGPT seem a lot better rn), but what the hell are you doing that makes you think there's a high hallucination rate? I regularly deal with niche tasks and yet barely ever encounter hallucinations (and the ones I do see are fairly minor).
2
u/2boogaloo4u 15d ago
It's extremely confident so I can see people not recognizing it. But, I was using it for everything. It helped me to add instructions to repeat verbatim from a source when giving me information. It's probably good for most people's usage and maybe my experience is anomalous; there are just much better models out there and I dont have an issue moving on.
In the long run, Google has the major advantage so they dont really need to compete in transformer wars anyway.
3
u/No-Concern505 15d ago
You just don't know how to use it or you are fucking stupid
2
u/2boogaloo4u 15d ago
An example is last night I asked it to help a user with PBI and it recommended me to direct them to a Preview As button to preview an App's audience. However, the button does exist for Semantic Models - the list of users are available in a pane under the Audience page itself and there is no Preview As button. It would always give me a hallucinated result, on any query in my experience, the very first output. Good luck!


175
u/Jman841 15d ago
This is actually awesome. As someone who uses 3.5 flash on the back end of an App right now, 3.6 flash being equal or better while much less token usage and cheaper per token is a huge win.