Top
Best
New

Posted by logickkk1 12 hours ago

Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber(blog.google)
https://console.cloud.google.com/agent-platform/publishers/g...
624 points | 501 commentspage 6
goldenarm 10 hours ago|
LLM reception is truly extreme, even worse than AAA game releases.

Ever frontier lab lived it at least once : missing the frontier by a few months triggers extremly negative reactions, then you take back the lead for 2 weeks, and the hype cycle repeats.

Andrex 6 hours ago|
I spin a mental roulette on whether the reception on a new release will be "OMG best model by far, no one will be able to catch up for months!" or "OMG this is already outdated, RIP company X, they might as well just give up now, there's no coming back from this."

It's quite a fun game. I click into the comments and see if the roulette wheel was right.

ComputerGuru 12 hours ago||
So 3.6 Flash is a somewhat of an admission that Google miscalculated by charging 3-5x for 3.5 Flash what it did for 3.0 Flash (3x input and output costs plus large token inefficiency changes) despite only modest improvements?

3.5 Flash Lite is only a hair cheaper than 3.0 Flash, but I think 3.0 Flash is a massively more capable model?

mfkrause 12 hours ago||
Pretty underwhelming, as expected honestly. I don't want to know what morale is like at DeepMind right now.
WarmWash 12 hours ago||
Especially when Google owns 15% of anthropic and serves them compute. Double especially when your boss (Hassibis) is also an early investor in Anthropic. Hell his NW might be more Anthropic than Google.
drob518 12 hours ago||
Yep, agreed. They still are not releasing anything frontier-class (Gemini Pro) at this point. Feels to me that they keep getting scooped by others (e.g. Kimi 3) and then are retrenching.
dvduval 12 hours ago||
It does seem like their releases are getting closer together. I get the feeling they realized they were trying to roll out to their entire ecosystem and now they’re focusing more just directly on the AI model itself. I think give it a little time and they’ll start to be one of the competitors too.
MILP 10 hours ago||
I'm a big fan of the Flash-Lite models. They're exceedingly fast and deliver great outputs for high volume use cases where you need to process requests at scale. Can't wait to try the newer version.
Havoc 11 hours ago||
Flash Lite: 0.3/m and 2.5/m

Deepseek Pro: 0.435/m 0.87/m

That's wildly ambitious pricing by Google. You can maybe get away with spicy pricing at the SOTA edge but at the lower tiers everything is a lot more price sensitive.

JacobAsmuth 10 hours ago|
You need to compare cost per task buddy boy. Cost per token doesn't tell you much when you don't know how many tokens a model will use to accomplish a task
Havoc 9 hours ago||
>boy

Seriously?

dumberquestions 12 hours ago||
"..and in some benchmarks like DeepSWE by Datacurve, we observe up to 65%, all at a lower cost per output token."

"3.6 Flash delivers higher precision with fewer unwanted code edits and reduced execution loops, as seen in DeepSWE (49% vs. 37%)"

So which one is it? 65% or 49%?

petu 12 hours ago|
First sentence is about token efficiency.
dumberquestions 12 hours ago||
You're right, should've gotten some LLM to summarize it instead of skimming.
semilin 12 hours ago||
Or you could have read it more closely before posting a comment saying it didn't make sense. You know, the old school way.
dumberquestions 12 hours ago||
If I'm going to read it wrong might as well have an LLM to blame.
XCSme 11 hours ago||
tl;dr: 3.6 flash is a bit smarter than 3.5 flash, but also a bit more expensive.

My results [0] put Gemini 3.6 Flash at the top.

3.6 Flash high has same $1.5 input price as 3.5 Flash, but output is cheaper from $9.0 to $7.5.

Google said 3.6 Flash is more token efficient, but in my tests it's actually LESS token efficient[1] than 3.5 Flash, so despite the output price reduction, it still costs more.

[0]: https://aibenchy.com/compare/google-gemini-3-6-flash-medium/...

[1]: https://aibenchy.com/compare/google-gemini-3-6-flash-high/go...

XCSme 11 hours ago||
I was expecting 3.6 Pro. It's been so long since the last Pro model...
thebigspacefuck 11 hours ago|
They are working on coming up with a better code name. You know, something like ”Fable” or ”Sol”, gotta have one these days. Personally I think they should go with “Mafia”. How cool would that sound? 3.6 Mafia.
dpacmittal 8 hours ago||
3.6 Gangsta Pro
pietz 12 hours ago|
Are they comparing 3.6 Flash to 5.6 Luna and losing? That's ruff.
polski-g 12 hours ago|
Why wouldn't they? Luna isn't a Flash model. OpenAI hasn't released a flash-equivalent model since gpt-oss-120b.
pietz 11 hours ago|||
Did you ask me a question and then answered it yourself in the very next sentence?

Anyway, given that both Gemini and OpenAI have 3 sizes of models, one would think Google compares their medium size to OpenAIs.

More comments...