Top
Best
New

Posted by logickkk1 11 hours ago

Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber(blog.google)
https://console.cloud.google.com/agent-platform/publishers/g...
614 points | 487 commentspage 5
mrandish 6 hours ago|
I often use Gemini free web chat because it's generally quite good at web search-related questions (apparently it has direct token-level access to the Google Search index) but I noticed in the last two weeks output quality of 3.5 Flash seriously degraded. Maybe they were switching over systems.
culi 6 hours ago|
Google's Knowledge Graph is a massive advantage no other competitor has. I don't think they've fully utilized its full potential but I don't know if any other company could've built something like Scholar Labs
Andrex 5 hours ago||
Knowledge Graph + automated transcriptions of almost every YouTube video = giant untapped moat of data
weird-eye-issue 1 hour ago||
It's not exactly untapped, my AI company has scraped YouTube transcripts for 3 years now for RAG
mythz 10 hours ago||
Always happy to see new Gemini releases as IMO Antigravity Pro 16.67/mo plan (Annual) is still the best plan available and have been pretty happy with Antigravity IDE.

If it wasn't for Gemini/Antigravity I'd have to go with a Max Claude plan, as it stands now I can get by with just a Claude Pro plan to get Opus when I need it, whilst using Antigravity as my day-to-day workhorse.

Unfortunately Gemini Flash became too expensive to use as a general purpose model (i.e. for AI features in Apps), luckily there are plenty of cheaper Chinese models to fill that gap now.

Andrex 5 hours ago||
What's the current outlook on Antigravity IDE vs. 2.0? How long will they begrudgingly keep it going before kicking everyone to 2.0/3.0?

(I actually use a mix of both for some offline projects, nothing serious.)

ValentineC 10 hours ago||
Why do you think it's the best plan available?
JacobAsmuth 9 hours ago||
(Rate limits * capability of the model) / cost
zwaps 5 hours ago||
Here's the issue:

GLM 5.2 is better, also cheaper, and almost as fast.

So essentially, a big L for Google. Combine this with them not being able to produce a frontier model this generation... hmm implications

HDBaseT 1 hour ago||
Counter-point, cost per task is almost the same ($0.47 vs $0.50) between GLM 5.2 and Gemini 3.6 Flash. [0] Not to mention the subscription plans likely produce 10x value compared to GLM 5.2 API, unsure the rate limits on a equal subscription vs subscription, but Google subscriptions offer tons of other value, including 1 year of Free Gemini for Education accounts.

[1] https://artificialanalysis.ai/models/gemini-3-6-flash

WarmWash 5 hours ago||
3.6 is roughly 50% faster, which isn't totally insignificant for being marginally more expensive.[1]

[1]artificialanalysis.ai

zwaps 5 hours ago||
Sure, but there's no sota alternative from Google. That's it, and its beaten by GLM 5.2 on every measure except somewhat speed.

I find that quite staggering. GLM is open weights

Narkov 3 hours ago||
Speed is definitely a marketable quality. All these things are a trade-off and solely measuring against SOTA I don't feel is always helpful.
parsimo2010 10 hours ago||
Feels like they released this to ride the wave of press of GPT-5.6, Kimi K3, and Qwen 3.8. Doesn't feel like Google has much substance with this post except a bump in version and tweaked their pricing.
sagex 10 hours ago||
Don't know why are they even pursuing Gemini. Just download the Kimi, call it Kimini and serve it on your GPU. Maybe then train next architecture based on this!
JeremyHerrman 9 hours ago||
Gemini 2.5 Flash-Lite has been my go to for cheap document processing at scale (especially with 50% off batch mode), but they are really boiling the frog with pricing increases with each version:

gemini-2.5-flash-lite: $0.10 input / $0.40 output

gemini-3.1-flash-lite: $0.25 input / $1.50 output

gemini-3.5-flash-lite: $0.30 input / $2.50 output (a 6.25x increase over 2.5!)

Now watch them deprecate Gemini 2.5 Flash-Lite in the coming months...

JacobAsmuth 8 hours ago||
How has your experience been with Gemma 4?
tjwebbnorfolk 9 hours ago||
gemma4 is the same price as 2.5-flash-lite, and performs better.
lambda 10 hours ago||
3.6 Flash scores exactly the same as 3.5 Flash on the Artificial Analysis index. Better on some tasks, worse on others. Mostly within what I'd consider the noise window. Looks pretty much indistinguishable from 3.5 Flash, at least on these benchmarks: https://artificialanalysis.ai/models/gemini-3-6-flash
summerlight 10 hours ago||
Looks like 3.6 Flash is the first model with their newest pretraining run (cutoff date is 2026/03), long after 2.5 series.
kilroy123 11 hours ago||
I deeply wish Google would focus on models like Gemma. Small, powerful, open-weight models you can run on phones or regular computer hardware.
mediaman 10 hours ago||
Gemma 4 was released in April. It's a good series of multimodal models.
accountrequired 9 hours ago||
gemma 4 thinks joe biden is president
kzrdude 6 hours ago|||
I asked Gemma 4 E2B, and if you use it as a reasoning model, it will give a better answer (that it doesn't know; it was also using the date information from the prompt.)
mediaman 8 hours ago|||
Small open source models shouldn't be used for world knowledge, that's not their purpose.
Petersipoi 8 hours ago||
Why not? Seems like a cop out.

Being able to ask questions to small open models seems.... obviously useful?

mediaman 7 hours ago||
Because they don't have a lot of parameters to store general Wikipedia knowledge. They're small. Use big models that have high parameter capacity to store general information. Or build a harness around the small model that searches a knowledge base/internet.

Use the right tool for the job. It's like asking why a screwdriver isn't good at sawing wood, or calling C a terrible language because it's hard to make CRUD apps with it.

lanthissa 11 hours ago||
i mean eventually then will, losing means open source, vertically integrated hardware means you can opensource and win on cost
thebigspacefuck 10 hours ago|
IMO Gemini has the best free tier models/app for everyday use. Muse-Spark is perhaps just slightly better, but has none of the connectivity to my GApps (for things like “create a recipe in my Google Docs from this image”).

Plus they are probably running these things on every Google search so saving tokens is a huge win for them.

copperx 9 hours ago|
Free? Did I misread the pricing details?
thebigspacefuck 2 hours ago||
Free plan, the default tier without requiring a subscription. If you use through the Gemini App or gemini.google without paying anything, the model used is 3.6 Flash.

Rankings for text are here https://arena.ai/leaderboard/text

For comparison of Free Tiers: - Gemini serves 3.6-flash (rank 12) - ChatGPT serves 5.5-Instant (rank 23) - Claude serves Sonnet 5 (rank 27) - Meta AI serves muse-spark-1.1 (rank 5)

While Meta AI serves the better ranked model, it doesn't end up working that well for other things. For example, if I ask "help me buy a new raincoat", it ends up suggesting a Cambodian website, whereas Google is well integrated with Google shopping. It doesn't have the same integration with GApps outside of Gmail/Calendar. A few other email connectors are available.

Claude has one of the best interfaces with connectors, skills, and plugins galore, but the model and limits are restrictive on the free tier.

Gemini, as far as I know, I've never hit a rate limit on Flash.

I believe Gemini is going to gain market share through the free tier funnel while serving models as cost-effectively as possible. People are going to use Gemini because they use GApps and Google.

ChatGPT and Anthropic are going to be competing for the API/Business users, but for everyone else they are going have to become Google before Google becomes them.

More comments...