Top
Best
New

Posted by logickkk1 21 hours ago

Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber(blog.google)
https://console.cloud.google.com/agent-platform/publishers/g...
701 points | 532 commentspage 13
npn 20 hours ago|
tested the models on aistudio. despite that the knowledge cut off is march 2026 it still knows nothing about 2025!

you can check by asking "list notable world events in 2025, only list unplanned" on aistudio. or you can ask for Charlie Kirk, it also does not know. I tried it multiple time to ensure that I didn't not get routed to older models!

> but google has search

irrelevant, without deeper knowledge about cutting edge technologies or latest libraries, all of it suggestions are crap. even you ask it to search it will still use outdated keyword thus only getting outdated information.

in other word, what a disaster!

npn 4 hours ago|
update: the knowledge cut off date is "unknown" now.

funny because some people downvoted me believed that there is no relation between knowledge cut off date and real world events. that's not how it works!

kthinckley 20 hours ago||
Google desperately needs to make some leadership changes within their Gemini team now that they've been surpassed by 3-5 open weight models and risk loosing frontier status all together in the near future.
alephnerd 20 hours ago|
Open weight models aren't likely to be open weight in the long-term. China has started considering export controlling and limiting access to model weights [0].

[0] - https://www.ft.com/content/6049a031-9e9b-464c-97bb-414da04d5...

ErneX 20 hours ago|||
That contradicts this:

https://www.wsj.com/tech/ai/chinas-xi-touts-open-source-ai-a...

So who even knows.

logicchains 18 hours ago|||
They don't need to be open weight in the long term; once there's an open-weight Fable-level model with 1M context it'll be pretty much good enough for all coding tasks, no need for new models.
game_the0ry 19 hours ago|
At this point, I think google should consider becoming a hyper scaler for anthropic and open ai, and I predict that that is exactly what they do. The model is no longer the most valuable part of the stack.