Top
Best
New

Posted by Philpax 17 hours ago

GLM-5.3-Flash(z.ai)
https://news.ycombinator.com/item?id=49450353
972 points | 497 commentspage 4
Aboutplants 12 hours ago|
When do Chinese models surpass US models? I thought there was at least be a 2 year runway but now I think they surpass it within 12 months, if not sooner.
danieltk76 7 hours ago||
tbh I wasnt that impressed by it. initial benchmarks were trying to say it was AGI but i told it to re-build Palantir in 1 pass and it gave me a non working prototype
garo-pro 17 hours ago||
> Combined with our latest 30T-token multimodal pre-training corpus [...]

Is the optimal formula still 20x the amount of model params in tokens for training? Could this mean we're getting a GLM with 1.5t params?

yousif_123123 15 hours ago||
Will we need all the data centers being built or will improvements in software and hardware allow the majority of AI workloads to run locally or in the cloud but way more efficiently than was projected when all the plans were laid out?

Like were executive at Google and AWS and Microsoft expecting this kind of performance from models smaller than what openai/anthropic have been doing? Are we really in a "compute desert"?

bakies 12 hours ago|
If it gets more efficient it'll be more enticing to expand use case. Personally I'm hoping to do a lot at home but I'm not counting the datacenter building as a bad move at this moment. It may and up that way.
jatins 15 hours ago||
I was quite surprised that Zai had deep pockets to serve this free for a week. My first guess was this was an American lab like xai or google
pohl 11 hours ago||
Does the word "flash" mean a specific thing when it comes to LLM models? I noticed that this word is used by gemini, qwen, and z.ai and I'm curious does it mean the same thing for each one, or did they all just accidentally brand similarly?
Doohickey-d 11 hours ago|
It seems like it has come to mean "fast, small, cheap" models these days, and seems well enough understood as such that different AI labs are adopting it.
Tepix 12 hours ago||
GLM 5.3 Flash: 320B parameters with 18B activated

Qwen 3.8 Next Flash: 125B + 51B = 176B parameters with 6B activated

DeepSeek V4 Flash: 284B with 13B activated

The new Qwen model is the most promising for one or two Strix Halo 128GB with the low number of active parameters. On paper it's much stronger than Qwen 3.8 27B.

epolanski 17 hours ago||
I'm starting to think that this whole sanctioning China may motivate and prompt them to do more and better in every field.

It's too big, bright and resourceful of a country to choose confrontation instead of collaboration.

ricardobeat 17 hours ago||
Starting? This was obvious way back in 2019, when the US decided to give China a little push developing their own silicon industry.
abroszka33 10 hours ago|||
> It's too big, bright and resourceful of a country to choose confrontation instead of collaboration.

It's not like we didn't try it. China first have to learn to make deals where both party benefits.

esperent 17 hours ago|||
This has been clearly stated as what would happen going back several decades at least.
pshirshov 15 hours ago|||
> I'm starting to think

That's good. Keep going.

himata4113 17 hours ago||
Well the big problem with china is that they do not respect international law when it comes to technology theft. But that argument is very weak when it appears that a lot of what they do is out in the open for anyone to replicate.
nananana9 16 hours ago|||
That's how you catch up when you're behind.

Now the US is behind in EVs can you guess what they're doing? [1]

[1] https://evwire.com/p/video-ford-ceo-jim-farley-says-they-fly...

himata4113 16 hours ago||
"argument is very weak" regardless as I said.
epolanski 16 hours ago||||
No major power respects nor cares about international law.

Intellectual property is part of WTO agreements but enforcement is domestic.

US companies do it too, regularly, they simply hire and poach staff from competitors.

Proving it to be IP theft is difficult unless you can prove documents being passed. But often all you need is the know-how of the hired talent.

cyanydeez 16 hours ago||||
yeah, America is totally out there respecting international law.

"problem" indeed.

fwip 16 hours ago|||
There isn't one global "international law" for copyright. There are treaties that countries negotiate with each other.

If the USA wanted a copyright treaty with China bad enough, we would negotiate one. China is not breaking any laws here, international or otherwise.

BeetleB 15 hours ago||
The key difference between this and all other GLM models is it's multimodal. You cannot send images to the other GLM models.
mrinterweb 15 hours ago|
I really wish GLM models had vision capabilities. I've worked around that in the past to use a vision MCP in my harness that GLM can call. It is not the same, but it allows the model to query images.
BeetleB 15 hours ago||
Well, now one of them does!
mrinterweb 12 hours ago||
That's wonderful. I was going off an older version of the Artificial Analysis page for GLM-5.3-Flash https://artificialanalysis.ai/models/glm-5-3-flash. The page is updated now to show that it does support multi-modal image inputs.
rahimnathwani 17 hours ago|
Related: https://news.ycombinator.com/item?id=49446422

(281 points, 118 comments)

More comments...