Top
Best
New

Posted by pella 14 hours ago

GLM-5.3: Frontier coding with emergent cyber capabilities(z.ai)
949 points | 477 commentspage 3
mraza007 14 hours ago|
Such an interesting times we are in,

We just had amazing releases this past two months

kimi k3, glm5.3 qwen3.8 and now glm5.3

These open models are getting really good

w4yai 11 hours ago|
You wrote GLM5.3 two times :)
InsideOutSanta 5 hours ago|||
An LLM so nice, they named it twice.
czottmann 9 hours ago||||
Because it's doubly good.
mraza007 5 hours ago|||
Sorry , It was 5.2 :)
moinism 11 hours ago||
Google: Here is the next iteration of our flash model series, with a discount. please use. thx.

Z.ai: Here is our next iteration, neck and neck with Fable/Sol. weights releasing in two weeks.

jamesponddotco 5 hours ago||
Is there a plan somewhere that gives access to Kimi K3 and GLM-5.3? I was thinking of testing both to run security reviews of my code.

I know OpenCode Go has both, but their limits seem kinda low, so I'm not sure how feasible it is to run such a task with them.

CuriouslyC 5 hours ago||
These results look pretty good, given the smaller model size and the GLM family's historic robustness. Cheaper than Kimi and more robust than DeepSeek. The question in my mind is if you're going cheap, are you going to stop here or go all the way down to DeepSeek Flash?
alienbaby 8 hours ago||
One htought I had; if The chinese allow unfettered access to cyber capabilties while th US does it's best to neuter it's model releases, from China's point of view they have the US all tied up in knots dealing with problems they don't give people the tools to solve. China giggles as it watches the US under threat from people using it's models. The US is restricting citizens from owning this particular kind of weapon, while China is handing it out to the wrolds citizens freely. It feels like the US would only come out worse overall?
onlyrealcuzzo 8 hours ago|
I suspect Anthropic wanted the US gov to ban Mythos for marketing.

If it turns out to be bad for them, the US gov will likely suddenly unban models.

swalsh 8 hours ago||
I suspect mythos demonstrated a fully autonomous offensive hack in a similar way Open AI's models performed, and the government is reacting to it the same way we reacted to blackhat.

The threat is real.

ikari_pl 4 hours ago||
Such a smart model and didn't warn them how confusing the headline is to anyone who understands what "cyber" means?
andai 5 hours ago||
We got nukes capable of having existential crises, before GTA 6...
maxdo 7 hours ago||
They just ignore in their benchmarks opus 5 for some reason :) also grok 4.6 . I wonder why
bigyabai 3 hours ago|
Opus 5 has nerfed cybersecurity performance, making it hard to benchmark: https://support.claude.com/en/articles/14604842-real-time-cy...
scottfits 4 hours ago|
what i appreciate most about this post is the level of transparency in how they built and scaled an RL pipeline. my friends at the big labs are so cagey about everything, and Zai is just putting out a great crash course for free.
More comments...