Posted by whiteros_e 6 days ago
> Why would I pick GLM over Claude?
To support the company that makes their model weights available for download, while Anthropic lobbies to restrict access.GLM: 80 USD (pro), 168 USD (max) -> with "limited-time event" discount this becomes 56 USD and 117.6 USD
I also don't understand why are they so much costlier, and I would also like to give it a try.
Anthropic's Pro is $20 and corresponds to Z.ai's Lite at $18
Anthropic's 5x Max is $100 and corresponds to Z.ai's Pro at $80
Anthropic's 20x Max is $200 and corresponds to Z.ai's Max at $168
I have both plans. Claude monthly €20 and Z’s €18 monthly. Running GLM-5.3 high on their monthly plan will hit quotas absurdly fast compared to Opus 5 High on Claude code. It’s almost unusable for AI driven development. I ended up using the Z plan for using GLM-5.3 as a detailed security reviewer and adversarial feedback. For that, it is much better than Opus which will flag and bail out for even simple security tasks that are aimed at defense.
But, it was enough for a customer like me who tried them at good faith to walk away and find their competitors..
I like the diversity of LLMs as of today and prefer to not tie myself to one big plan with any vendor. If they don’t prefer me as a customer, then I will accept that, and move away.
GLM's "Max" plan is (was?) equivalent to 3x Claude's 20x ($200) plan.
There's also a new free stealth model available that's more likely than not in the GLM family. This seems to happen every few months for a week or two and represents a good savings opportunity.
Creators of known unreliable programs be surprised their programs are unreliable.
All I can recall reading from OpenAI about what they have actually done in the name of "RSI" is using one of their models to help automate the training process.
That's more than Anthropic has done though.
Ziphu seem much more matter of fact about it. To me it' a shame that they've decided to the use this "RSI" name, but at least they are being fairly specific about what they mean by it, while the western companies seem to want to invite you to think it's more than just dogfooding and automation.
Signed, a customer.
Also maybe we can stop saying "we can't slow down because China will never slow down" - I don't really think slowing down is right, BUT if slowing down is correct then maybe we should be talking about China slowing down instead of just saying "won't happen" without any evidence that Chinese labs don't have similar concerns.
One drive gives about 2 tok/s; striped across four drives it reaches 3.5 tok/s with byte-identical output, and our best internal build with a not-yet-published patch does 4.2.
Method and numbers: https://github.com/argonautlabsai/argodrive (built on antirez/ds4).