Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

100% Positive

Analyzed from 439 words in the discussion.

Trending Topics

#glm#plan#model#more#flash#coding#own#month#providers#bit

Discussion (23 Comments)Read Original on HackerNews

dada216•37 minutes ago
We built a complete production-grade inference service from scratch on a cluster of more than 100,000 Chinese-made AI accelerators. All production inference for GLM-5.3-Flash runs on this system.
tefkah•26 minutes ago
> Today, GLM-5.3 has become an indispensable daily coding partner for everyone on the team, and it is moving steadily toward replacing us. If this trend continues, given enough compute and enough time, its endpoint is a system that can design and train its own successor entirely autonomously. This is known as Recursive Self-Improvement, or RSI.

Statements dreamed up by the utterly deranged.

embedding-shape•34 minutes ago
I was gonna ask how people found their coding plans, and realize, have they massively ramped up the prices? Seems the middle plan is ~$80/month now, didn't that used to be like $20/month? Cheapest plan is ~$20/month currently.

They must have hit really hard scaling limits if the prices were hiked so much so quickly.

asp_hornet•21 minutes ago
The way I look at it, their coding plan doesn’t retain data or use it for training making it one of the cheaper plans for me.

https://docs.z.ai/legal-agreement/privacy-policy

andy_ppp•18 minutes ago
You believe any of these companies care about the law? They care about winning and building the self improving AI as quickly as possible.
asp_hornet•15 minutes ago
I too am sceptical but I’ll take my chances. At least it’s helping the open weights.
Daviey•22 minutes ago
I paid $360 annual for Max plan and currently averaging about 1BN tokens a day with their frontier GLM-5.3 model. This was clearly unsustainable for them and they've dropped this package.
broodbucket•31 minutes ago
Yeah it went from a great deal to unviable compared to other providers imo. They really need to find a healthy middle ground
bbor•22 minutes ago
It's hard to know, since no one advertises the actual token limits (partially cause they're prolly complex / adaptive). So it seems much more likely that they just offer different pricing tiers than you're used to. Like, the $80 plan is still ~$80 of subscription quota, regardless of what else is offered.

For [API usage](https://openrouter.ai/z-ai/glm-5.3-flash#providers) they charge a bit more than the very cheapest providers of GLM-5.3-Flash, but not so much that a big price difference would make sense.

bbor•27 minutes ago
Well, other than the infrastructure they got from illegally routing millions of paying customers' requests through Anthropic's Opus 4.8 in a distillation attack...
jensb1•24 minutes ago
What is "illegal" about it?
rob74•9 minutes ago
This article left me with one immediate question: "WTF is GLM?".

Honestly, I have no idea what z.ai is either (I'm aware of an AI-enabled editor called Zed, but that's under zed.dev), so it's a bit presumptuous from them to assume that everyone is familiar with their product...

jbonatakis•3 minutes ago
z.ai is a fairly well known AI lab out of China and their GLM models are probably the most popular outside of Anthropic or OpenAI’s. I don’t think it’s presumptuous for them to not introduce themselves in a post on their own blog, I think you’re just a bit out of the loop here.
fxwin•4 minutes ago
It's presumptuous for them to assume that a reader of their blog is familiar with their product?

Also I feel like the obvious way to read the very first sentence is that GLM is a language model

> As we develop GLM, the model sometimes exhibits capabilities that surprise us

drbscl•3 minutes ago
>As we develop GLM, the model sometimes exhibits capabilities that surprise us, and even unsettle us.

Come on now

Also, why would they introduce themselves on their own blog?