Live data from Hacker News

GLM-5.3 is now open-weight

huggingface.co

31–40 of 296 posts

Re: GLM-5.3 is now open-weight

#31
I previously posted that DS4Flash was _good_ but not _great_ on two DGX Sparks, but I have to say that GLM-5.3 is pretty amazing. It's been able to tackle all the random hard problems I've thrown at it and it has the intuition that DS4Flash seems to lack.

We're nowhere near a Fable-class model IMO, but things are going to get interesting in this next year.

Re: GLM-5.3 is now open-weight

#32
post #2

I've been using it more and more. Feels like Opus 4.8, in the best possible way.

I'm starting to think Opus 4.8 is significantly smaller than most people assume. If it's significantly larger than GLM 5.3 (I've heard some insane guesstimates out there like upwards of 5T params or more), that would prove rather embarrassing for Anthropic.

It seems like there is tradeoff between model size and the need for tool use, which - in my mind - is quite costly in terms of time and tokens. More detailed world knowledge requires an exponential increase in model size, but most knowledge can be acquired ad hoc using search or database queries. This will fail for questions where the model lacks the knowledge to ask the right questions, but maybe this could be solved by a handful small inquiry models with different knowledge encoded in their weights?

Re: GLM-5.3 is now open-weight

#33

GLM 5.3 is probably the sweet spot open weights model if you want to go beyond deepseek flash or the new glm flash. I used it with pi and had a fairly good time, especially since it’s less touchy about cyber and whatnot than the US guys. It’s slightly behind Kimi in ability but it’s a lot easier to run it, I’d expect prices (and speed!) from third parties to be noticeably better. Assuming you’re willing to drop a fat…

Well if you did get the m5 ultra could you obliterate the guardrails and then your wife can ask it pertinent but unsafe questions about how to punish you. Seems doable.

Re: GLM-5.3 is now open-weight

#34
post #22
post #14

Earlier quoted context omitted.

z.ai is using all Chinese hardware for flash: https://thenewstack.io/glm-5-3-flash-chinese-chips/ There are other providers with much faster inference, like BaseTen at >100t/s: https://openrouter.ai/z-ai/glm-5.3-flash#performance

How do I find out where the openrouter model providers' servers are located?

If you click on the provider name, the panel that pops up shows a "Region" value. Not every provider lists their region, however.

Re: GLM-5.3 is now open-weight

#35
post #25

I'd like to ask Sam Altman if he still thinks that it's too dangerous to publish GPT-3. I mean, no one would use it, but what is his reasoning for not publishing it now, in 2026?

>but what is his reasoning for not publishing it now, in 2026?

What's the point of publishing it when it'll likely be outclassed by gpt-oss?

Re: GLM-5.3 is now open-weight

#36
post #25

I'd like to ask Sam Altman if he still thinks that it's too dangerous to publish GPT-3. I mean, no one would use it, but what is his reasoning for not publishing it now, in 2026?

They already publish gpt-oss which is several generations better than gpt-3

Re: GLM-5.3 is now open-weight

#37
post #2

I've been using it more and more. Feels like Opus 4.8, in the best possible way.

Do you use it to write HTML/CSS? Javascript? C++? There's a huge difference in ways people use models and if you are not specific about it then your comment means nothing, unfortunately.

Re: GLM-5.3 is now open-weight

#38
post #25

I'd like to ask Sam Altman if he still thinks that it's too dangerous to publish GPT-3. I mean, no one would use it, but what is his reasoning for not publishing it now, in 2026?

I think it would be an important historical document as well. We are potentially looking at the dawn of AGI and one of the most important models ever created. Each model is also a kind of ultimate time capsule, containing a snapshot of the entire human collective mind. If you wanted to ask a 2002 person what they thought about future historical events you can just ask them directly.

Re: GLM-5.3 is now open-weight

#39
post #13
post #5

GLM-5.3-Flash is actually cheaper than deepseek and better than deepseek but no one is talking about yet :)

It’s cheaper sure, but it’s very slow. It’s not a drop in replacement

I think we don't have a good draft model for better speculative decoding yet (e.g. DFlash 2). Once we do, it will be faster.
Post reply on HN