Live data from Hacker News

GLM-5.3 is now open-weight

huggingface.co

121–130 of 297 posts

Re: GLM-5.3 is now open-weight

#121

Earlier quoted context omitted.

It's not that GLM5.3 in full precision unquantized is any smaller, it's 141 * 5.4GB files at approx 770GB which is about the same size as 5.2.

Hold on.. the routed experts are in FP8 now? Previously they were in BF16. Nice, this shall cut my download time by half!

This time they just made FP8 "default", accompanied by "-BF16" model/page (previously "-FP8" was released alongside).

Re: GLM-5.3 is now open-weight

#122

I previously posted that DS4Flash was _good_ but not _great_ on two DGX Sparks, but I have to say that GLM-5.3 is pretty amazing. It's been able to tackle all the random hard problems I've thrown at it and it has the intuition that DS4Flash seems to lack. We're nowhere near a Fable-class model IMO, but things are going to get interesting in this next year.

do you mean GLM 5.3 flash?

Re: GLM-5.3 is now open-weight

#124
post #29
post #11

h/t to DeepInfra for being the first 3rd party provider for it on OpenRouter ( https://openrouter.ai/z-ai/glm-5.3?endpoint=b711bea7-3994-49... ).

on my TrustedRouter: z-ai/glm-5.3: also Z.ai, Novita, Atlas Cloud, IO.NET

How am I supposed to navigate around there? For example the pricing page is empty or is that how it was supposed to look? On models and providers pages there are lists but no way to filter or get any kind of meaningful info. Or is this a WIP/POC?

Re: GLM-5.3 is now open-weight

#125
post #25

I'd like to ask Sam Altman if he still thinks that it's too dangerous to publish GPT-3. I mean, no one would use it, but what is his reasoning for not publishing it now, in 2026?

There’s not such a straightforward relationship between safety and model sis. According to the book The Thinking Game, lower quality models at that time were considered less safe, because they could be easily tricked into doing harmful stuff. In the book, Dario (of Anthropic) was the head of safety at openAI and was responsible for pushing for 10x scaling in training to make the models safer . It does make sense, a s…

Models are quite safe when they're useless, actually.

In the times of GPT-3 I'd scoff at the idea of an LLM doing any hacking; today, I'm running several AIs on my code before publishing, and they are finding (and demonstrating!) RCEs on my localhost server.

For example, one found a missing check in a third party JWT library which allowed full account takeover, which I'd have never even looked at.

Hence I don't believe a single word coming out of these people's mouths. Their "beliefs" are just marketing.

Re: GLM-5.3 is now open-weight

#126
post #119

I previously posted that DS4Flash was _good_ but not _great_ on two DGX Sparks, but I have to say that GLM-5.3 is pretty amazing. It's been able to tackle all the random hard problems I've thrown at it and it has the intuition that DS4Flash seems to lack. We're nowhere near a Fable-class model IMO, but things are going to get interesting in this next year.

I assume that's about 5.3 Flash, not full?

Yes, sorry 5.3 flash.

Re: GLM-5.3 is now open-weight

#127
post #91

Does this mean it'll be on Bedrock soon? I hear great things about this model but I want AWS data handling practices...

I doubt it - AWS hasn't added any non-western models since GLM 5 and MiniMax M2.5 in February, afaik. Might be a deal with OpenAI (GPT 5.4 was the first to be available via Bedrock, in April) or might just be that there isn't a lot of demand due to corporate skittishness around models trained in China.

Re: GLM-5.3 is now open-weight

#129

Earlier quoted context omitted.

When we consider: * LLM usage is new for the world * Models are evolving quickly with high worldwide competition * Hardware is evolving despite RAM shortages Is investing a huge sum of money in equipment for local inference a wise use of money? Or are M5 Ultra and equivalently priced local inference hardware future-proof enough to be worth it relative to how the market is evolving? Maybe it’s all a question of what y…

Tools vs services in my mind. There is no guarantee any provider will continue to do what they are doing for you at the price they are doing it. The object permanence of not having to reinvent the world every time a model gets sunsetted has value.

> Tools vs services in my mind. There is no guarantee any provider will continue to do what they are doing for you at the price they are doing it.

with open models, there is ecosystem/market of providers, where you can easily switch to provider you like

Re: GLM-5.3 is now open-weight

#130

Earlier quoted context omitted.

When we consider: * LLM usage is new for the world * Models are evolving quickly with high worldwide competition * Hardware is evolving despite RAM shortages Is investing a huge sum of money in equipment for local inference a wise use of money? Or are M5 Ultra and equivalently priced local inference hardware future-proof enough to be worth it relative to how the market is evolving? Maybe it’s all a question of what y…

Tools vs services in my mind. There is no guarantee any provider will continue to do what they are doing for you at the price they are doing it. The object permanence of not having to reinvent the world every time a model gets sunsetted has value.

Do You have guarante any electricity price?
Post reply on HN