Live data from Hacker News

GLM-5.3 is now open-weight

huggingface.co

21–30 of 296 posts

Re: GLM-5.3 is now open-weight

#22
post #14

Earlier quoted context omitted.

I've been using it quite a bit too. My main complaint is that it can be really slow sometimes — like, really slow — and the speed feels pretty inconsistent.

z.ai is using all Chinese hardware for flash: https://thenewstack.io/glm-5-3-flash-chinese-chips/ There are other providers with much faster inference, like BaseTen at >100t/s: https://openrouter.ai/z-ai/glm-5.3-flash#performance

How do I find out where the openrouter model providers' servers are located?

Re: GLM-5.3 is now open-weight

#23
post #8
post #5

GLM-5.3-Flash is actually cheaper than deepseek and better than deepseek but no one is talking about yet :)

It's actually slightly more expensive ($0.50 vs $0.48), but there's a temporary 50% discount. I've seen dozens of conversations about it in last 24 hours, and every major inference provided added in first 24 hours. I think it's gaining plenty of traction.

It's interesting that OpenCode Go is treating it as 2x more expensive than DeepSeek Flash, even factoring in the 50% discount

Re: GLM-5.3 is now open-weight

#24
How much usage do you find you get on these kinda models (I know the pricing changes a bit) compared to a $20 sub say for Google AI Pro in anti gravity?

I hate how difficult it is to compare prices when looking at subscriptions.

Would $20 in open router, using models like GLM get me more or less?

Re: GLM-5.3 is now open-weight

#25
I'd like to ask Sam Altman if he still thinks that it's too dangerous to publish GPT-3. I mean, no one would use it, but what is his reasoning for not publishing it now, in 2026?

Re: GLM-5.3 is now open-weight

#27
post #2

I've been using it more and more. Feels like Opus 4.8, in the best possible way.

I'm starting to think Opus 4.8 is significantly smaller than most people assume. If it's significantly larger than GLM 5.3 (I've heard some insane guesstimates out there like upwards of 5T params or more), that would prove rather embarrassing for Anthropic.

> that would prove rather embarrassing for Anthropic

Not really, in that you just work with different constraints.

Anthropic and US labs in general has maybe 100s to 1000s of GPUs per person to experiment. Zai and Chinese labs in general have 1-10.

The priorities are different.

Re: GLM-5.3 is now open-weight

#28
post #8

Earlier quoted context omitted.

It's actually slightly more expensive ($0.50 vs $0.48), but there's a temporary 50% discount. I've seen dozens of conversations about it in last 24 hours, and every major inference provided added in first 24 hours. I think it's gaining plenty of traction.

It's interesting that OpenCode Go is treating it as 2x more expensive than DeepSeek Flash, even factoring in the 50% discount

Go has API pricing + this weird scaling of how much is it worth. Some models get $60 of usage, some $30 and some $15 etc.

Re: GLM-5.3 is now open-weight

#29
post #11

h/t to DeepInfra for being the first 3rd party provider for it on OpenRouter ( https://openrouter.ai/z-ai/glm-5.3?endpoint=b711bea7-3994-49... ).

on my TrustedRouter:

z-ai/glm-5.3: also Z.ai, Novita, Atlas Cloud, IO.NET

Re: GLM-5.3 is now open-weight

#30

Earlier quoted context omitted.

I'm starting to think Opus 4.8 is significantly smaller than most people assume. If it's significantly larger than GLM 5.3 (I've heard some insane guesstimates out there like upwards of 5T params or more), that would prove rather embarrassing for Anthropic.

I hear the argument here, but isn't it possible it has dramatically more knowledge and when you get outside the common cases many of us use it for, it'll have completely different capabilities? I feel like most benchmarks cluster on a reasonably limited area of human knowledge

Sort of depends on how well the core reasoning works. It’s not a big effort to connect an LLM to a search provider.

You do pay for the tokens, but in theory on a smaller model each token is cheaper.

Post reply on HN