Live data from Hacker News

GLM-5.3-Flash

z.ai

41–50 of 605 posts

Re: GLM-5.3-Flash

#43

Earlier quoted context omitted.

> get myself four sparks at a decent price Wow, if you don't mind me asking. How and where?

I bought 4x Asus GX10 with the 1TB option. I don't understand why, but it's the only model in the whole lineup that isn't priced insanely. They were briefly on sale with a $200-off coupon, but they show up on warehouse deals from time-to-time as well.

> it's the only model in the whole lineup that isn't priced insanely

$4,000 isn't priced insanely? ye gads

Re: GLM-5.3-Flash

#44
post #31

> 320B total parameters and just 18B active parameters This is pretty hefty for a "flash" model, even a 256 GB setup is insufficient at q4 - and q4 is already the worst-but-still-acceptable quant in my experience. The benchmarks look great, especially since GLM tends to be more honest than the average Chinese lab, but you’ll need to splurge to run it at home. @edit: so many releases that I forgot to math. This fits j…

Looks like the M5 Ultra Studio wait times are going to increase again. Already at 10-12 weeks, I wonder how long it'll go?

I guess like the M3 Ultra, at some point normal customers won’t be able to buy it.

Re: GLM-5.3-Flash

#45

from the article, pareto frontier for open source models is completely dominated by GLM now.

Well, it will be interesting to see where Qwen3.8-Flash-Next ends up landing, also released today. These are exciting times!

Re: GLM-5.3-Flash

#46

Earlier quoted context omitted.

I bought 4x Asus GX10 with the 1TB option. I don't understand why, but it's the only model in the whole lineup that isn't priced insanely. They were briefly on sale with a $200-off coupon, but they show up on warehouse deals from time-to-time as well.

> it's the only model in the whole lineup that isn't priced insanely $4,000 isn't priced insanely? ye gads

I thought 4000 in sum. No wait, 4000 per, plus tax. Or EUR pricing to similar accord. Ouch.

Re: GLM-5.3-Flash

#47
When reading this type of announcements, always have keen eyes on graphs.

e.g. "Agent Coding Performance by Effort Level" cuts Y-axis from 0~20.

- This makes it as if GLM-5.3-Flash made a bigger jump than it claimed as the Y-axis does not increase much (stupid trick used in biz reports)

I did mention that ox was working ok for me, and having an open-weight comparable to close to SOTA makes it very compelling for me to try it out locally (well, only if I got more VRAM)

Re: GLM-5.3-Flash

#48

Earlier quoted context omitted.

> get myself four sparks at a decent price Wow, if you don't mind me asking. How and where?

I bought 4x Asus GX10 with the 1TB option. I don't understand why, but it's the only model in the whole lineup that isn't priced insanely. They were briefly on sale with a $200-off coupon, but they show up on warehouse deals from time-to-time as well.

~$4000 USD each on Amazon, $175 for the cable.

Re: GLM-5.3-Flash

#49
post #8
post #3

Standard API Pricing for GLM-5.3-Flash (per 1M tokens) - Input: $0.15 - Output: $0.50 - Cached input: $0.03

Is that cheaper than DS4 flash?

It's even cheaper than DS4's off-peak pricing. Seems like DeepSeek have some stiff competition now

Re: GLM-5.3-Flash

#50
post #38
post #5

> with all of this traffic served on Chinese AI chips RIP Nivida shareholders

Another self-inflicted own courtesy of US government policy. While I think China would always get to hardware self-sufficiency eventually, all export controls have done is (1) accelerate China's development, and (2) divert revenue that would've otherwise gone to NVIDIA/AMD/etc instead.

The export controls were revoked before it triggered Chinese protectionism: https://www.silicon.co.uk/e-innovation/artificial-intelligen... / https://archive.vn/B2pah
Post reply on HN