Live data from Hacker News

GLM-5.3-Flash

z.ai

301–310 of 605 posts

Re: GLM-5.3-Flash

#301

Earlier quoted context omitted.

I bought 4x Asus GX10 with the 1TB option. I don't understand why, but it's the only model in the whole lineup that isn't priced insanely. They were briefly on sale with a $200-off coupon, but they show up on warehouse deals from time-to-time as well.

> it's the only model in the whole lineup that isn't priced insanely $4,000 isn't priced insanely? ye gads

They only went up from 3000€ to 4000€ which isn't a lot.

For comparison the cheapest Strix Halo 128GB went from 1600€ to 2600€ in the same timeframe.

Re: GLM-5.3-Flash

#302
post #232
post #228

Earlier quoted context omitted.

I've had the exact opposite experience. I've been using 3.8 for my daily driver since last week, and I've gradually been giving it more and more complex tasks as it continues to deliver high quality results. Now I am basically handing off large complex features, and 3.8 is doing the planning, task breakdown, implementation and review with just a few notes from my side. The tradeoff is time (especially on RDMA4 hardwa…

Fascinating that after years of for me this for me that, literally no one writes 2-3 other words like “i do react frontend” or whatever just for us to know why the results are different

Your assumption is incorrect.

Re: GLM-5.3-Flash

#304

> 320B total parameters and just 18B active parameters This is pretty hefty for a "flash" model, even a 256 GB setup is insufficient at q4 - and q4 is already the worst-but-still-acceptable quant in my experience. The benchmarks look great, especially since GLM tends to be more honest than the average Chinese lab, but you’ll need to splurge to run it at home. @edit: so many releases that I forgot to math. This fits j…

Speaking as someone who isn't really well versed in this, does 18B active parameters mean that you could potentially hold only the 18B parameters in RAM and stream the rest from a fast NVMe SSD for acceptable performance similar to how Colibri works?

https://github.com/JustVugg/colibri

Re: GLM-5.3-Flash

#305
GLM 5.3 Flash: 320B parameters with 18B activated

Qwen 3.8 Next Flash: 125B + 51B = 176B parameters with 6B activated

DeepSeek V4 Flash: 284B with 13B activated

The new Qwen model is the most promising for one or two Strix Halo 128GB with the low number of active parameters. On paper it's much stronger than Qwen 3.8 27B.

Re: GLM-5.3-Flash

#307

Earlier quoted context omitted.

I don't understand how people don't consider this. Plus you're spec'd out of near-SOTA level in months. The only reasons to actually do this are a) you have a lot of dispensable income and are a hobbyist/tinkerer, b) you have real, legitimate privacy concerns or, relatedly, c) you're doing something you don't want to get flagged

"you don't want to get flagged" Ding!

What is actually getting you flagged by the openweights inference providers? Thus far I haven't hit any of the reverse engineering or infosec guardrails that Anthropic is so keen on

Re: GLM-5.3-Flash

#308

Earlier quoted context omitted.

I really wish GLM models had vision capabilities. I've worked around that in the past to use a vision MCP in my harness that GLM can call. It is not the same, but it allows the model to query images.

Well, now one of them does!

That's wonderful. I was going off an older version of the Artificial Analysis page for GLM-5.3-Flash https://artificialanalysis.ai/models/glm-5-3-flash. The page is updated now to show that it does support multi-modal image inputs.

Re: GLM-5.3-Flash

#309
post #31

Earlier quoted context omitted.

Looks like the M5 Ultra Studio wait times are going to increase again. Already at 10-12 weeks, I wonder how long it'll go?

I guess like the M3 Ultra, at some point normal customers won’t be able to buy it.

That M3 had an older type of RAM. Apple hopefully secured sufficient supply of the newer variant for the M5 Ultra.

Re: GLM-5.3-Flash

#310
post #306

Good bicycle, good pelican: https://tools.simonwillison.net/markdown-svg-renderer#url=ht...

I only see a mostly blank page with a "Paste" button, a "URL" button, and and a "Preview" label.
Post reply on HN