Live data from Hacker News

GLM-5.3-Flash

z.ai

71–80 of 605 posts

Re: GLM-5.3-Flash

#71
post #53

When reading this type of announcements, always have keen eyes on graphs. e.g. "Agent Coding Performance by Effort Level" cuts Y-axis from 0~20. - This makes it as if GLM-5.3-Flash made a bigger jump than it claimed as the Y-axis does not increase much (stupid trick used in biz reports) I did mention that ox was working ok for me, and having an open-weight comparable to close to SOTA makes it very compelling for me t…

they also conspicuously omitted GPT 5.6 Luna from comparison. It scores lower, but is also cheaper. MiMo 2.5 is not a valid comp at this point edit: nevermind. it is there in the artifical analysis scatter plot, but is greyed-out. MUCH more interesting is that in that chart, their cost is WAY off. The actual chart shows GLM 5.3 Flash at $0.09, but their chart shows $0.045...

The web page says 5.3 flash is discounted right now.

Re: GLM-5.3-Flash

#72
post #46

Earlier quoted context omitted.

I thought 4000 in sum. No wait, 4000 per , plus tax. Or EUR pricing to similar accord. Ouch.

Yeah, that little cluster costs about the same as a brand-new Dacia Sandero.

Yeah, yeah. BUT, will the Sandero be ... load-bearing? :)

Re: GLM-5.3-Flash

#74
post #8

Earlier quoted context omitted.

Is that cheaper than DS4 flash?

Slightly more expensive than the (post-price hike) DS4 flash pricing, but in the ballpark. https://openrouter.ai/compare/deepseek/deepseek-v4-flash-073...

Comparison should be to 0731

Re: GLM-5.3-Flash

#75
post #7

Weights on HF here: https://huggingface.co/zai-org/GLM-5.3-Flash I decided to take the plunge and get myself four sparks at a decent price (and bought the QSFP cables from AliExpress because they are literally 1/2 the price of Amazon), even knowing Apple was going to release new hardware and there's probably a spark 2 on the horizon. It looks like this is going to be a decent fit for what I need. I've been experiment…

If you used the bare API pricing, 1M tokens @ 30% input/70% output/50% cached, you'd pay $0.05805. Even with four discounted sparks, how much are you paying for the same tokens/distribution?

There's soooo much by way of experiments, explorations, tinkering, and even projects that you can't possibly pursue through a some SaaS API.

The more reasonable comparison is against rented GPU's, while looking at tradeoffs in latency and upload/download/storage/instance management overhead.

Buying hardware for local models is meeting a wholly different need than buying tokens through OpenRouter or whatever.

Re: GLM-5.3-Flash

#76
Why is their own coding plan always the last place z.ai release their models? Its even online, you just have to guess the model settings.

Re: GLM-5.3-Flash

#78

I'm starting to think that this whole sanctioning China may motivate and prompt them to do more and better in every field. It's too big, bright and resourceful of a country to choose confrontation instead of collaboration.

Well the big problem with china is that they do not respect international law when it comes to technology theft. But that argument is very weak when it appears that a lot of what they do is out in the open for anyone to replicate.

yeah, America is totally out there respecting international law.

"problem" indeed.

Re: GLM-5.3-Flash

#79
Chinese labs are so used to manipulating benchmarks to try to flatter inferior models that when they finally have one that's really pretty good I think the official announcement here undersells it.

https://deepswe.datacurve.ai/

That's pretty solid. Smarter and cheaper than Luna xhigh, not as smart but less expensive than Luna max. Smashes deepseek v4 flash, and even worse it matches v4 pro at a tiny fraction the cost. Roughly equivalent to sol medium, at a fraction the cost.

They should've just lead with real, up to date data, because it's good, not the silly old tactics like comparing to Opus 4.8 when 5.0 is out in many of their charts.

Congrats to them!

Re: GLM-5.3-Flash

#80
post #53

Earlier quoted context omitted.

they also conspicuously omitted GPT 5.6 Luna from comparison. It scores lower, but is also cheaper. MiMo 2.5 is not a valid comp at this point edit: nevermind. it is there in the artifical analysis scatter plot, but is greyed-out. MUCH more interesting is that in that chart, their cost is WAY off. The actual chart shows GLM 5.3 Flash at $0.09, but their chart shows $0.045...

The web page says 5.3 flash is discounted right now.

Seems disingenuous to draw frontier graphs with starter pricing.
Post reply on HN