Live data from Hacker News

GLM-5.3 is now open-weight

huggingface.co

41–50 of 298 posts

Re: GLM-5.3 is now open-weight

#42
just tested (zai-org/GLM-5.3-Flash via together.ai) against latest DeepSeek-V4-Flash for a very specific task and thought i'd report here...

- price: DS4 wins... $0.0235 vs $0.0242 for ten tasks

- latency: GLM wins... 108s total against 154s

this is for a personal use-case where i'm detecting ads in a written transcript. sticking with ds4-flash for now since latency is not a critical factor

Re: GLM-5.3 is now open-weight

#43

GLM 5.3 is probably the sweet spot open weights model if you want to go beyond deepseek flash or the new glm flash. I used it with pi and had a fairly good time, especially since it’s less touchy about cyber and whatnot than the US guys. It’s slightly behind Kimi in ability but it’s a lot easier to run it, I’d expect prices (and speed!) from third parties to be noticeably better. Assuming you’re willing to drop a fat…

I have just built an Epyc with 512gb DDR4 3200 RAM for a "reasonable" price and I'm hoping to have a setup with GLM as the architect and Qwen 27b/Next Flash as the implementer. This is 1/5 of the price of the Mac, but also probably 1/5 of the speed lol.

Re: GLM-5.3 is now open-weight

#44
post #2

I've been using it more and more. Feels like Opus 4.8, in the best possible way.

I really like how it doesn't have that Claude talk. It just does the thing without Claude's "load-bearing honesty." It's probably my favorite model to interact with, even if it isn't the best or most reliable.

Re: GLM-5.3 is now open-weight

#45
post #36
post #25

I'd like to ask Sam Altman if he still thinks that it's too dangerous to publish GPT-3. I mean, no one would use it, but what is his reasoning for not publishing it now, in 2026?

They already publish gpt-oss which is several generations better than gpt-3

At release, GPT-OSS was arguably a few generations behind the open frontier.

Re: GLM-5.3 is now open-weight

#47
post #22

Earlier quoted context omitted.

How do I find out where the openrouter model providers' servers are located?

If you click on the provider name, the panel that pops up shows a "Region" value. Not every provider lists their region, however.

I think the region is just the HQ of the provider. So z.ai's region is Singapore but it's quite likely that their servers are actually in China

Re: GLM-5.3 is now open-weight

#48
post #12

Earlier quoted context omitted.

I'm starting to think Opus 4.8 is significantly smaller than most people assume. If it's significantly larger than GLM 5.3 (I've heard some insane guesstimates out there like upwards of 5T params or more), that would prove rather embarrassing for Anthropic.

You can't compare models released 6+ months apart. GLM 5.2 was same architecture as 5.3 and not nearly as good. Takes time to build frontier intelligence and distill down to smaller sizes.

[deleted]

Re: GLM-5.3 is now open-weight

#50

GLM 5.3 is probably the sweet spot open weights model if you want to go beyond deepseek flash or the new glm flash. I used it with pi and had a fairly good time, especially since it’s less touchy about cyber and whatnot than the US guys. It’s slightly behind Kimi in ability but it’s a lot easier to run it, I’d expect prices (and speed!) from third parties to be noticeably better. Assuming you’re willing to drop a fat…

One could also run it locally on a used dual xeon (or amd-equivalent) server with 512GB RAM, albeit slower, if you have a useful workflow for it that's like "take this day's efforts and run it through various analysis agents", combined with giving it one-shot tasks/modules to build overnight. You would want a place like a garage or basement to put the server because it'll be loud.

You'd also likely spend far more in electricity than the API cost of processing the prompt(s)
Post reply on HN