Live data from Hacker News

GLM-5.3 is now open-weight

huggingface.co

291–296 of 296 posts

Re: GLM-5.3 is now open-weight

#291
post #25

I'd like to ask Sam Altman if he still thinks that it's too dangerous to publish GPT-3. I mean, no one would use it, but what is his reasoning for not publishing it now, in 2026?

I think it would be an important historical document as well. We are potentially looking at the dawn of AGI and one of the most important models ever created. Each model is also a kind of ultimate time capsule, containing a snapshot of the entire human collective mind. If you wanted to ask a 2002 person what they thought about future historical events you can just ask them directly.

> containing a snapshot of the entire human collective mind.

I say this with kindness: Anyone who believes this absolutely needs to turn off their computer for the week, go outside, travel a bit, and experience reality with other humans outside their regular bubble.

The “entire human collective mind” is not digital. It’s not on the internet. These models could’ve syphoned literally every piece of digital media in existence and still wouldn’t have it. People don’t exist inside computers, and it is naive to believe the sum of what’s online makes the sum of the human experience. It doesn’t.

Re: GLM-5.3 is now open-weight

#293
post #284
post #275

Earlier quoted context omitted.

Could you give me an example of a similar task you asked Fable and a different model to do where Fable did a better job? I have a hard time getting models like GLM 5.3 to not perform on my tasks but I might be biased.

Changes in a complex codebase. Opus can do it, but needs more handholding. Making the plan with fable and let opus implement it worked out well.

That's actually not quite what I was asking for.

Re: GLM-5.3 is now open-weight

#294

Earlier quoted context omitted.

You do, there's like 20 providers for any model on openrouter. You can also just spin bedrock or gcp and download the weights for later if you're worried. It's never going to make cost sense when the token rate is so low with how expensive ram is

What if the internet goes away?

Starlink? It's never gone anymore

Re: GLM-5.3 is now open-weight

#295

Earlier quoted context omitted.

I think that's overly pessimistic. Here's [1] a video of somebody running it on a ~$6000 rig and getting around 14T/s for complex prompts (about double that for simpler prompts). Payback time is going to depend on your electric cost/consumption. In most domains cloud providers end up charging a significant premium rather than a offering a scale enabled discount, relative to local at retail costs. That will almost cer…

With roughly 2.7 million seconds per month, times 14 tokens per second, you are getting 38.5 million tokens a month at most. That’s less than 164USD worth of GLM5.3 tokens on the inference market. So that 6000 USD rig will take 3 years to break even - and only if it runs continuously. And this is being generous, as it’s not even taking quantisation into account.

> That’s less than 164USD worth of GLM5.3 tokens on the inference market.

I can cherry pick stats too.

The other day I heard mention of someone paying $200/mo for Claude Code.

At those rates my local LM setup pays for itself in a single year.

Post reply on HN