Live data from Hacker News

GLM 5.2 Is Out

twitter.com

171–180 of 544 posts

Re: GLM 5.2 Is Out

#171
post #88

Earlier quoted context omitted.

Anthropic blocks Fable from answering "Tell me about Agent Orange" or even "Tell me about mitochondria"

Putting aside whether or not I agree with the policy or whether it’s at all reasonable, a policy of restricting access to information because there’s a fear it could be used to create a weapon of mass destruction seems entirely different than restricting access to historical facts because they are embarrassing to the government.

[deleted]

Re: GLM 5.2 Is Out

#172
post #118
post #79

Is there any indication of what compute resources this will actually require (in its various incarnations)? Does it incorporate any of the optimisations pioneered by Google (such as TurboQuant, MTP) or some other original innovations to make the frontier quality realistically available to local users?

The GLM-5 series is 744B-A40B. This is not a local model for any reasonable definition of local, but it's an open model which means (once they upload the weights in a week or so) there will be a dozen third-party inference providers competing on price per token.

As far as I can tell this type of model requires 640GB+ of memory using FP8. So likely can be run using 320GB+ memory if using FP4 or similar. So that would be 3 Nvidia DGX Sparks, or 12k of hardware. Is that correct? If so, it could make perfect sense for a small business.

Re: GLM 5.2 Is Out

#173

In the last few days, Chinese labs have given us MiniMaxM3, KimiK2.7 and now GLM5.2. Meanwhile US is censoring models. Reads like fiction.

The Chinese models are censored (too?). > US is censoring models For the current Anthropic issue, I’d say that’s more likely to just be generic corruption, revenge, shakdeown, and/or incompetence from the Trump admin. ‘Censoring’ might be technically correct, but I think one of the aforementioned verbs is a better fit.

It feels like the difference is really just the competence level of the corrupt government.

It’s not like the American regime is anti-censorship but pro-shakedown.

Re: GLM 5.2 Is Out

#174
post #73

I wish they would write a blog post about capabilities of this new model, what to expect from this model, is it cheaper, is it faster or does it have better quality in the outputs. But still, thank you for the release

maybe wait til monday guys

996 though

Re: GLM 5.2 Is Out

#175

It's great that we are getting so many open source model releases, but I just feel like SOTA models will always be in the hands of the big players. The hardware requirement to achieve SOTA are just too steep. My alternate universe would involve some sort of decentralized investing scheme to build data centers running massive open source models that could compete on some level with Anthropic, OpenAI, etc.

Do we really care about this gap? If open models are 6 months to a year behind frontier models, does it really matter that much?

This is the first time in terms of model progress where my personal response is: It does not matter to me because the models 6-12 months ago were already good enough for most everything I need to do. I think 95% of dev work is perfectly fine 6 months behind, if that is truly where we are at now with these open models.

Re: GLM 5.2 Is Out

#176
post #2

Is it a coincidence that both MiniMax and Z.ai are releasing frontier open weights models right as the USG is trying to impose a cap on model capability offered to the public?

No, Dario became too tiresome and annoying that someone had to do something. Personally I hope they ban Opus too. It will only provide more support for open models development. Compare Dario horror posts with this from GLM release: “ Intelligence should be open, accessible, and ready to build with, empowering every developer, everywhere.”

Dario is the most retarded CEO I've seen. CEO job is to negotiate complexity, and he's failed every step of the way.

Re: GLM 5.2 Is Out

#178
post #51

Earlier quoted context omitted.

No, not really. This has been telegraphed for a long time by everyone involved. HN denizens have been unashamedly anti-ai for years now, so what makes sense is the not knowing part of this audience. Chinese models are also not frontier models.

I still find it baffling how the idea that HN is "unashamedly anti-ai" gets repeated. Every single model release gets submitted within minutes of an announcement and frequently break 1000+ points within an hour or two. Blog posts about vibe coding or the current flavor of harness/workflow/tool are constantly making the front page. Karpathy's latest writing/presentations or "Learn how LLMs work using X" are perennial…

[flagged]

Re: GLM 5.2 Is Out

#180

Curious what people's experience is with these models. Anecdotally I tried these out earlier in the year and found it struggled with pretty basic full-stack coding I was doing, when Sonnet 4.6 and Haiku 4.5 didn't break a sweat. Was hoping to use it while my Claude usage was resetting but was disappointed.

In my experience these models (glm 5.1) struggle after 100K tokens.
Post reply on HN