Live data from Hacker News

GLM 5.2 Is Out

twitter.com

121–130 of 544 posts

Re: GLM 5.2 Is Out

#121

Curious what people's experience is with these models. Anecdotally I tried these out earlier in the year and found it struggled with pretty basic full-stack coding I was doing, when Sonnet 4.6 and Haiku 4.5 didn't break a sweat. Was hoping to use it while my Claude usage was resetting but was disappointed.

I've been using GLM-5/5.1 for about 6 months and it has been a fairly capable model. I've seen a lot of mixed opinions that tend to align with harness usage so it is worth trying out a couple with a model before writing it off. For example, I'm using crush and have had a good experience while others using CC have had a much more mixed experience. For task complexity, I treat it as I would sonnet with the same care in building out plans/prompts before firing it off and letting it go.

I use intelliJ for much of my development and also set the built in AI tools to use my GLM sub (BYOK) and it has worked out well albeit a bit slow.

Overarll, it's my main model and has been getting better with each release.

Re: GLM 5.2 Is Out

#122
post #88

Earlier quoted context omitted.

Anthropic blocks Fable from answering "Tell me about Agent Orange" or even "Tell me about mitochondria"

But you can see the CBRN weapon nexus in your examples that's missing from the Tiananmen prompt, right? Do American models refuse to tell you about COINTELPRO, Kent State, or My Lai, for instance?

Well, one did suddenly develop the need to tell users continuously about apparent white genocide in South Africa.

Re: GLM 5.2 Is Out

#123

Released at the exact same time, 5:21 pm (Chinese time), as when Anthropic received the letter from the government banning Fable, and explicitly citing other models becoming unusable.

... really? are you sure about the timezones? That's kind of odd, isn't it? Maybe the post was edited afterwards?

It's not odd, they explicitly reference the time and the closing of Fable. See the submitted post now as the link has changed to the founder's announcement.

Re: GLM 5.2 Is Out

#124
post #88

Earlier quoted context omitted.

Anthropic blocks Fable from answering "Tell me about Agent Orange" or even "Tell me about mitochondria"

But you can see the CBRN weapon nexus in your examples that's missing from the Tiananmen prompt, right? Do American models refuse to tell you about COINTELPRO, Kent State, or My Lai, for instance?

American models are restricted from telling you inconvenient truths just as much, you just erroneously assume to know what those truths are in the first place.

Which is of course circular thinking: why would they restrict things you already know about? Why would they do it in such a clumsy and obvious way?

Look at MKULTRA, you know next to nothing about it and much less do you know what they do in that direction now.

For a current psyops, look at www.war.gov/UFO/ and marvel at how they tell you nothing, reinforcing your false belief to already know everything.

There is much more and you know much less about it.

Re: GLM 5.2 Is Out

#125
post #118
post #79

Is there any indication of what compute resources this will actually require (in its various incarnations)? Does it incorporate any of the optimisations pioneered by Google (such as TurboQuant, MTP) or some other original innovations to make the frontier quality realistically available to local users?

The GLM-5 series is 744B-A40B. This is not a local model for any reasonable definition of local, but it's an open model which means (once they upload the weights in a week or so) there will be a dozen third-party inference providers competing on price per token.

> This is not a local model for any reasonable definition of local

That's true for now. I am hopeful that once the hardware markets have recovered from OpenAI's sabotage, we will see more hardware dedicated to local inference that can handle these big models.

Also, I'm thinking about the unique MoE routing that Apple is using with their new Apple Foundation Model. The model is trained and architected so that experts are not swapped for every token, but only occasionally. This suggests that e.g., a 744B parameter model in the future could have experts offloaded to SSD and still run with the effective computing requirements of a 40B model.

Re: GLM 5.2 Is Out

#126

It would be so extremely awesome if this ai would have been a Claude killer alternative and 90% of Europe cancels Claude subscriptions and subscribe on this one. It would be the dumbest move of the year by the US.

I'm actually interested in doing that.

What would be the most favorable model/company to move to for scientific programming and engineering questions?

Re: GLM 5.2 Is Out

#127

Crossing fingers for a 5.2 flash release - it’s been a while but I still feel like 4.7 flash is one of the strongest local coding models

Really? I had a terrible experience with 4.7-flash. Qwen-3.5 is still the best local model for me. (3.6 pushed VRAM usage just out of 24GB and then you're not using a consumer GPU any more)

Re: GLM 5.2 Is Out

#129
post #83

Earlier quoted context omitted.

[flagged]

The good news is if there are multiple frontier AI models from multiple countries with non overlapping sets of restricted answers, we can just use a couple of them to get open answers.

Not really non-overlapping though: both refuse to talk much about certain widely common activity between people (or even by yourself). That activity has shaped humanity quite a bit throughout its entire history. It's hard to imagine AI can understand humans fully if everything about it is excluded from the training data.

Re: GLM 5.2 Is Out

#130
post #88

Earlier quoted context omitted.

Anthropic blocks Fable from answering "Tell me about Agent Orange" or even "Tell me about mitochondria"

But you can see the CBRN weapon nexus in your examples that's missing from the Tiananmen prompt, right? Do American models refuse to tell you about COINTELPRO, Kent State, or My Lai, for instance?

try to ask even grok about some stuff happenning right now in middle east or related to epstein files - its more and more censored and only sometimes will answer if you ask know what detailed question to ask. One year ago grok wasn't that bad and its supposed to be the less censored.
Post reply on HN