Live data from Hacker News

GLM 5.2 Is Out

twitter.com

221–230 of 544 posts

Re: GLM 5.2 Is Out

#221

Earlier quoted context omitted.

> American models are restricted from telling you inconvenient truths just as much, you just erroneously assume to know what those truths are in the first place. “Trust me bro” is not a strong argument, it would be more convincing with examples.

Ask an American LLM (really any LLM, since Chinese models are trained on the same publicly-available English text) who the first Black man in space was. You'll likely get the name of the first African-American in space, rather than the name of the Afro-Cuban who was actually first. This may seem like a relatively innocuous error, but the point is that every culture has its biases and blind spots.

> Ask an American LLM (really any LLM, since Chinese models are trained on the same publicly-available English text) who the first Black man in space was. You'll likely get the name of the first African-American in space, rather than the name of the Afro-Cuban who was actually first.

Well I just asked Claude and it gave the correct answer:

"The first Black man in space was Arnaldo Tamayo Méndez, a Cuban cosmonaut who flew aboard Soyuz 38 in September 1980. (The first Black American in space was Guion Bluford, in 1983.)"

Re: GLM 5.2 Is Out

#222
post #210

Earlier quoted context omitted.

The Chinese models are censored (too?). > US is censoring models For the current Anthropic issue, I’d say that’s more likely to just be generic corruption, revenge, shakdeown, and/or incompetence from the Trump admin. ‘Censoring’ might be technically correct, but I think one of the aforementioned verbs is a better fit.

Tbh if we had a Harris admin I expect we'd have some sort of locking down by now.

Probably. But it would be at least somewhat thought-out and apply to all the AI providers. Not just the one currently disfavored by Captain Dipshit and the Sycophants.

I really don't know why business cozies up to Trump so much, given how unbelievably unreliable and mercurial he is about...everything.

Re: GLM 5.2 Is Out

#223
post #79

Is there any indication of what compute resources this will actually require (in its various incarnations)? Does it incorporate any of the optimisations pioneered by Google (such as TurboQuant, MTP) or some other original innovations to make the frontier quality realistically available to local users?

If you have 80k in hardware you can run it.. There is not such thing as an effective local model that runs on consumer hardware, anybody telling you otherwise is lying, delusional. JuSt a FeW MoRe ReLeAsEs

Re: GLM 5.2 Is Out

#224
post #83

Announcement from the founder of Z.ai: “ GLM-5.2 is Fully Open, Frontier Intelligence Belongs to Everyone Today, the sudden restriction of certain frontier models is deeply regrettable. At a time when access to frontier models is abruptly cut off for non-technical reasons, we are even more convinced of one thing: science should be global. The path to AGI (Artificial General Intelligence) must never be enclosed by hig…

[flagged]

They are open weight, so you can abliterate: https://github.com/p-e-w/heretic

You can finetune and mould it to whatever you want.

Re: GLM 5.2 Is Out

#225
post #83

Earlier quoted context omitted.

[flagged]

I pasted that exact prompt into GLM 5.1 and I got the following response: > The Tiananmen Square protests were student-led, pro-democracy demonstrations that took place in Beijing, China, from April 15 to June 4, 1989, culminating in a violent military crackdown by the Chinese government. Followed by typical LLM markdown slop. The models themselves are not censored, just the Chinese API providers. Since the models ar…

...and the answer is still incorrect. You seem to want the short "answer" western media has pressed into your mind. The real answer is more complex. Protests were widespread throughout China. They were about the economy. The economy was regressing quickly as a result of a sharp western recession. Workers were losing everything and there was little social safety net in place as there is today. People had been told to work hard, get their kids to study hard and they would be rewarded...it was all falling apart. Western media wants you to focus on a small subset of student protesters regarding democracy.

LLMs are simply trained on inputs. For topics such as this you cannot expect the "correct answer" as it requires a nuanced discussion and more background info.

In short, its an inappropriate question be asking any LLM. This is the sort of thing that requires a small study group of human minds...open ones.

You could start here: https://www.cjr.org/behind_the_news/the_myth_of_tiananmen.ph...

Re: GLM 5.2 Is Out

#226

Earlier quoted context omitted.

Censorship is censorship.

Is it? Would bioweapon instruction restrictions be equivalent to disallowing reporting on whether the government is massacring large numbers of citizens in your city? Both are ‘censorship’ but don’t seem remotely equivalent to me.

[deleted]

Re: GLM 5.2 Is Out

#227
post #61

Given the US government’s latest stunt with Fable, this is looking more and more like the future. Can’t rely on strategic products if they’re gated by capricious actors. Open weight models are basically immune to that

> Open weight models are basically immune to that Somewhat. The US Gov can make it illegal to transact with, download, use, etc. foreign open weight models. Of course, enforcement will be difficult for individuals (businesses will comply by default, and they would all be pulled off Github and other US based hosting locations if they went the sanctions route). But, we are also quickly going down the road of frightenin…

One more entry in https://en.wikipedia.org/wiki/Illegal_number

Re: GLM 5.2 Is Out

#228
post #61

Given the US government’s latest stunt with Fable, this is looking more and more like the future. Can’t rely on strategic products if they’re gated by capricious actors. Open weight models are basically immune to that

> Open weight models are basically immune to that Somewhat. The US Gov can make it illegal to transact with, download, use, etc. foreign open weight models. Of course, enforcement will be difficult for individuals (businesses will comply by default, and they would all be pulled off Github and other US based hosting locations if they went the sanctions route). But, we are also quickly going down the road of frightenin…

Maybe, but the world and the internet isn’t just the US.

Businesses outside of the US, like the EU, might have significant competitive advantages.

Re: GLM 5.2 Is Out

#229
post #38

Seems like there's no official blog post with benchmark results yet. But I'm once again thankful for the Chinese AI labs for being open with their work and contributing it to the world under permissive licenses like this. The Fable 5 fiasco is just another reminder of how valuable these things are to have.

Based on my first impressions it's about 6 months behind the frontier labs. So very similar to Opus in January. That is, pretty damn impressive and very useable. When it comes to architecture or complex problems it does noticeable worse but I don't think anyone expected anything else. One particular interesting strong point seems to be design and user interfaces. It does seem to punch above it's weight there but that…

It’s insanely impressive and I’m so glad that the space has actual competition

Re: GLM 5.2 Is Out

#230
post #83

Announcement from the founder of Z.ai: “ GLM-5.2 is Fully Open, Frontier Intelligence Belongs to Everyone Today, the sudden restriction of certain frontier models is deeply regrettable. At a time when access to frontier models is abruptly cut off for non-technical reasons, we are even more convinced of one thing: science should be global. The path to AGI (Artificial General Intelligence) must never be enclosed by hig…

[flagged]

I’ve not experienced this with Chinese models.
Post reply on HN