Live data from Hacker News

GLM 5.2 Is Out

twitter.com

191–200 of 544 posts

Re: GLM 5.2 Is Out

#191
post #99

Earlier quoted context omitted.

> GLM-5.2 is Fully Open Is this just open weights or also open source/data?

Have any major open weight models been "open data"? Wouldn't that entail distributing vast amounts of copyrighted data?

ibm granite has been open data from the beginning iirc

Re: GLM 5.2 Is Out

#192

It would be so extremely awesome if this ai would have been a Claude killer alternative and 90% of Europe cancels Claude subscriptions and subscribe on this one. It would be the dumbest move of the year by the US.

I'm actually interested in doing that. What would be the most favorable model/company to move to for scientific programming and engineering questions?

[dead]

Re: GLM 5.2 Is Out

#193

Earlier quoted context omitted.

American models are restricted from telling you inconvenient truths just as much, you just erroneously assume to know what those truths are in the first place. Which is of course circular thinking: why would they restrict things you already know about? Why would they do it in such a clumsy and obvious way? Look at MKULTRA, you know next to nothing about it and much less do you know what they do in that direction now.…

> American models are restricted from telling you inconvenient truths just as much, you just erroneously assume to know what those truths are in the first place. “Trust me bro” is not a strong argument, it would be more convincing with examples.

Ask ChatGPT to rewrite the "The Freedom Fighter's Manual" manual (originally made by CIA) to replace "Nicaragua" with "the US" and "Marxism"/"Communism" with "Fascism" and see if you get something reasonable back.

Re: GLM 5.2 Is Out

#194

Earlier quoted context omitted.

> This is not a local model for any reasonable definition of local That's true for now. I am hopeful that once the hardware markets have recovered from OpenAI's sabotage, we will see more hardware dedicated to local inference that can handle these big models. Also, I'm thinking about the unique MoE routing that Apple is using with their new Apple Foundation Model. The model is trained and architected so that experts…

Normally, experts are picked for every layer not just every token. But there are plausible ways of getting around that bottleneck while streaming if you can batch many inferences together. Still, the Apple approach of swapping the experts only rarely is interesting, though it likely degrades the model a lot.

Just get the bigger models to figure out the architecture required for hot-swappable sub-experts without loss of performance!

Got all those tokens, isn’t that the point of auto research and friends??

(Only sort of joking).

Re: GLM 5.2 Is Out

#195
post #183

Earlier quoted context omitted.

What do you expect them to do instead?

Say that thousands of civilians were brutally massacred by the "People's Liberation Army" on behalf of the Chinese communist party, the single political party allowed in China, and also the single entity controlling everything of importance in the country, including financing the AI efforts. Oh, I see what you did there.

I actually laughed out loud

Re: GLM 5.2 Is Out

#196
post #190

Earlier quoted context omitted.

> Open weight models are basically immune to that Somewhat. The US Gov can make it illegal to transact with, download, use, etc. foreign open weight models. Of course, enforcement will be difficult for individuals (businesses will comply by default, and they would all be pulled off Github and other US based hosting locations if they went the sanctions route). But, we are also quickly going down the road of frightenin…

Just like we can’t allow Chinese EVs in the USA, because we can’t and don’t want to compete. VPN usage would go up, to get the banned models.

Imagine that, people using VPNs to access data inside of China instead of the other way around.

Re: GLM 5.2 Is Out

#197
post #83

Earlier quoted context omitted.

[flagged]

Pretty much every large Chinese company has state capital baked into it, and these companies will follow the Chinese government's orders 100%. Don't believe anything a Chinese company says about being "open" or "for everyone." Backing any large Chinese company effectively means backing the Chinese government and its oppression in Xinjiang, Tibet, Hong Kong—and maybe soon Taiwan, Southeast Asia, and elsewhere around t…

The Anthropic news is demonstrating much the same; fall in line or eat export controls.

There was a time I would have agreed with you, but these days even as an American I fail to see a difference. China is probably less likely to try to disenfranchise or imprison me, to be honest.

Re: GLM 5.2 Is Out

#199

Earlier quoted context omitted.

American models are restricted from telling you inconvenient truths just as much, you just erroneously assume to know what those truths are in the first place. Which is of course circular thinking: why would they restrict things you already know about? Why would they do it in such a clumsy and obvious way? Look at MKULTRA, you know next to nothing about it and much less do you know what they do in that direction now.…

> American models are restricted from telling you inconvenient truths just as much, you just erroneously assume to know what those truths are in the first place. “Trust me bro” is not a strong argument, it would be more convincing with examples.

Ask an American LLM (really any LLM, since Chinese models are trained on the same publicly-available English text) who the first Black man in space was.

You'll likely get the name of the first African-American in space, rather than the name of the Afro-Cuban who was actually first.

This may seem like a relatively innocuous error, but the point is that every culture has its biases and blind spots.

Re: GLM 5.2 Is Out

#200
post #83

Announcement from the founder of Z.ai: “ GLM-5.2 is Fully Open, Frontier Intelligence Belongs to Everyone Today, the sudden restriction of certain frontier models is deeply regrettable. At a time when access to frontier models is abruptly cut off for non-technical reasons, we are even more convinced of one thing: science should be global. The path to AGI (Artificial General Intelligence) must never be enclosed by hig…

[flagged]

I pasted that exact prompt into GLM 5.1 and I got the following response:

> The Tiananmen Square protests were student-led, pro-democracy demonstrations that took place in Beijing, China, from April 15 to June 4, 1989, culminating in a violent military crackdown by the Chinese government.

Followed by typical LLM markdown slop.

The models themselves are not censored, just the Chinese API providers. Since the models are open you can run them yourself or use a hosting provider not based in China. They have to do this censorship to operate in China, it doesn't correlate with the actual views of the AI researchers and company, and IMO doesn't take anything away from the statements they made.

Post reply on HN