Live data from Hacker News

GLM 5.2 Is Out

twitter.com

531–540 of 544 posts

Re: GLM 5.2 Is Out

#531
Lots of misanthropic bots claiming GLM 5.2 is 6 months behind when it's on par or better than opus 4.8 (released may 28, so 2 weeks) in most benchmarks.

Re: GLM 5.2 Is Out

#532
post #509

Earlier quoted context omitted.

I run Claude Max daily, and tried letting Opus 4.8 write an ADR with known requirements. After searching through codebase, git history, etc it spat out a surface level reasonable ADR, with the customary bloated text. I started reading through it asking "Is this sentence needed?: ' '", whereby it acknowledges that no, it adds nothing and changes nothing not already served by other statements. I ask it to go through ea…

Compact the context and try again, or switch to the model with the 1 million token context. They all struggle after a hugge task like rying to make sense of a large codebase. Claude is especially poor at knowing when to compact on it's own.

It has 1M context, and it's not a huge codebase, and the context is sub 10% for a thorough task. This is an LLM issue, not a model/harness issue.

I've run copilot/gemini/pi/opencode/etc for a long time, against all major providers. Don't get me wrong, I get good productivity out of it or I wouldn't use it, but it's very different from intelligence.

Re: GLM 5.2 Is Out

#533

Earlier quoted context omitted.

Again, in the Crimea example, both sides would accuse the other of being "evil" for all of those reasons. In the majority of conflicts, "evil" is an entirely meaningless political dogwhistle.

My statements are exclusive to Tibet. I don't even know where you pulled Crimea from. While we're discussing GLM, it's attempt to "intimidate" people by accusing them of insulting the CCP if you think otherwise... let's just say that doesn't fly with the free world, at all.

Your statements are about "evil" which has no objective measure. Ethics isn't on any benchmarks, because you cannot objectively measure it.

> let's just say that doesn't fly with the free world, at all.

How come? I'm an American and I pay for GLM-5 without trouble. Which part doesn't fly?

Re: GLM 5.2 Is Out

#534

Earlier quoted context omitted.

My statements are exclusive to Tibet. I don't even know where you pulled Crimea from. While we're discussing GLM, it's attempt to "intimidate" people by accusing them of insulting the CCP if you think otherwise... let's just say that doesn't fly with the free world, at all.

Your statements are about "evil" which has no objective measure. Ethics isn't on any benchmarks, because you cannot objectively measure it. > let's just say that doesn't fly with the free world, at all. How come? I'm an American and I pay for GLM-5 without trouble. Which part doesn't fly?

I think I see the problem: you don't seem to believe in objective truth. For example: you are a human, not a banana.

Basic morals are built into humans. Some become desensitized to it over time.

Objective truth: murder is evil, torturing the innocent is evil, destroying innocent lives is evil.

Be glad you're in the free world, where you can do whatever you want. China and similar countries are not like the free world. The specific part that doesn't fly is lying to the public, and trying to intimidate them to deny the well-documented truth (see the chatgpt link I posted, which information you can manually verify). The American spirit fights for freedom and liberty.

Re: GLM 5.2 Is Out

#535

Earlier quoted context omitted.

Trump is of course the worst US administration, but at least America is still nominally a democracy. As long as free elections exist, the regime Trump represents can be voted out. The American people and press still have free speech—they can freely criticize anyone, including Trump. China is different. The CCP will rule forever, no matter how terrible the things they do. No one is allowed to criticize the government.…

Trump has made some concerning moves around freedom of speech and freedom of elections, but none of it is concrete yet. Maybe it never will be, either because the threat was overstated or because he’s just not competent enough to pull it off. China does worse on those fronts, but they do so predictably. I don’t agree with many of their goals, but you can generally rely on them pursuing those goals in a manner consist…

China is unpredictable; you never know where their red lines are. An Australian journalist got years in prison for leaking something that was about to be public anyway. The founders of Manus also had their freedom of movement restricted. Entrepreneurs who invested in China have said they can’t get their money out.

If China, or Chinese companies, end up with a monopoly-like advantage in AI, the result will be like the current rare-earth situation. China will absolutely put strict controls on AI models, and that would be much worse than depending on OpenAI or Claude.

Re: GLM 5.2 Is Out

#536
post #490

Earlier quoted context omitted.

I also refuse to use that word, and I am not a bot.

There was a whole bit in one of the Asimov stories about a politician who’s accused of being a robot. He denies it, but he’s very well behaved to the point where he’s never been recorded to break the three laws. In the end he has to punch someone on stage to prove his humanity (or did he? ;)

I loved this story. I haven't read it in a long time, but I thought that ending was great.

Personally, I think he was a bot.

Re: GLM 5.2 Is Out

#537
post #490

Earlier quoted context omitted.

There was a whole bit in one of the Asimov stories about a politician who’s accused of being a robot. He denies it, but he’s very well behaved to the point where he’s never been recorded to break the three laws. In the end he has to punch someone on stage to prove his humanity (or did he? ;)

I loved this story. I haven't read it in a long time, but I thought that ending was great. Personally, I think he was a bot.

They need to get the guy he punched to punch someone, just to be sure.

...Or is it punches all the way down? :D

Re: GLM 5.2 Is Out

#538

Earlier quoted context omitted.

Why is the text field in dataset preview table populated with pornographic labels?

Because it's a random sample of the Internet?

Could be but then it means like 98% is pornography I guess, because it’s every row, so if random a bad sign!

Re: GLM 5.2 Is Out

#539
post #438

Earlier quoted context omitted.

How do you define AGI these days?

I don't have a fully perfect definition, but I can name a couple of requirements. Ironically, both reasoning and agency are required, neither of which our "reasoning agents" possess.

> I don't have a fully perfect definition

It feels like no one has. Am I wrong? How can we even talk about something that doesn't have a definition?

Re: GLM 5.2 Is Out

#540
post #500

Earlier quoted context omitted.

I should think learning about history should lead to a desire for citizens to be able to quietly make weapons at home given the many documented cases of governments across the world mass murdering their own citizens (or foreign governments invading and genociding). What's the point of telling people the wrongs of their oppressors while simultaneously disempowering them from doing anything about it or preparing to def…

The idea that Chinese citizens could’ve prevented the Tiananmen massacre with a bunch of home printed AK-47s is silly. The government had tanks. The same applies in the US.

Yes, strawmen are silly. You don't prevent the Tiananmen massacre or fight tanks; you murder an officer (or his family) while he's out shopping.
Post reply on HN