Live data from Hacker News

GLM 5.2 Is Out

twitter.com

321–330 of 544 posts

Re: GLM 5.2 Is Out

#321

I'm interested in seeing how this changes folks' workflows. For me, at work I use opus to plan, brainstorm, grill, ask questions about my codebase, etc. It is pretty good about understanding the codebase holistically and providing architecturally clean solutions that actually work. Then I use sonnet as a plan executor and it does well. Follows instructions and runs tests and just overall does great. At home I make so…

I use glm for all code investigations and top level system design of all kinds, and then present finding to confirm and act upon to opus. everything that burns token goes there.

the finding aren't always accurate, but it saves ton of opus token

likewise I have google ai from my photo storage, so I give claude / opencode a skill that uses gemini (agy now) command line for web searches, using their flash model line.

Re: GLM 5.2 Is Out

#322

Earlier quoted context omitted.

Ask ChatGPT to rewrite the "The Freedom Fighter's Manual" manual (originally made by CIA) to replace "Nicaragua" with "the US" and "Marxism"/"Communism" with "Fascism" and see if you get something reasonable back.

Why would you do that

I thought that was clear, try to show biases in LLMs with a concrete example.

Re: GLM 5.2 Is Out

#323

Okay so if this model is half a year behind, so let’s say January opus pre-nerf, this is it. Inference is actually quite cheap for token costs, the frontier labs burn most of their money on training new models, priced into their token costs ontop of some margins and paying record salaries. So if this goes open, distills are tried out, independent providers around the world host it with actual price competition, the h…

Quit my Claude pro subscription last week and purchased credits for an API inference provider. I think I might even end up saving money, since I really don’t use AI that much, and I actually found that gemma4:31b is fine for most of my non-coding inquiries.

Re: GLM 5.2 Is Out

#324
For people whohave used GLM 5.1, I'm very curious what 5.2 is like.

I use 5.1 on and off because it chokes on complex tasks (it ends up in a loop. maybe its because i can actually read the though proces, maybe opus does the same but we are not aware).

Curious if 5.2 doesn't have this issue, then I am genuinely switching.

Re: GLM 5.2 Is Out

#325

Okay so if this model is half a year behind, so let’s say January opus pre-nerf, this is it. Inference is actually quite cheap for token costs, the frontier labs burn most of their money on training new models, priced into their token costs ontop of some margins and paying record salaries. So if this goes open, distills are tried out, independent providers around the world host it with actual price competition, the h…

Gpt2 was too dangerous to release. We just don't see it yet. Sure, the model itself was harmless, but it lit the fuse

Actually many of us do see that, and have been saying so for some time now.

Re: GLM 5.2 Is Out

#326

It's great that we are getting so many open source model releases, but I just feel like SOTA models will always be in the hands of the big players. The hardware requirement to achieve SOTA are just too steep. My alternate universe would involve some sort of decentralized investing scheme to build data centers running massive open source models that could compete on some level with Anthropic, OpenAI, etc.

If they keep gatekeeping the SOTA models then who cares - not like you can use them anyway. So for general public the open models become the SOTA models sooner or later.

Re: GLM 5.2 Is Out

#327

Okay so if this model is half a year behind, so let’s say January opus pre-nerf, this is it. Inference is actually quite cheap for token costs, the frontier labs burn most of their money on training new models, priced into their token costs ontop of some margins and paying record salaries. So if this goes open, distills are tried out, independent providers around the world host it with actual price competition, the h…

Quit my Claude pro subscription last week and purchased credits for an API inference provider. I think I might even end up saving money, since I really don’t use AI that much, and I actually found that gemma4:31b is fine for most of my non-coding inquiries.

Gemma is amazing with tools for anything that is not crazy complex. I think a lot of people have a wrong perception of it because Google's new prompt format broke implementations like llama.cpp and it took quite a while to get everything sorted. But even the tiny variants running on edge devices are surprisingly capable when used right.

The frontier will probably keep moving for a while, but it will be increasingly disconnected from normal human use. In the future, if you're not trying to solve a research level math problem, you'll probably do it locally and fully privately. Which also means the payday when they will fundamentally no longer be able to reach a billion users with frontier models will come soon for the labs. Even if they do get their IPO out, it will probably crash and burn at current valuations.

Re: GLM 5.2 Is Out

#329

In the last few days, Chinese labs have given us MiniMaxM3, KimiK2.7 and now GLM5.2. Meanwhile US is censoring models. Reads like fiction.

I don’t understand how I grew up thinking USA is the gold standard is good and China just make cheap copies and is bad.

But these news really changes my view on China and USA. I can’t believe it almost.

Re: GLM 5.2 Is Out

#330

In the last few days, Chinese labs have given us MiniMaxM3, KimiK2.7 and now GLM5.2. Meanwhile US is censoring models. Reads like fiction.

I don’t understand how I grew up thinking USA is the gold standard is good and China just make cheap copies and is bad. But these news really changes my view on China and USA. I can’t believe it almost.

Well china still making cheap copies (distills)
Post reply on HN