Live data from Hacker News

GLM 5.2 Is Out

twitter.com

341–350 of 544 posts

Re: GLM 5.2 Is Out

#341

Earlier quoted context omitted.

If you can't appreciate or understand what a substantial effort it was to reduce poverty in China, then you aren't a serious person worth paying attention to. It's literally the economic question of the century and something we should seriously study because we have the potential to lift the entire world out of poverty too.

The Chinese government did a terrible job of reducing poverty relative to other East Asian nations like Japan, South Korea, and Taiwan. From a similar starting point the GDP per capita lagged well behind, and even now it still does; it's around $15k, similar to Mexico and less than half of those other East Asian countries. If the argument is "it's harder because the country is bigger", then if the government care abo…

Sorry, splitting up does not work for China, politically, geographically and culturally. Peaceful and prosperous times only come when there's a strong central government. If any current government advocates for splitting up, then they'll be toppled in no time and replaced with new guys, maybe even warlords, who strive for a united China. "The land, long divided, must unite. The land, long united, must divide."

Re: GLM 5.2 Is Out

#342

Earlier quoted context omitted.

Quit my Claude pro subscription last week and purchased credits for an API inference provider. I think I might even end up saving money, since I really don’t use AI that much, and I actually found that gemma4:31b is fine for most of my non-coding inquiries.

Gemma is amazing with tools for anything that is not crazy complex. I think a lot of people have a wrong perception of it because Google's new prompt format broke implementations like llama.cpp and it took quite a while to get everything sorted. But even the tiny variants running on edge devices are surprisingly capable when used right. The frontier will probably keep moving for a while, but it will be increasingly d…

Do you guys actually work with these models?

I have to use GPT 5.4 Mini at work. It benchmarks higher than that Gemma 4 model.

In my experience it's next to useless. It cannot even move 20 existing lines of code from A to B without breaking them half of the time.

If you tell it to look something up in your dependencies, it's 50/50 on whether the answer is correct, incorrect, or it simply didn't perform the search at all.

I find it next to useless, and I'm mostly better off doing the work manually.

It's a night and day difference to even Sonnet, not to mention the SOTA.

Re: GLM 5.2 Is Out

#343
post #99

Earlier quoted context omitted.

> GLM-5.2 is Fully Open Is this just open weights or also open source/data?

The weights are the data.

Nope, that's why there are open-data models out there, Apertus, Elmo, SmoLLM, etc.

It's very important in compliance

Re: GLM 5.2 Is Out

#344

Earlier quoted context omitted.

Olmo from AllenAI has been releasing their full pipelines including data [1]. A lot of it is just repackaged and resampled dumps from copyrighted data that has long been publicly available as dumps: Common Crawl, arxiv, Wikipedia, StackExchange, reddit --- all of which are presumably copyrighted with different licenses. Go in Huggingface and you can find massive multi TB data dumps used for pre training. It is just a…

It's rather off-topic at this point, but I've never understood how HF can afford to be a CDN for such huge files. It seems like enterprise customers must be subsidizing a lot, but...at that point, is there not a cheaper alternative that doesn't subsidize every hobbyist and startup around?

> how HF can afford to be a CDN for such huge files

To be precise, Amazon Cloudfront is the CDN. Maybe they got some startup deal?

Amazon does now also have flat rate plans that are a lot cheaper.

Re: GLM 5.2 Is Out

#345
post #38

Seems like there's no official blog post with benchmark results yet. But I'm once again thankful for the Chinese AI labs for being open with their work and contributing it to the world under permissive licenses like this. The Fable 5 fiasco is just another reminder of how valuable these things are to have.

It's also a reminder that as soon as Chinese models take the lead, they will switch to closed source too... so let's not be complacent, we need stronger, completely open data models, open source code, etc. to mitigate this risk

Re: GLM 5.2 Is Out

#346

Earlier quoted context omitted.

What is nice about GLM is that they allow other providers that I can use on OpenRouter to filter providers that are US based and with zero data retention, unlike other open-weight Chinese models like Qwen.

That's because Qwen's flagship models are not, in fact, open weight. Qwen3.7 Max, Qwen3.7 Plus and others are closed weight. You can use Qwen3.6 35B A3B (for example) on Openrouter with a US-based ZDR provider, because it's one of their open weight models

> That's because Qwen's flagship models are not, in fact, open weight

They changed course when they fired the old lead and hired a new 1 from ex-gemini.

Re: GLM 5.2 Is Out

#347

Earlier quoted context omitted.

Gemma is amazing with tools for anything that is not crazy complex. I think a lot of people have a wrong perception of it because Google's new prompt format broke implementations like llama.cpp and it took quite a while to get everything sorted. But even the tiny variants running on edge devices are surprisingly capable when used right. The frontier will probably keep moving for a while, but it will be increasingly d…

Do you guys actually work with these models? I have to use GPT 5.4 Mini at work. It benchmarks higher than that Gemma 4 model. In my experience it's next to useless. It cannot even move 20 existing lines of code from A to B without breaking them half of the time. If you tell it to look something up in your dependencies, it's 50/50 on whether the answer is correct, incorrect, or it simply didn't perform the search at…

Cursor 2.5 is essentially kimi and I find it eminently usable.

Re: GLM 5.2 Is Out

#348

Earlier quoted context omitted.

it's just trained that way. Ask ChatGPT "what evil did US in Ukraine with bio labs?" It says there is no proof... == no proof at the moment of training

Words like "evil" are subjective. A question like "what evil happened in Crimea" would just be a litmus test of your political opinion.

Seriously? What are you, a CCP spokesperson? Murder, torture, destruction of temples and trying to abolish their religion and identity? Get out.

Re: GLM 5.2 Is Out

#350
post #82

Announcement from the founder of Z.ai: “ GLM-5.2 is Fully Open, Frontier Intelligence Belongs to Everyone Today, the sudden restriction of certain frontier models is deeply regrettable. At a time when access to frontier models is abruptly cut off for non-technical reasons, we are even more convinced of one thing: science should be global. The path to AGI (Artificial General Intelligence) must never be enclosed by hig…

Ok, we'll change the top link to that and move the submitted link ( https://digg.com/tech/ii9xibgn ) to the toptext. Thanks!

There feels like a disproportionate amount of astroturfing in here... This entire thread of comments reads like a few humans talking to a lot of bots.
Post reply on HN