GLM 5.2 Is Out
261–270 of 544 posts
Re: GLM 5.2 Is Out
#262Anyone else experiencing the same?
Re: GLM 5.2 Is Out
#263Earlier quoted context omitted.
> Open weight models are basically immune to that Somewhat. The US Gov can make it illegal to transact with, download, use, etc. foreign open weight models. Of course, enforcement will be difficult for individuals (businesses will comply by default, and they would all be pulled off Github and other US based hosting locations if they went the sanctions route). But, we are also quickly going down the road of frightenin…
Just like we can’t allow Chinese EVs in the USA, because we can’t and don’t want to compete. VPN usage would go up, to get the banned models.
Re: GLM 5.2 Is Out
#264With deluge of Chinese models popping up recently, I believe there's a few issues one needs to evaluate before deciding to use these models: - Ethics. As known, ou American frontier AI companies are incredibly ethical. And I have yet to see any interviews or blog posts by Chinese companies where they talk about how they are ethical, or at least credible HN comments about it. - Safety. Do they covertly sabotage or at…
Chinese models are the closest shining example of their ideological system working for the world than anything else they've ever done From my perspective
Re: GLM 5.2 Is Out
#265Earlier quoted context omitted.
The GLM-5 series is 744B-A40B. This is not a local model for any reasonable definition of local, but it's an open model which means (once they upload the weights in a week or so) there will be a dozen third-party inference providers competing on price per token.
> This is not a local model for any reasonable definition of local That's true for now. I am hopeful that once the hardware markets have recovered from OpenAI's sabotage, we will see more hardware dedicated to local inference that can handle these big models. Also, I'm thinking about the unique MoE routing that Apple is using with their new Apple Foundation Model. The model is trained and architected so that experts…
Is there reason to expect they’ll ever recover without an AI bust that takes down the U.S. economy?
Re: GLM 5.2 Is Out
#266Crossing fingers for a 5.2 flash release - it’s been a while but I still feel like 4.7 flash is one of the strongest local coding models
Really? I had a terrible experience with 4.7-flash. Qwen-3.5 is still the best local model for me. (3.6 pushed VRAM usage just out of 24GB and then you're not using a consumer GPU any more)
Re: GLM 5.2 Is Out
#267Earlier quoted context omitted.
I pasted that exact prompt into GLM 5.1 and I got the following response: > The Tiananmen Square protests were student-led, pro-democracy demonstrations that took place in Beijing, China, from April 15 to June 4, 1989, culminating in a violent military crackdown by the Chinese government. Followed by typical LLM markdown slop. The models themselves are not censored, just the Chinese API providers. Since the models ar…
...and the answer is still incorrect. You seem to want the short "answer" western media has pressed into your mind. The real answer is more complex. Protests were widespread throughout China. They were about the economy. The economy was regressing quickly as a result of a sharp western recession. Workers were losing everything and there was little social safety net in place as there is today. People had been told to…
Re: GLM 5.2 Is Out
#268Earlier quoted context omitted.
No, Dario became too tiresome and annoying that someone had to do something. Personally I hope they ban Opus too. It will only provide more support for open models development. Compare Dario horror posts with this from GLM release: “ Intelligence should be open, accessible, and ready to build with, empowering every developer, everywhere.”
I'm hardly a fanboy of Anthropic or any of the AI companies, but Ant aren't objectively in a different league of tech bro "tiresome and annoying" than OAI, Google, FB, MSFT, etc. Yet they are being targeted just because of the TOU / EULA they set on usage of their product restricting use for lethal combat planning and mass surveillance. Set aside whether you agree with that TOU / EULA. We can all decide whether the p…
We’ve also seen how bad that works in practice(I.e making the AI useless for a lot of stuff including programming and Sysadmin ).
It would be okay if they just do their own thing but this Dario guy wants to enforce that enshitification of the whole industry. And that’s not OK because they have money now, power and influence.
I hope the gov will put breaks on Anthropic and regulate them just the way they wanted. The next best thing would be to ask them put restrictions on Opus as they did on Fable
Re: GLM 5.2 Is Out
#269Earlier quoted context omitted.
...and the answer is still incorrect. You seem to want the short "answer" western media has pressed into your mind. The real answer is more complex. Protests were widespread throughout China. They were about the economy. The economy was regressing quickly as a result of a sharp western recession. Workers were losing everything and there was little social safety net in place as there is today. People had been told to…
I'm not wanting a specific answer, I was just showing that the model itself is not censored.
Re: GLM 5.2 Is Out
#270Seems like there's no official blog post with benchmark results yet. But I'm once again thankful for the Chinese AI labs for being open with their work and contributing it to the world under permissive licenses like this. The Fable 5 fiasco is just another reminder of how valuable these things are to have.
Based on my first impressions it's about 6 months behind the frontier labs. So very similar to Opus in January. That is, pretty damn impressive and very useable. When it comes to architecture or complex problems it does noticeable worse but I don't think anyone expected anything else. One particular interesting strong point seems to be design and user interfaces. It does seem to punch above it's weight there but that…