Earlier quoted context omitted.
Gpt2 was too dangerous to release. We just don't see it yet. Sure, the model itself was harmless, but it lit the fuse
Actually many of us do see that, and have been saying so for some time now.
GLM 5.2 Is Out
411–420 of 544 posts
Re: GLM 5.2 Is Out
#412Genuine question: How safe is it to use Chinese models via their services? Surely Anthropic and OpenAI are ingesting what I push there as well, but they're at least vaguely allied with my home country geopolitically. China on the other hand seems to be interested in supporting countries like Iran and Russia.
I suggest to not look at how each company is expressing themselves on the media, look at how they are actually behaving. When I first tried out Z.ai last year, I too was concerned regarding where my data goes. I vaguely remember from their ToS (please verify yourself too) that they followed a zero data retention policy for its AI services. This of course applied to their paying customers. I do not know if it applies…
edit: this is a comment about suing and enforcing judgments against Chinese companies in the US, especially software companies, not necessarily about how trustworthy the Chinese labs are.
Re: GLM 5.2 Is Out
#413Earlier quoted context omitted.
The Chinese models are censored (too?). > US is censoring models For the current Anthropic issue, I’d say that’s more likely to just be generic corruption, revenge, shakdeown, and/or incompetence from the Trump admin. ‘Censoring’ might be technically correct, but I think one of the aforementioned verbs is a better fit.
> corruption, revenge, shakdeown, and/or incompetence Sadly, I think it's all four at once.
It’s not just the models. Try copy pasting stuff out of the claude app, or sharing a conversation. It’s completely broken now.
Re: GLM 5.2 Is Out
#414Earlier quoted context omitted.
I don’t understand how I grew up thinking USA is the gold standard is good and China just make cheap copies and is bad. But these news really changes my view on China and USA. I can’t believe it almost.
> I don’t understand how I grew up thinking USA is the gold standard is good and China just make cheap copies and is bad You did not grow up in the 80s ... Where it was the same about US vs Japan. Look how it turned out for several of the US industries. The US tends to sleep, look down on other countries, and then it loses key industries because of that attitude.
I guess they’ll just milk the ICE assembly lines until they are bailed our or go under, Detroit-style.
Re: GLM 5.2 Is Out
#415Earlier quoted context omitted.
Gemma is amazing with tools for anything that is not crazy complex. I think a lot of people have a wrong perception of it because Google's new prompt format broke implementations like llama.cpp and it took quite a while to get everything sorted. But even the tiny variants running on edge devices are surprisingly capable when used right. The frontier will probably keep moving for a while, but it will be increasingly d…
Do you guys actually work with these models? I have to use GPT 5.4 Mini at work. It benchmarks higher than that Gemma 4 model. In my experience it's next to useless. It cannot even move 20 existing lines of code from A to B without breaking them half of the time. If you tell it to look something up in your dependencies, it's 50/50 on whether the answer is correct, incorrect, or it simply didn't perform the search at…
Re: GLM 5.2 Is Out
#416Seems like there's no official blog post with benchmark results yet. But I'm once again thankful for the Chinese AI labs for being open with their work and contributing it to the world under permissive licenses like this. The Fable 5 fiasco is just another reminder of how valuable these things are to have.
Based on my first impressions it's about 6 months behind the frontier labs. So very similar to Opus in January. That is, pretty damn impressive and very useable. When it comes to architecture or complex problems it does noticeable worse but I don't think anyone expected anything else. One particular interesting strong point seems to be design and user interfaces. It does seem to punch above it's weight there but that…
So it's not really similar to opus in January?
Re: GLM 5.2 Is Out
#417Given the US government’s latest stunt with Fable, this is looking more and more like the future. Can’t rely on strategic products if they’re gated by capricious actors. Open weight models are basically immune to that
> Open weight models are basically immune to that Somewhat. The US Gov can make it illegal to transact with, download, use, etc. foreign open weight models. Of course, enforcement will be difficult for individuals (businesses will comply by default, and they would all be pulled off Github and other US based hosting locations if they went the sanctions route). But, we are also quickly going down the road of frightenin…
It’d force people to run inference locally, and that’d expose the actual $/perf of the models instead of keeping it secret then propping it up with circular revenue and blatant securities fraud.
If we don’t do something like that, we won’t have much of an AI industry post-bubble.
Anyone else remember solyndra?
Re: GLM 5.2 Is Out
#418Earlier quoted context omitted.
I wouldn't bet on it. Chinese live the free market ideals instead of just preaching them but rent-seeking and seeking regulatory capture at the first opportunity. In China business doesn't control politics. Dynamics is completely different and so might be the outcomes.
Well I do hope you're right - that's a brighter future for all
Re: GLM 5.2 Is Out
#419Re: GLM 5.2 Is Out
#420Earlier quoted context omitted.
Quit my Claude pro subscription last week and purchased credits for an API inference provider. I think I might even end up saving money, since I really don’t use AI that much, and I actually found that gemma4:31b is fine for most of my non-coding inquiries.
Got a link to that API inference provider?