> beats Claude in our Cyber Benchmarks Beats which model in Claude? Whenever a "benchmark" doesn't put precise model numbers in their headlines I am immediately skeptical. Either they don't know the difference (bad) or they are benchmarking against weaker models (misleading, also bad). It's like when studies say "AI is bad at X" and they used GPT-3.5 in current year.
GLM 5.2 beats Claude in our benchmarks
21–30 of 559 posts
Re: GLM 5.2 beats Claude in our benchmarks
#22Earlier quoted context omitted.
> is it going to be a crime to have them, run them, make them available? Now you're getting it! Commerce will call it a munition and those harboring it as harboring illegal/foreign munitions. No business will take the hit, so they will quickly deplatform the models. No end user has the GPU capacity to use GLM 5.2 or similar models at full precision so the government will call the problem "mostly solved." But they mig…
Or we use the models to work on fixing vulns and stop over-blowing the doom scenarios. Gotta save the kids and kill the terrorists though! I'm for making software better instead of banning it based on what the rich and powerful claim. I suspect the real fear is that open weight models undermine the financials and token prices they thought were going to pay off their ludicrous spending because they have all raced and…
That would be the rational thing to do.
> financials and token prices
I do not think the government thinks this deeply. Market manipulation might be a rational, if unethical reason to ban open source models.
But this admin banned Anthropic models to "own the libs." They will continue to ban what they want for whatever reason they want. I don't think those reasons will be particularly coherent.
Re: GLM 5.2 beats Claude in our benchmarks
#23GLM export controls incoming? I predict Commerce will force OpenRouter, HuggingFace to take some open models down within the next few months. Not that it would make any sense.
If that happens it'll be an absolute disaster. Imagine a scenario where Anthropic and OpenAI prohibit most US companies from using their latest models because of safety.. And meanwhile attackers use equivalent open source models to attack US companies. Any prohibition on open source models will do nothing to fix the problem.. since attackers will never feel bound to the law. All advanced models must be available for…
Re: GLM 5.2 beats Claude in our benchmarks
#24GLM export controls incoming? I predict Commerce will force OpenRouter, HuggingFace to take some open models down within the next few months. Not that it would make any sense.
>GLM export controls incoming? US imposing export restrictions on a model from China?
Re: GLM 5.2 beats Claude in our benchmarks
#25Earlier quoted context omitted.
While unlikely , it is not without precedent , there are restrictions on ASML a Dutch company to sell EUV machines
ASML complies as an ally, why would China comply? The weights are already available and downloaded, is it going to be a crime to have them, run them, make them available? Constitutional rights still exist (I hope)
Yeah. Illegal numbers.
Re: GLM 5.2 beats Claude in our benchmarks
#26Claude Code is an agent harness, not an LLM.
Claude is a brand (or group of LLMs), not an LLM.
Re: GLM 5.2 beats Claude in our benchmarks
#27Earlier quoted context omitted.
If that happens it'll be an absolute disaster. Imagine a scenario where Anthropic and OpenAI prohibit most US companies from using their latest models because of safety.. And meanwhile attackers use equivalent open source models to attack US companies. Any prohibition on open source models will do nothing to fix the problem.. since attackers will never feel bound to the law. All advanced models must be available for…
Right, but is there any evidence of intelligence behind any of these (government) decisions? It’s just regulatory capture + marketing (plus some people living out an imaginary fantasy that they’re in Neuromancer or something), absolutely no reason to think they won’t try and target open models as part of this.
If the real motive is profit, then open source models are likely simply not a viable means to that end.
Re: GLM 5.2 beats Claude in our benchmarks
#28GLM export controls incoming? I predict Commerce will force OpenRouter, HuggingFace to take some open models down within the next few months. Not that it would make any sense.
If that happens it'll be an absolute disaster. Imagine a scenario where Anthropic and OpenAI prohibit most US companies from using their latest models because of safety.. And meanwhile attackers use equivalent open source models to attack US companies. Any prohibition on open source models will do nothing to fix the problem.. since attackers will never feel bound to the law. All advanced models must be available for…
But that's the whole point.
Fall out of favor with the admin and you lose access to the good American models, aren't allowed to use Chinese ones, and fall prey to the attackers and behind your competitors.
Re: GLM 5.2 beats Claude in our benchmarks
#29Re: GLM 5.2 beats Claude in our benchmarks
#30It reads like an ad. Secondly these are "just" IDORs, arguably the easiest class of vulnerabilities. Thirdly it compares to GPT 5.5 and Opus 4.8. No, we don't have Mythos at home.