It reads like an ad. Secondly these are "just" IDORs, arguably the easiest class of vulnerabilities. Thirdly it compares to GPT 5.5 and Opus 4.8. No, we don't have Mythos at home.
GLM 5.2 beats Claude in our benchmarks
11–20 of 559 posts
Re: GLM 5.2 beats Claude in our benchmarks
#12Earlier quoted context omitted.
>GLM export controls incoming? US imposing export restrictions on a model from China?
While unlikely , it is not without precedent , there are restrictions on ASML a Dutch company to sell EUV machines
The weights are already available and downloaded, is it going to be a crime to have them, run them, make them available? Constitutional rights still exist (I hope)
Re: GLM 5.2 beats Claude in our benchmarks
#13Beats which model in Claude? Whenever a "benchmark" doesn't put precise model numbers in their headlines I am immediately skeptical. Either they don't know the difference (bad) or they are benchmarking against weaker models (misleading, also bad).
It's like when studies say "AI is bad at X" and they used GPT-3.5 in current year.
Re: GLM 5.2 beats Claude in our benchmarks
#14GLM 5.2 is already capable enough to assist in self-training which is similar to what we saw happen with frontier models and they appear to be getting there at a significantly lower cost than openai/anthropic.
Re: GLM 5.2 beats Claude in our benchmarks
#15Earlier quoted context omitted.
While unlikely , it is not without precedent , there are restrictions on ASML a Dutch company to sell EUV machines
ASML complies as an ally, why would China comply? The weights are already available and downloaded, is it going to be a crime to have them, run them, make them available? Constitutional rights still exist (I hope)
Now you're getting it! Commerce will call it a munition and those harboring it as harboring illegal/foreign munitions.
No business will take the hit, so they will quickly deplatform the models.
No end user has the GPU capacity to use GLM 5.2 or similar models at full precision so the government will call the problem "mostly solved." But they might choose to "make examples" out of a few people using p2p software to download the weights if they choose to.
Re: GLM 5.2 beats Claude in our benchmarks
#16GLM export controls incoming? I predict Commerce will force OpenRouter, HuggingFace to take some open models down within the next few months. Not that it would make any sense.
>GLM export controls incoming? US imposing export restrictions on a model from China?
Re: GLM 5.2 beats Claude in our benchmarks
#17> beats Claude in our Cyber Benchmarks Beats which model in Claude? Whenever a "benchmark" doesn't put precise model numbers in their headlines I am immediately skeptical. Either they don't know the difference (bad) or they are benchmarking against weaker models (misleading, also bad). It's like when studies say "AI is bad at X" and they used GPT-3.5 in current year.
Re: GLM 5.2 beats Claude in our benchmarks
#18Earlier quoted context omitted.
ASML complies as an ally, why would China comply? The weights are already available and downloaded, is it going to be a crime to have them, run them, make them available? Constitutional rights still exist (I hope)
> is it going to be a crime to have them, run them, make them available? Now you're getting it! Commerce will call it a munition and those harboring it as harboring illegal/foreign munitions. No business will take the hit, so they will quickly deplatform the models. No end user has the GPU capacity to use GLM 5.2 or similar models at full precision so the government will call the problem "mostly solved." But they mig…
I'm for making software better instead of banning it based on what the rich and powerful claim.
I suspect the real fear is that open weight models undermine the financials and token prices they thought were going to pay off their ludicrous spending because they have all raced and raised hardware prices.
Re: GLM 5.2 beats Claude in our benchmarks
#19GLM export controls incoming? I predict Commerce will force OpenRouter, HuggingFace to take some open models down within the next few months. Not that it would make any sense.
Any prohibition on open source models will do nothing to fix the problem.. since attackers will never feel bound to the law. All advanced models must be available for defensive purposes.
Re: GLM 5.2 beats Claude in our benchmarks
#20Earlier quoted context omitted.
> is it going to be a crime to have them, run them, make them available? Now you're getting it! Commerce will call it a munition and those harboring it as harboring illegal/foreign munitions. No business will take the hit, so they will quickly deplatform the models. No end user has the GPU capacity to use GLM 5.2 or similar models at full precision so the government will call the problem "mostly solved." But they mig…
Or we use the models to work on fixing vulns and stop over-blowing the doom scenarios. Gotta save the kids and kill the terrorists though! I'm for making software better instead of banning it based on what the rich and powerful claim. I suspect the real fear is that open weight models undermine the financials and token prices they thought were going to pay off their ludicrous spending because they have all raced and…