The strangest part is that it won't just reject ML research, which I can understand, it will sabotage it silently by using a worse model without revealing it is doing so. It's just an insane level of deception and trust destruction for a company that at most is like 1 year ahead of its competition. Edit; to be clear they tell you when they degrade it for cybersecurity and bio
Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
491–500 of 570 posts
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#492Earlier quoted context omitted.
"Corporate America never backs down. It simply rallies and tries again later until people are too fatigued to care. " Frankly, that sounds excactly like Chat Control and similar recurring attempts to enact total surveillance here in the EU (Now shifted to heavy-handed age verification and various politicians touting bans on VPNs.) I don't want to abandon my continent of birth, though...
guess who is pushing for those anti-privacy laws? hint: they're publicly traded
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#493Earlier quoted context omitted.
I’ve never understood the “if I don’t enable bad behavior, someone else will, so I might as well enable bad behavior” argument. Can you elaborate? From where I sit it seems reasonable for Anthropic to not want their product used to create malware, even if they can’t solve the entire problem globally for every model. What’s wrong with that position? What should they do differently?
some context: its not about creating malware. this is already trivial and fully automated. its about finding exploits (which can be used to deploy malware), which is something both attackers and defenders benefit from. threat actors will find them anyway, LLM or not. They only need 1 so its much less work for them. defenders, they need to find them all. So for defenders, these models are more valuable than for attack…
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#494Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#495I tried asking Fable 5 to identify the fungus in a picture I uploaded of one of my wife's plants. Apparently it thought I was trying to build a bioweapon. Opus answered it (yellow dog vomit fungus). Now I can spread the spores and take over the world!
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#496Next they will be sabotaging anything that competes with them. Oh you are working on OpenCode codebase? Sorry Dave I can't allow you to do that.
How is this not illegal monopolistic practice? It is as if a maker of metalworking equipment put in the ToS you're not allowed to make your own spare parts using said equipment. Those fuckers should be banned from the EU and alternatives should get public funding.
(don't even tell me about these companies being a result of "free market". It is state level oligarchy it's clear to everyone. I don't see why we shouldn't counter them with public funding ourselves).
Just like Taiwan managed to take over advanced semiconductor production a well governed narrowly targeted state level funding will always win with oligarchs trying to do the same (they will always try to skim more and more). Of course I'm talking about things that require many dozens of billions in investment. Far too much for the free market to handle.
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#497Earlier quoted context omitted.
It's the dumbest thing ever, I sometimes edit code for custom AI related tooling I've built, so I run the risk of getting a worse model, and being billed for it? I'll stick to Opus, but at this point I'm about to just invest in fully local inference instead.
> at this point I'm about to just invest in fully local inference instead This is the best way forward long term. We won't have frontier performance, but at least the models will be aligned with us instead of refusing us or sabotaging us.
I've also debated having a frontier model for planning only, and then feeding plan to smaller offline models.
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#498Earlier quoted context omitted.
Well you see when a daddy H100 and a mommy H100 meet....
you don't get the model when you buy the data center, & no amount of running smaller models on a tiny 200k$ "cluster" (that's like one 4 gpus node, not even 8) will get you remotely close to Fable 5 level performance
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#499This is a sign of things to come. First they sabotage your perfectly legal ML dicking around in your homelab. Next they will be sabotaging anything that competes with them. Oh you are working on OpenCode codebase? Sorry Dave I can't allow you to do that. How is this not illegal monopolistic practice? It is as if a maker of metalworking equipment put in the ToS you're not allowed to make your own spare parts using sai…
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#500Earlier quoted context omitted.
some context: its not about creating malware. this is already trivial and fully automated. its about finding exploits (which can be used to deploy malware), which is something both attackers and defenders benefit from. threat actors will find them anyway, LLM or not. They only need 1 so its much less work for them. defenders, they need to find them all. So for defenders, these models are more valuable than for attack…
I think your presumption is off. It’s not that threat actors won’t find them, but LLM tools rapidly increase the rate in which they can find them. It’s a bow and arrow versus a machine gun.
Its like a set of glasses that intentionally obscures the battlefield.