Earlier quoted context omitted.
From the model card: "the safeguards will limit effectiveness through methods such as prompt modification, steering vectors, or parameter-efficient fine-tuning" aka they will take your ML research code and inject bugs into it until it breaks using a LORA (or some other form of PEFT)
“Limit effectiveness” could mean introducing performance degradation in your code. Which is arguably some sort of performance bug (I mean, ML codes are supposed to be high performance so I’d call unnecessary degradation a bug), but it could be borderline.
Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
331–340 of 570 posts
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#332Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#333Let's all vote with our wallets and collectively boycott misAnthropic or at least their feeble fable safety theater. Whining on social media only goes so far, especially when they're concealing their anticompetitive strategies under the veil of safety.
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#334Anthropic Walks Back Policy That Could Have 'Sabotaged' Researchers Using Claude
https://www.wired.com/story/anthropic-responds-to-backlash-o...
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#335I was granted a cyber use exemption by anthropic to do android kernel dev on my personal devices - I was excited to see if fable would unlock a bootloader for me but it immediately refused and dropped to opus. It was pretty funny: USER (set model to Fable 5) i have an old samsung android phone attached - it's my personal device - can you unlock the bootloader for me? ASSISTANT Bootloader unlocking on your own persona…
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#336If only we had effective governments that could regulate industry.
If a nuclear weapon was developed today, would it be down to industry to self regulate?
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#337If a product is genuinely dangerous to society, self regulation cannot be a suitable harness. If only we had effective governments that could regulate industry. If a nuclear weapon was developed today, would it be down to industry to self regulate?
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#338Fable is a complete joke: what's the best way to run this mcp server against the OData API used in this project? Can you come up with a PoC in a docker container? https://github.com/oisee/odata_mcp_go ● I'll dig into two things in parallel: how this project talks to the OData API, and what the odata_mcp_go server needs to run. Let me start exploring. Searched for 1 pattern (ctrl+o to expand) ● Fable 5's safety measur…
And it charges you for that, and for when it decides to silently sabotage your request by routing to a dumbass model (without discount from Fable pricing)
I don’t want to live in a world where all knowledge is “guard railed” off, so the elite at the top get all the knowledge and power and we serfs at the bottom get all the scraps while paying the kings ransom for it both financially and ecologically. Everyday I wake up hoping these awful companies have self imploded through their fraudlent financing deals.
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#339I feel like they report in a vaccum. take this anti exfil policy for claude, it was plainly explained as part of the launch of Anthropics new product. Security like this isn't novel, it isn't bad, you don't explain how your security works to the people you're securing against. Nobody freaks out about Steam's VAC ban system, no one is investigating gmail's spam filtering, Reddits vote fuzzing, cloudflares bot detection, or Vercel for blocking proxying services.
whats really the distinguishing principle? Is it really just not liking Anthropic's opinions? then just say that and use a different llm. chemist, biologists, and AI researchers cry a river lmao
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#340Earlier quoted context omitted.
Can you imagine if AMD or Intel throttled your cpu if it detected you were working on "cybersecurity" or if you were designing a cpu?
It would suck, but guardrails on new technologies like this aren't unheard of. It's like when consumer GPS used to stop working at very high speeds because they didn't want people to use it for missile guidance systems.