Earlier quoted context omitted.
Anthropic's claim was that Deepseek collected ~150k conversations. https://www.anthropic.com/news/detecting-and-preventing-dist... I think the extent of distillation by Deepseek specifically is overstated. For comparison, Minimax collected over 13m 'exchanges', which starts to sound a lot more like large-scale distillation.
Ah, dang it. My college professors warned me about this: the Wikipedia page I read the other day is wrong!
Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
371–380 of 570 posts
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#372Earlier quoted context omitted.
I read the same announcement. Or more precisely, I read at least two slightly different revisions of the announcement (it was updated between my two passes). Our org has ZDR, and has had it since the contract was signed. Yesterday two things held true at the same time: 1. Fable was available if you had at least .170 CLI client; and 2. ZDR was no longer on By the time West Coast woke up, the admin panel apparently had…
You mean off as in no Data Retention? Or in we turned off your ZDR Policy so we collect all your data now?
Somewhere along the line we also used the self-service toggle to turn ZDR back on. I am not 100% certain of the exact timeline of interleaving events, many of the actions were taken by our Western US folks. Sorry. It's been a bit hectic over the past ~36h...
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#373Earlier quoted context omitted.
Imo that's a big win. The LLM just gaslighting you into suboptimal approaches was insane.
I guess, but yesterday Anthropic had their version of Google removing the "Don't be evil" from their motto. They destroyed a metric ton of goodwill they'll never regain.
I mean, did nobody ever get the vibes, never see a pattern emerging? (well they don't or they wouldn't be so amazed by pattern recognition machines on steroids)
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#374Earlier quoted context omitted.
Trains made by Newag were programmed to brick themselves if they detected a non-Newag workshop was repairing them. https://news.ycombinator.com/item?id=38638865 https://news.ycombinator.com/item?id=38628635 https://news.ycombinator.com/item?id=38567687 https://news.ycombinator.com/item?id=38530885
And that was correctly perceived to be illegal by antitrust regulators.
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#375News just broke in this Wired story: "Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude" https://www.wired.com/story/anthropic-responds-to-backlash-o... > “We’re changing Fable 5’s safeguards for frontier LLM development to make them visible.” Anthropic said in a statement to WIRED. “We made the wrong tradeoff and we apologize for not getting the balance right.” Sounds like the wides…
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#376Let's all vote with our wallets and collectively boycott misAnthropic or at least their feeble fable safety theater. Whining on social media only goes so far, especially when they're concealing their anticompetitive strategies under the veil of safety.
Tastes like... astroturf.
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#377News just broke in this Wired story: "Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude" https://www.wired.com/story/anthropic-responds-to-backlash-o... > “We’re changing Fable 5’s safeguards for frontier LLM development to make them visible.” Anthropic said in a statement to WIRED. “We made the wrong tradeoff and we apologize for not getting the balance right.” Sounds like the wides…
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#378The question is: If biological, computer security, and ML research are so bad, why do they even train on the relevant data? The only answer that makes sense is they wanted the model to be competent and usable in these fields, just not by you , which is why they had to bolt on a badly functioning crippling device after the fact.
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#379Long live static websites without any Javascript.
Re: Cybersecurity researchers aren't happy about the guardrails on Anthropic's Fable
#380So I suspect Anthropic started A/B testing or just plain testing this a while ago, Tell HN: Claude flags biology / biotech questions https://news.ycombinator.com/item?id=47929885 Today, it's flagging population research questions, Using only the dataset you constructed, assess two questions: 1. **Mortality:** do [GROUP] show mortality that differs from (a) your comparison groups and (b) era- and sex-matched US popula…
I was digging into some orbital mechanics questions and I assume it decided I was trying to backyard-science my way into an orbital-bombardment weapon. Kind of wild how this product's impression has gone from "wow, this is pretty neat" to "irreverent sack of dog shit you" in 24 hours almost solely on the back of a half-baked moderation system.