Live data from Hacker News

Anthropic walks back policy that could have 'sabotaged' researchers using Claude

wired.com

1–10 of 41 posts

Re: Anthropic walks back policy that could have 'sabotaged' researchers using Claude

#2
It's kinda weird to think the Chinese AI labs might be more trust worthy than the US labs.

- Anthropic is ran by a bunch of nut jobs.

- OpenAI is ran by a guy you can't trust.

I don't even know if we should include DeepMind, Meta, or xAi in the conversation of AI labs at this point since they can't produce models better than Chinese labs.

Re: Anthropic walks back policy that could have 'sabotaged' researchers using Claude

#3
Too late.

We now all know that Anthropic CAN do that if they want to. The fact that they told you upfront about it shows that their arrogance on this self-sabotage against their customers is at stratospheric levels.

Believe them the first time, and they are not your friends at all.

Re: Anthropic walks back policy that could have 'sabotaged' researchers using Claude

#4

It's kinda weird to think the Chinese AI labs might be more trust worthy than the US labs. - Anthropic is ran by a bunch of nut jobs. - OpenAI is ran by a guy you can't trust. I don't even know if we should include DeepMind, Meta, or xAi in the conversation of AI labs at this point since they can't produce models better than Chinese labs.

To be fair, nerfing Claude on frontier research tasks is consistent with Anthropic's stated beliefs. So in that sense you can trust them to always behave consistently if strangely. But this launch was done very poorly with the lack of transparency on when the frontier research policy was violated.

Re: Anthropic walks back policy that could have 'sabotaged' researchers using Claude

#5
Call me a cynic, but I don't believe this is a genuine change of heart at all. It feels much more like a panicked response to something that might undermine their IPO.

Even if you trust Anthropic today (which I don't), they clearly don't want competition and there's no telling what other shady moves they'll pull in future.

The only sustainable way forward is to support open models. I was already on the fence about whether or not to keep my Max subscription (the extra cost over something like DeepSeek V4 didn't really feel justifiable). This is the tipping point for me, I'll be cancelling my sub before it renews at the end of the month.

Re: Anthropic walks back policy that could have 'sabotaged' researchers using Claude

#7
The damage is done. They designed Fable to be dishonest and sabotage-y. Look at this report of Claude now randomly changing AI software without being asked to:

https://xcancel.com/hammer_mt/status/2064839924398825798

This is so completely dishonest. But it also shows how deeply anti competitive Anthropic is. They will talk about safety but it’s not actually about safety: building features like this seems intended to hurt competition in the AI space. They don’t mind if AI helps YOUR competitor but if it means competition for them, they suddenly have a problem with it.

I don’t care that they walked this back. They’ve shown who they are. And what they’re capable of.

Re: Anthropic walks back policy that could have 'sabotaged' researchers using Claude

#8
post #5

Call me a cynic, but I don't believe this is a genuine change of heart at all. It feels much more like a panicked response to something that might undermine their IPO. Even if you trust Anthropic today (which I don't), they clearly don't want competition and there's no telling what other shady moves they'll pull in future. The only sustainable way forward is to support open models. I was already on the fence about wh…

I think they are legitimately convinced that this model is so dangerous it could destroy the world and that they genuinely have the responsibility to prevent it from assisting other models to destroy the world.

I don't think I agree that I should be forbidden from e.g. patching a binary to work on the latest macOS since the company behind it died and intentionally installed a time-based kill-switch (FUCK ADOBE for popularizing that practice). But ooOOooOOoo working with machine code is so cybersecurity and therefore suspicious.

Re: Anthropic walks back policy that could have 'sabotaged' researchers using Claude

#9

It's kinda weird to think the Chinese AI labs might be more trust worthy than the US labs. - Anthropic is ran by a bunch of nut jobs. - OpenAI is ran by a guy you can't trust. I don't even know if we should include DeepMind, Meta, or xAi in the conversation of AI labs at this point since they can't produce models better than Chinese labs.

To be fair, nerfing Claude on frontier research tasks is consistent with Anthropic's stated beliefs. So in that sense you can trust them to always behave consistently if strangely. But this launch was done very poorly with the lack of transparency on when the frontier research policy was violated.

Yeah and their belief are fucking crazy and dangerous. They are literally sabotaging their users. They built in malware into their model if you prompt it about training a fucking AI model. It doesn't tell you, no it literally sabotages you by editing your prompt and intentionally goes against your request.

You want fucking nut jobs like this building models?

It's one thing to build safeguards on your model and have it prompt the user back. I'm sorry I can't help you with this request. Chinese models do this for some requests.

It's another thing to actively try to make the model perform worst for your user on purpose because it asked the model to do something you, the model creator, didn't like.

Imagine someone is asking a logical medical question and the model swaps the prompt and purpose being less intelligent and gives bad advice to this person.

How do these people not understand they are stupid.

Re: Anthropic walks back policy that could have 'sabotaged' researchers using Claude

#10
post #5

Call me a cynic, but I don't believe this is a genuine change of heart at all. It feels much more like a panicked response to something that might undermine their IPO. Even if you trust Anthropic today (which I don't), they clearly don't want competition and there's no telling what other shady moves they'll pull in future. The only sustainable way forward is to support open models. I was already on the fence about wh…

I think they are legitimately convinced that this model is so dangerous it could destroy the world and that they genuinely have the responsibility to prevent it from assisting other models to destroy the world. I don't think I agree that I should be forbidden from e.g. patching a binary to work on the latest macOS since the company behind it died and intentionally installed a time-based kill-switch (FUCK ADOBE for po…

The company was founded basically out of the effective altruism movement.
Post reply on HN