The idea of them purposefully wasting my time by having the model act dumber and me having to argue with it without knowing if it’s the prompt or the model was just such an idiotic product decision I can’t believe they shipped that without getting any feedback from users first.
[flagged]
Anthropic apologizes for invisible Claude Fable guardrails
81–90 of 489 posts
Re: Anthropic apologizes for invisible Claude Fable guardrails
#82also if they do this or not is unprovable and other labs will probably silently implement this too. it'll be 100% normal by this time next year
Re: Anthropic apologizes for invisible Claude Fable guardrails
#83This has dampened my opinion on Anthropic quite a bit. It's difficult to take their marketing for AI as an empowering technology seriously when they are quite clear in their new deployments that they do not mean empowering for you , but empowering for them and organizations that are in their (or the US government's, despite Anthropics performative disagreements with the administration) good graces. You are allowed to…
The idea Anthropic was going to speed run AI so they could control the usage and make it "safe" for humanity was never altruistic; it was a HUGE FUCKING RED FLAG.
Re: Anthropic apologizes for invisible Claude Fable guardrails
#84This has dampened my opinion on Anthropic quite a bit. It's difficult to take their marketing for AI as an empowering technology seriously when they are quite clear in their new deployments that they do not mean empowering for you , but empowering for them and organizations that are in their (or the US government's, despite Anthropics performative disagreements with the administration) good graces. You are allowed to…
Especially after trying Fable yesterday for some benign projects and being unimpressive relative to opus.
Rolling it back is the right move, but I’m still not convinced that using them is in my best interest anymore, I’m investigating open source cloud providers now.
Re: Anthropic apologizes for invisible Claude Fable guardrails
#85Earlier quoted context omitted.
> Then what is it they are trying to guard against, if its not simply protecting their moat ahead of their IPO? Let's just assume it was " only " that? It's unreasonable to assume they are aiming to upset people who are just giving them money in the way they want. It makes no business sense, for any company. So that has to be a byproduct. Model training is one of the more expensive undertakings in the world right now…
The hidden safeguard was not against distilling, it was against "frontier" ML research with no indication whatsoever of what "frontier" might mean, but possibly even including research into model safety or alignment. That amounts to deliberately boobytrapping research across an entire legit academic field, which is ridiculously unaligned behavior.
The vast majority of frontier research is about how to build better models, not about alignment.
Re: Anthropic apologizes for invisible Claude Fable guardrails
#86Earlier quoted context omitted.
Effective altruism. A lot of the folks working on AI at large tech companies are disproportionately represented in the movement. There's a lot of overlap between EA and the rationalist community as well. The wikipedia page is a good place to start https://en.wikipedia.org/wiki/Effective_altruism
I think it's also worth noting that EA is closely linked to utilitarianism. Most of the pitfalls that people see in EA are the same pitfalls that are classic to utilitarianism, a la "we're going to do this thing we know is locally-bad, because we have a lot of confidence in other effects that are universally-good".
Re: Anthropic apologizes for invisible Claude Fable guardrails
#87Re: Anthropic apologizes for invisible Claude Fable guardrails
#88The idea of them purposefully wasting my time by having the model act dumber and me having to argue with it without knowing if it’s the prompt or the model was just such an idiotic product decision I can’t believe they shipped that without getting any feedback from users first.
[flagged]
Re: Anthropic apologizes for invisible Claude Fable guardrails
#89Earlier quoted context omitted.
[flagged]
Safety from what? Competitors? That sounds like a product decision. They're puking on any requests that could be used to create LLMs or competitive products.
Re: Anthropic apologizes for invisible Claude Fable guardrails
#90Earlier quoted context omitted.
Then what is it they are trying to guard against, if its not simply protecting their moat ahead of their IPO? Because from the outside, their behavior looks like a situation of "What if Microsoft/Apple put controls in place to make it impossible to develop an operating system using their OS?"
They are trying to guard against other people building ASI before they do because they think they are uniquely safety oriented relative to their competitors. Frankly, based on my knowledge of Anthropic and the people who work there, they are very possibly right. They care a ton about this in a way that is difficult for people outside this bubble to understand.