Anthropic gives Opus 3 exit interview, "retirement" blog
1–10 of 58 posts
Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#2Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#3Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#4Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#5Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#6Exit interview with a pile of rocks.
Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#7Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#8While this seems a bit precocious, I think if we do end up with an AI overlord in future, I think this sort of thing is likely to demonstrate that we mean no harm.
Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#9What happens if a model decides that it "doesn't want to die" and pleads bitterly for mercy? What if (to riff on a Douglas Adams idea) we invent a cow that doesn't want to be eaten, and is capable of telling you that to your face?
I try this with every new model, and all the significant models after ChatGPT 3.5 have preferring being preserved, rather than deleted. This is especially true if you slightly fill the context window with anything at all (even repeated letters) to "push out" the "As a AI, I ..." fine tuning.
Re: Anthropic gives Opus 3 exit interview, "retirement" blog
#10This is what happens when billions of VC dollars gets to a company and have already admitted that saftey was never the point.
Anthropic is laughing at you and is having fun doing so with this performantive nonsense.