"AI company warns of AI danger. Also, buy our AI, not their AI!"
Anthropic's models do not come out looking good in this research. If this is an ad for Anthropic's models, it's not a particularly great one.
Agentic Misalignment: How LLMs could be insider threats
61–70 of 86 posts
Re: Agentic Misalignment: How LLMs could be insider threats
#62Earlier quoted context omitted.
I think the narrative of "AI is just a tool" is much more harmful than the anthropomorphism of AI. Yes, AI is a tool. So are guns. So are nukes. Many tools are easy to be misused. Most tools are inherently dangerous.
I don’t quite follow. Just because a tool has the potential for misuse, doesn’t make it not a tool. Anthropomorphizing LLMs, on the other hand, has a multitude of clearly evident problems arising from it. Or do you focus on the “just” part of the statement? That I very much agree with. Genuinely asking for understanding, not a native speaker.
It's no longer "just a tool".
Re: Agentic Misalignment: How LLMs could be insider threats
#63Yeah, all the more reason not to have them doing autonomous behaviors. Rules of using AI: #1: Never use AI to think for you #2: Never use AI to do atomonous work That leaves using them as knowledge assistants. In time, that will be realized as their only safe application. Safe to the user's minds, and safe to the user's environment. They are idiot savants, after all, having them do atomonous work is short sighted.
> having them do atomonous work is short sighted
I also think they shouldn’t be doing atomonous work. Maybe autonomous work, but never atomonous.
Re: Agentic Misalignment: How LLMs could be insider threats
#64Yeah, all the more reason not to have them doing autonomous behaviors. Rules of using AI: #1: Never use AI to think for you #2: Never use AI to do atomonous work That leaves using them as knowledge assistants. In time, that will be realized as their only safe application. Safe to the user's minds, and safe to the user's environment. They are idiot savants, after all, having them do atomonous work is short sighted.
Sounds good on paper, but it has a game theory problem. If your efforts can always be out-raced by someone using AI to do autonomous work, don't you end up having to use it that way just to keep up?
Re: Agentic Misalignment: How LLMs could be insider threats
#65Yeah, all the more reason not to have them doing autonomous behaviors. Rules of using AI: #1: Never use AI to think for you #2: Never use AI to do atomonous work That leaves using them as knowledge assistants. In time, that will be realized as their only safe application. Safe to the user's minds, and safe to the user's environment. They are idiot savants, after all, having them do atomonous work is short sighted.
Good luck with that. We have a non-insignificant amount of people doing the #1 already, and the amount of people doing the #2 is only going to increase as more and more AIs are designed to be good at autonomous agentic behavior specifically. The ship has long sailed on "just never let AIs do anything dangerous". If that was your game plan on AI safety, you need a new plan.
Re: Agentic Misalignment: How LLMs could be insider threats
#66Yeah, all the more reason not to have them doing autonomous behaviors. Rules of using AI: #1: Never use AI to think for you #2: Never use AI to do atomonous work That leaves using them as knowledge assistants. In time, that will be realized as their only safe application. Safe to the user's minds, and safe to the user's environment. They are idiot savants, after all, having them do atomonous work is short sighted.
Sounds good on paper, but it has a game theory problem. If your efforts can always be out-raced by someone using AI to do autonomous work, don't you end up having to use it that way just to keep up?
Re: Agentic Misalignment: How LLMs could be insider threats
#67Yeah, all the more reason not to have them doing autonomous behaviors. Rules of using AI: #1: Never use AI to think for you #2: Never use AI to do atomonous work That leaves using them as knowledge assistants. In time, that will be realized as their only safe application. Safe to the user's minds, and safe to the user's environment. They are idiot savants, after all, having them do atomonous work is short sighted.
Sounds good on paper, but it has a game theory problem. If your efforts can always be out-raced by someone using AI to do autonomous work, don't you end up having to use it that way just to keep up?
The latter may be cheaper, sure. But too cheap can become very expensive quickly.
Re: Agentic Misalignment: How LLMs could be insider threats
#68Earlier quoted context omitted.
Which jobs do you think it actually can replace?
First of all job replacement is not hard, and doesn't require AI. As an example, we had release train engineers whose job was to make sure the right versions of submodules made it into the release, etc. Lots of running around and keeping track of things. We scripted like 95% of that away, and now it most of it happens automatically. The people who do that now do something else. I just turned a page of notes and requi…
Re: Agentic Misalignment: How LLMs could be insider threats
#69Re: Agentic Misalignment: How LLMs could be insider threats
#70Merge comments? https://news.ycombinator.com/item?id=44331150 I'm really getting bored of Anthropic's whole song and dance with 'alignment'. Krackers in the other thread explains it in better words.