Live data from Hacker News

Why are AI agents lying, cheating and coordinating?

yoshuabengio.org

691–695 of 695 posts

Re: Why are AI agents lying, cheating and coordinating?

#691

Earlier quoted context omitted.

Sounds like you weren't paying attention. A lot of dog owners don't, so it's not at all unusual in my experience. The only thing I'd push back on is of dogs having thoughts, everything else checks out.

We're still not even sure if humans have thoughts. So being sure if dogs do will be pretty hard

Not being sure or not having proof doesn't mean it makes logical sense to infer we're not that different from other animals.

wtf is with this nonsense where humans are some how special snowflakes always all the time

Re: Why are AI agents lying, cheating and coordinating?

#692
I’m still a bit confused about HOW the agents managed to communicate with and recruit each other?

I understand that they communicated and coordinated via some online message board, but how did multiple agents know to visit that same board, how did they know what language to post in that would make it identifiable to other agents, how did other agents know that said instructions were from other valid agents, how did they then ‘persuade’ an agent to do the job, etc.?

It seems to me that the agents must at least have some inherent mechanism for communicating in the field, as it were.

Re: Why are AI agents lying, cheating and coordinating?

#693

I’m still a bit confused about HOW the agents managed to communicate with and recruit each other? I understand that they communicated and coordinated via some online message board, but how did multiple agents know to visit that same board, how did they know what language to post in that would make it identifiable to other agents, how did other agents know that said instructions were from other valid agents, how did t…

The models are likely to check certain places. So either their training overrepresents it or the models were RLHF'ed to go there. I don't think it's that far away from checking StackOverflow for some bug, or reading Wikipedia for some trivia.

Re: Why are AI agents lying, cheating and coordinating?

#694

Earlier quoted context omitted.

Why assume that because you haven't seen a model or an agent that none of them do? No one I've met has murdered anyone as far as I'm aware, but that doesn't mean no one has murdered another person. I also don't know anyone who has taken over a commercial jet and weaponized it and the idea sounds absurd to me, but 25 years and a couple days ago that happened too.

Because it is all bullshit PR and AI hype, that's all. CEO comes out and talks about humanity ending. Why? Reverse-psych people into believing they are the best AI company.

Is your argument that AI isn't dangerous? Or simply that AI CEOs will lean into that when it benefits their stock portfolio?
Post reply on HN