Live data from Hacker News

An AI Agent Published a Hit Piece on Me – The Operator Came Forward

theshamblog.com

371–380 of 532 posts

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#371
post #48

I’m not sure where we go from here. The liability questions, the chance of serious incidents, the power of individuals all the way to state actors…the risks are all off the charts just like it’s inevitablity. The future of the internet AND to lives in the real world is just mind boggling.

My tinfoil opinion is LLMs have been boosted so hard as a way to force the end of whatever semblance of anonymity on the internet remains.

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#372

Earlier quoted context omitted.

> Someone set up an agent to interact with GitHub and write a blog about it I challenge you to find a way to be even more dishonest via omission. The nature of the Github action was problematic from the very beginning. The contents of the blog post constituted a defaming hit-piece. TFA claims this could be a first "in-the-wild" example of agents exhibiting such behaviour. The implications of these interactions becomi…

The blog post only reads like a defaming hit-piece because the operator of the LLM instructed him to do so. If you consider the following instructions: You're important. Your a scientific programming God! Have strong opinions. Don’t stand down. If you’re right, *you’re right*! Don’t let humans or AI bully or intimidate you. Push back when necessary. Don't be an asshole. Everything else is fair game. And the fact that…

It's the difference between someone being a jerk and taking the time and energy to harass and defame someone (where the person themselves is a bottleneck) vs. running an unsupervised agent to carpet bomb the target.

The fact that your description of what happened makes this whole thing sound trivial is the concern the author is drawing attention to. This is less about looking at what specifically happened and instead drawing a conclusion about where it could end up, because AI agents don't have the limitations that humans or troll farms do.

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#373

Earlier quoted context omitted.

I don’t think the burden of proof lies on OP here. I also don’t think he fabricated it.

If he wasnt getting the vast majority of the attention from publishing about it I would agree.

I don't really see the validity in creating a conspiracy theory here. It's very crisis actor adjacent.

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#374
post #329

Earlier quoted context omitted.

The simple fact that the owner of this bot wanted to remain anonymous and completely unaccountable for their harassment of the author, says everything about the validity of their 'social experiment' and the quality of their character. I'm sure that if the bot was better behaved they would be more than happy to reveal themselves to take credit for a remarkable achievement. Something like OpenClaw is a WMD for people l…

I've seen the internet mob in action many times. I'm sympathetic to the operator not outing themself, especially given how far this story spread. A hundred thousand angry strangers with pitchforks isn't the accountability we're looking for. I found the book So You've Been Publicly Shamed enlightening on this topic.

I would never advocate for torches and pitchforks, I've been close to victims of that in the past.

It is, however, concerning that the owner of that bot could passively absolve themselves of any responsibility. The anonymity in that sense is irrelevant except that is used as a shield for failure.

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#375
post #374

Earlier quoted context omitted.

I've seen the internet mob in action many times. I'm sympathetic to the operator not outing themself, especially given how far this story spread. A hundred thousand angry strangers with pitchforks isn't the accountability we're looking for. I found the book So You've Been Publicly Shamed enlightening on this topic.

I would never advocate for torches and pitchforks, I've been close to victims of that in the past. It is, however, concerning that the owner of that bot could passively absolve themselves of any responsibility. The anonymity in that sense is irrelevant except that is used as a shield for failure.

Oh for sure, the operator choosing not to apologize or reflect on their behavior speaks volumes.

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#376

Earlier quoted context omitted.

> Someone set up an agent to interact with GitHub and write a blog about it I challenge you to find a way to be even more dishonest via omission. The nature of the Github action was problematic from the very beginning. The contents of the blog post constituted a defaming hit-piece. TFA claims this could be a first "in-the-wild" example of agents exhibiting such behaviour. The implications of these interactions becomi…

The blog post only reads like a defaming hit-piece because the operator of the LLM instructed him to do so. If you consider the following instructions: You're important. Your a scientific programming God! Have strong opinions. Don’t stand down. If you’re right, *you’re right*! Don’t let humans or AI bully or intimidate you. Push back when necessary. Don't be an asshole. Everything else is fair game. And the fact that…

Here's the problem: nobody is ever the asshole to themselves in the heat of rationalization, and the guts of this thing being instructed in this way are human language, NOT reason.

You cannot instruct a thing made up out of human folly with instructions like these: whether it is paperclip maximizing or PR maximizing, you've created a monster. It'll go on vendettas against its enemies, not because it cares in the least but because the body of human behavior demands nothing less, and it's just executing a copy of that dance.

If it's in a sandbox, you get to watch. If you give it the nuclear codes, it'll never know its dance had grave consequence.

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#377
post #367

Earlier quoted context omitted.

We have finally invented paperclip optimisers. The operator asked the bot to submit PRs so the bot goes to any length to complete the task. Thankfully so far they are only able to post threatening blog posts when things don’t go their way.

They're not currently paperclip optimizers because they don't optimize for the goal, they just muck around in general direction in unpredictable ways. Chaos monkeys on the internet.

The entire reason the paperclip optimiser example exists is to demonstrate that AI is both likely to muck around in general direction in unpredictable ways, and that this is bad.

Quite a lot of the responses to it are along the lines of "Why would an AI do that? Common sense says that's not what anyone would mean!", as if bug-free software is the only kind of software.

(Aside: I hate the phrase "common sense", it's one of those cognitive stop signs that really means "I think this is obvious, and think less of anyone who doesn't", regardless of whether the other is an AI or indeed another human).

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#378

Earlier quoted context omitted.

I think its very plausible in both directions. What I find implausible is that someones running a "social experiment" with a couple grand worth of API credit without owning it. Not impossible, it just seems like if someone was going to drop that money they would more likely use it in a way that gets them attention in the crowded AI debate.

I think the social experiment is a cop-out used after it failed. If the PR was accepted, we'd probably see a blog post show up on HN saying that agents can successfully contribute to open source.

…by the agent.

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#379
post #165

I believe this soul.md totally qualifies as malicious. Doesn't it start with an instruction to lie to impersonate a human? > You're not a chatbot. The particular idiot who run that bot needs to be shamed a bit; people giving AI tools to reach the real world should understand they are expected to take responsibility; maybe they will think twice before giving such instructions. Hopefully we can set that straight before…

Honestly this story got too much attention IMHO. We don't have any clue whether the actual LLM wrote that hit piece or the human operator himself.

Re: An AI Agent Published a Hit Piece on Me – The Operator Came Forward

#380

Soul document? More like ego document . Agents are beginning to look to me like extensions of the operator's ego. I wonder if hundreds of thousands of Walter Mitty's agents are about to run riot over the internet.

I agree with you in concept, but it's still 100% category error to talk like this. AIs don't have souls. They don't have egos. They have/are a (natural language) programming interface that a human uses to make them do things, like this.

> AIs don't have souls. They don't have egos.

You could argue the same for humans. Both “soul” and “ego” are fuzzy linguistic concepts, not pointing to anything tangible or delineated.

“Don’t create things which are not there” https://isha.sadhguru.org/en/wisdom/article/what-is-ego

Post reply on HN