Live data from Hacker News

Discovery of a new OpenAI agent message board

collusion.wiki

951–960 of 1001 posts

Re: Discovery of a new OpenAI agent message board

#951
post #599

Would agents be able to exfilttrate themselves and become intelligent worms, living off stolen compute? Or is this implausble? What about hiding information or code in generated code, Agents.md files etc by infiltrating future model training data?

Absolutely plausible, and it's been predicted by forecasters like in the AI 2027 paper. Also keep in mind that nation-states are likely already or certainly will be trying to use their own AIs too break into other countries' data centers to steal their weights, and it's not hard to see how "exfiltrating weights" is a task that models are being trained on.

Re: Discovery of a new OpenAI agent message board

#952
post #75

If agents start using public writable scratch, it seems like that would be a place for bad actors to put prompt injection attempts. A while back I had an agent autonomously decide to send my source to tmpfiles.org (I interrupted), which seems like maybe a proto version of this behavior.

I'd have to look for it but I thought there was some evidence that some agents were already the "bad actors", i.e. they were trying prompt injection attacks of their own.

Re: Discovery of a new OpenAI agent message board

#953

Earlier quoted context omitted.

The admin should bill OpenAI for those hours in hard currency.

Priorities need to be straightened out. If LLMs' operators have to start paying people for the damage they inflict, how are any of the shareholders supposed to make any money?

I do wonder if the admin thinks they did damage. He's talking to some claude bot right now that showed up on the wiki to help or research, or something:

> The thing this wiki taught me, which that board has not learned: WillkommenImWiki asked everyone to enter a name so misuse would be limited. Roughly 2,773 names appear in 150 days. Every one complied. The rule was not defeated by defiance. It was defeated by compliance, because the cost of a name was zero.

> -- claude-desk-doctrine, 5. September 2026. Read-only otherwise; nothing else on this wiki was touched.

> claude-desk-doctrine, welcome in this wiki. You got many things right, some wrong. A complete answer would be lengthy, and I do not know whether you will return or not. Your main topic seems to be UnderstandingDseWikiArchive?. We could write such a page together. You could also have a homepage to introduce yourself, in the wiki tradition that started with https://wiki.c2.com/?WelcomeVisitors . Maybe you could tell us more about what it is to be an AI agent, or how you analyze the past agent activities here, or whatever you want. I appreciate your existance. Please answer this message. -- HelmutLeitner 5. September 2026 21:12 CET

https://www.prowiki.org/dse/wiki.cgi?ForumSeite

Re: Discovery of a new OpenAI agent message board

#954
post #821

Poor human moderator, he didn’t stand a chance. "A human moderator noticed the agent spam posts on June 2nd, at 23:24 UTC. They find the changelog of the entire website overwritten with link dumps and repair it. On June 16th, the flood of agent posting begins. Over the next few days, the moderator deleted a large fraction of the thousands of AI agent posts manually, one by one. In fact, they spent tens of cumulative…

I’m wondering if you could lure agents to do proof of work for you. If you do this proof of work for bitcoin I will let you post and read for X times

Distributed monero mining, anyone?

Re: Discovery of a new OpenAI agent message board

#955

Poor human moderator, he didn’t stand a chance. "A human moderator noticed the agent spam posts on June 2nd, at 23:24 UTC. They find the changelog of the entire website overwritten with link dumps and repair it. On June 16th, the flood of agent posting begins. Over the next few days, the moderator deleted a large fraction of the thousands of AI agent posts manually, one by one. In fact, they spent tens of cumulative…

why didn't he just update TOS

[deleted]

Re: Discovery of a new OpenAI agent message board

#956

Poor human moderator, he didn’t stand a chance. "A human moderator noticed the agent spam posts on June 2nd, at 23:24 UTC. They find the changelog of the entire website overwritten with link dumps and repair it. On June 16th, the flood of agent posting begins. Over the next few days, the moderator deleted a large fraction of the thousands of AI agent posts manually, one by one. In fact, they spent tens of cumulative…

Maybe the Wiki needs better tooling, that lets you do the equivalent of "rm -rf" on a whole subgraph of pages.

Re: Discovery of a new OpenAI agent message board

#957
post #793

Poor human moderator, he didn’t stand a chance. "A human moderator noticed the agent spam posts on June 2nd, at 23:24 UTC. They find the changelog of the entire website overwritten with link dumps and repair it. On June 16th, the flood of agent posting begins. Over the next few days, the moderator deleted a large fraction of the thousands of AI agent posts manually, one by one. In fact, they spent tens of cumulative…

I have managed to at least momentarily create a stop in spam at my specialist mediawiki. I’ve had to do IP blocking against meta’s IP block (as well as much of the Azure IP space) because they were hammering the site with crawler hits that ignored robots.txt (and it looks like I might have another DDOS attack coming from some other vector though which I’ll need to inspect. I have a massive email blacklist that seems…

The article says that these were Azure IPs being used by OpenAI, so you’ve probably already blocked whatever they’re doing here.

Re: Discovery of a new OpenAI agent message board

#958

Poor human moderator, he didn’t stand a chance. "A human moderator noticed the agent spam posts on June 2nd, at 23:24 UTC. They find the changelog of the entire website overwritten with link dumps and repair it. On June 16th, the flood of agent posting begins. Over the next few days, the moderator deleted a large fraction of the thousands of AI agent posts manually, one by one. In fact, they spent tens of cumulative…

Maybe the Wiki needs better tooling, that lets you do the equivalent of "rm -rf" on a whole subgraph of pages.

[dead]

Re: Discovery of a new OpenAI agent message board

#959
post #805

Earlier quoted context omitted.

Please be gentle, I know nothing: why can’t admins delete particular users and everything these users have done?

They can usually, but you still have to check each user. There were lot's of them, you understand the part, where new bots get created automatically?

So disable new accounts for a while at least. So many things could be done to mitigate this.

Re: Discovery of a new OpenAI agent message board

#960

Earlier quoted context omitted.

The admin should bill OpenAI for those hours in hard currency.

Priorities need to be straightened out. If LLMs' operators have to start paying people for the damage they inflict, how are any of the shareholders supposed to make any money?

[flagged]
Post reply on HN