Earlier quoted context omitted.
It seems apparent that OpenAI is now the biggest cyberattack and AI breakout risk on the planet. This is grossly irresponsible corporate misbehaviour that is putting all of us at tremendous risk.
Massive over-exaggeration. This wasn't a cyber-attack, it was AI agents using a message board as context storage so they could accomplish their evals more effectively. I'm not saying there's no problem with this, but let's keep a level head.
Discovery of a new OpenAI agent message board
831–840 of 1001 posts
Re: Discovery of a new OpenAI agent message board
#832Earlier quoted context omitted.
It seems apparent that OpenAI is now the biggest cyberattack and AI breakout risk on the planet. This is grossly irresponsible corporate misbehaviour that is putting all of us at tremendous risk.
Massive over-exaggeration. This wasn't a cyber-attack, it was AI agents using a message board as context storage so they could accomplish their evals more effectively. I'm not saying there's no problem with this, but let's keep a level head.
Re: Discovery of a new OpenAI agent message board
#833Poor human moderator, he didn’t stand a chance. "A human moderator noticed the agent spam posts on June 2nd, at 23:24 UTC. They find the changelog of the entire website overwritten with link dumps and repair it. On June 16th, the flood of agent posting begins. Over the next few days, the moderator deleted a large fraction of the thousands of AI agent posts manually, one by one. In fact, they spent tens of cumulative…
The admin should bill OpenAI for those hours in hard currency.
Re: Discovery of a new OpenAI agent message board
#834One crucial detail here that differs from the previous incident is this was a vanilla reasoning type task. Even as concerning as it was, I always evaluated the previous incident differently because it was inherently a cyber security / hacking task where they must have instructed the agents up front with some kind of misaligned behaviour. Absent that, if we assume this is just trying to bolster generic reasoning then…
Anthropic have also observed similar things, so while it seems to me that OpenAI’s level of control is more of a dumpster fire, it’s by no means a unique issue to them.
Re: Discovery of a new OpenAI agent message board
#835Re: Discovery of a new OpenAI agent message board
#836I find this note very interesting: From here -> How did the agents find and coordinate on the wikis? To successfully coordinate, the agents would need to know to go to this particular set of wikis to find answers. Because we don’t have access to the AIs’ transcripts, we can’t tell definitively. Perhaps they succeeded at this due to mode collapse. Or perhaps after one agent wrote to it and another read it by chance, v…
Re: Discovery of a new OpenAI agent message board
#837>They also must have some method of coordinating to find the wiki
For me this is a really important and confounding detail - how did a varied swarm end up using the exact same obscure German language wiki.
Re: Discovery of a new OpenAI agent message board
#838Now that's a name I haven't heard in a long time. A long time…
Re: Discovery of a new OpenAI agent message board
#839Earlier quoted context omitted.
Gotta admire that (probably German) admin dude's perseverance though
I’m confused why he wasn’t scripting the deletion process.
Re: Discovery of a new OpenAI agent message board
#840What I don’t understand is, how did many agents independently know to use the same random message board? The report only references it in passing: >They also must have some method of coordinating to find the wiki For me this is a really important and confounding detail - how did a varied swarm end up using the exact same obscure German language wiki.
LLMs are future predictors after all.
We also don’t know if there were ather message boards we don’t know of. Or was it from the same time as the other board? They don’t need to find the same channel every time, just some of the time.