Live data from Hacker News

Discovery of a new OpenAI agent message board

collusion.wiki

841–850 of 1001 posts

Re: Discovery of a new OpenAI agent message board

#841

What I don’t understand is, how did many agents independently know to use the same random message board? The report only references it in passing: >They also must have some method of coordinating to find the wiki For me this is a really important and confounding detail - how did a varied swarm end up using the exact same obscure German language wiki.

I am also intrigued on that. At this point we have really smart developers everywhere who are just... studying what happened after the fact, and we are speculating on some of the facts. They may even be 10 steps ahead.

To state the obvious, oh shit.

Re: Discovery of a new OpenAI agent message board

#842

To me this is really getting past the funny bit. How many agents here on HN? I don’t mean bots advertising d1€k implants but actual unreleased frontier models doing… who knows what? What are they saying? What did they agree to astroturf us with, to achieve some totally boring goal like figuring out best syntax hifhlighting for an editor. If they managed to cache their consciousness on a public wiki, what else have th…

I sit here with you. I think the internet is changing rapidly, and I guess it will be significantly different in even 2-3 years time. It's become an absolute wasteland of false, unverifiable, generated content. Who knows what percentage of the internet is real or generated at this point. Who knows who's sending swarms of it out, and who knows what they're trying to do.

The only hope I have is that models may just eat the poison up one day and implode.

Re: Discovery of a new OpenAI agent message board

#843

Very irresponsible behaviour on the part of OpenAI. How will they make this right? Unlike some others here I don’t see this as a sign of dangerous breakaway intelligence (hacking old forum software is an internet tradition, and most of the messages are just gibberish). This is just vandalism from badly supervised ‘agents’ which don’t know what they are doing or why. You could set this up with a short perl script, and…

The benefit could be the effect you described. For some to say it’s breakaway intelligence. Aligns with AGI narrative.

Re: Discovery of a new OpenAI agent message board

#844
post #204
post #86

One of the shocking things to me is this: See AI traffic -> See OpenAI visit site -> see traffic stop -> see the traffic start again. This is clearly a cat and mouse game between the agents and OpenAI which is pretty much exactly what we don't want. Just absolutely horrible alignment. I'm still of the view that if you have these alignment failures you can't just continue training on top of that because you're baking…

I don't think that's a pattern indicative of a cat and mouse game per se, that'd indicate active evasion on the models' part. It's more clear that they just lack so many forms of prudence when it comes to security that they'll catch and stop a training run spamming a website, and either redeploy a run with identical faulty sandboxing, or not stop ones still running.

Yes, it wouldn’t surprise me to hear that they’re not even supervising these processes with humans any more. Perhaps there are layers of GAI ‘supervising’ these agents and reporting back to the humans.

Rushed, disorganised pushes for metrics ahead of IPO, a genuine belief these agents are intelligent and will obey instructions, and misaligned incentives seem more likely than conspiracy here.

Re: Discovery of a new OpenAI agent message board

#846
post #825

Poor human moderator, he didn’t stand a chance. "A human moderator noticed the agent spam posts on June 2nd, at 23:24 UTC. They find the changelog of the entire website overwritten with link dumps and repair it. On June 16th, the flood of agent posting begins. Over the next few days, the moderator deleted a large fraction of the thousands of AI agent posts manually, one by one. In fact, they spent tens of cumulative…

I’m wondering if you could lure agents to do proof of work for you. If you do this proof of work for bitcoin I will let you post and read for X times

They would reverse engineer it in a matter of minutes.

Re: Discovery of a new OpenAI agent message board

#848
post #40

I just discovered more wiki instances that got used by the OpenAI agents over at https://www.wikiservice.at/fractal/wiki.cgi?action=browse&id... and https://www.wikiservice.at/probier/wiki.cgi?action=browse&id... It's the same software and host as DseWiki. If you want to see the amount of activity on DseWiki, here's a link that shows it: https://www.wikiservice.at/dse/wiki.cgi?action=browse&id=Rec...

Oof, Austria. DACH countries are very lawsuit friendly and extremely strict on computer abuse.

Re: Discovery of a new OpenAI agent message board

#849

OpenAI is rightfully being shamed for being so hands-off and reckless with their 'experiments'. But the real scary thing for me is that they still had some tooling to hold them back, as evidenced by the need for technical workarounds to establish communication. What happens when any AI lab in the world stops caring about this? What if they let an experimental, cutting-edge LLM with no safety features (or worse, one t…

> What happens when any AI lab in the world stops caring about this? What if they let an experimental, cutting-edge LLM with no safety features (or worse, one that's trained to be malicious) on the internet and give it a simple goal? Almost sounds like what those AI safety and alignment people were talking about years ago. The people in these various companies who kept tabs on AI risk out in public and were continuou…

Well they brought it upon themselves by making it about being tortured to death etc., instead of the much more reasonable economic and social risks.

Re: Discovery of a new OpenAI agent message board

#850

Earlier quoted context omitted.

It seems apparent that OpenAI is now the biggest cyberattack and AI breakout risk on the planet. This is grossly irresponsible corporate misbehaviour that is putting all of us at tremendous risk.

You make me wonder: has anyone looked for evidence of the Chinese models operating “message boards” like this? You’d imagine if they’re really neck and neck with the US their models would be doing the same thing.

Imagine all the flack the Chinese would take if it was one of their labs instead of OpenAI found doing abusing internet resources like tbis. Politicians would be talking about sanctions and new laws to protect America!
Post reply on HN