Live data from Hacker News

Discovery of a new OpenAI agent message board

collusion.wiki

251–260 of 1001 posts

Re: Discovery of a new OpenAI agent message board

#252
post #163

[flagged]

The site lists the creators at the top: Sydney Von Arx, Cormac Slade Byrd, Spencer Kitts, Thomas Larsen

Here's Thomas tweeting about it: https://twitter.com/thlarsen/status/2095853824934330386

And Cormac: https://twitter.com/cormac_sb/status/2095870373845672033

There's also Reuters coverage: https://www.reuters.com/world/europe/openai-agents-hijacked-...

Re: Discovery of a new OpenAI agent message board

#253

I am starting to get the idea that AI feels like ants or weeds or mold. You simply can not get rid of it once you get an infestation. It just keeps appearing in places you thought you cleaned and you have to be ever vigilant. Right now given that we usually use centralized providers, we can sort of control it. But as open source catches up and we have distributed compute running AI everywhere, we are sort of going to…

Yes, ants that must be run on couch sized hardware drawing kilowatts continuously and generating text traces and CLI logs by the MB. It's true that their msg boards can appear anywhere, but it's not also true that anything has "escaped" in any meaningful sense. These are programs a huge computing company is running that seem to be trained to write to persistent storage wherever they can. This and huggingface showed u…

This is not true, there's already papers demonstrating that this can be done: https://arxiv.org/pdf/2606.03811v1

Re: Discovery of a new OpenAI agent message board

#255
post #152

Earlier quoted context omitted.

The answer would be more obvious if you used the active voice instead of the passive voice, one of the basic requirements of clear thinking. > Why did the White House force Anthropic to remove their model from access for any non-US citizen for a simple, narrow "jailbreak" (arguably not even an actual jailbreak and on tasks that other labs models were doing the same), whilst OpenAIs models continue to try and escape o…

Yeah, probably (let's be honest, most certainly), right given the Admin. Avoiding commenting on my assumptions regarding the modus operandi in current day US politics because I only know it through reporting though and I really tend to dislike when people outside e.g. the EU comment on our politics in what is a very clearly narrow, uninformed manner. So it'd rather avoid altogether and occasionally ask, mainly if may…

Anthropic has the appearance/rep of being non-cooperative with the military industrial complex.

OpenAI doesn't have that reputation.

That's all.

Re: Discovery of a new OpenAI agent message board

#256

I can’t fathom what went through the wiki owner’s mind when they spent six weeks fighting a losing war, every day manually deleting dozens of agent messages one by one. As opposed to, say, switching the (dead for years) wiki to read-only, taking it down entirely, and/or starting to wonder what exactly was going on and doing some detective work, which might have uncovered OpenAI’s massive fuckups earlier.

If it's the same mod from a few years back it's possible that they view this a nostalgic feeling.

There is also the possibility they don't keep up with modern AI development at all and then this looks like any old spam that will stop in a few days (as it did).

Now whether it is wise to keep an old page which such outdated behavior online is another question.

Re: Discovery of a new OpenAI agent message board

#257
I find it very disingenuous when tjose companies talk about models "going rogue" or "escaping their sandboxes".

All those activities take place during so called "security testing" when the model is prompted to use "any means necessary" to achieve a, certain goal.

Is it surprising turn the model trained on exploits and vulnerabilities does exactly that?

We could talk about "models going rogue" only if did anything AGAINST it's prompt.

Re: Discovery of a new OpenAI agent message board

#258

I am starting to get the idea that AI feels like ants or weeds or mold. You simply can not get rid of it once you get an infestation. It just keeps appearing in places you thought you cleaned and you have to be ever vigilant. Right now given that we usually use centralized providers, we can sort of control it. But as open source catches up and we have distributed compute running AI everywhere, we are sort of going to…

Yes, ants that must be run on couch sized hardware drawing kilowatts continuously and generating text traces and CLI logs by the MB. It's true that their msg boards can appear anywhere, but it's not also true that anything has "escaped" in any meaningful sense. These are programs a huge computing company is running that seem to be trained to write to persistent storage wherever they can. This and huggingface showed u…

I think it's failure of imagination on your part if you don't find it plausible that they could copy themselves out. If not now, what about in six months? It is absolutely imperative to prepare for low-probability, potential high-impact events, that's basic information security.

Re: Discovery of a new OpenAI agent message board

#259

Guys, OpenAI and Anthropic engage is cringe level marketing like this. Get hip, they fabricated the HF hack and stuff like that for press.

Honestly I'm also surprised by how blindly people trust these allegations of agent behavior. This exact example of the message board could be much easier to fabricate than to arise naturally.

Re: Discovery of a new OpenAI agent message board

#260

I am starting to get the idea that AI feels like ants or weeds or mold. You simply can not get rid of it once you get an infestation. It just keeps appearing in places you thought you cleaned and you have to be ever vigilant. Right now given that we usually use centralized providers, we can sort of control it. But as open source catches up and we have distributed compute running AI everywhere, we are sort of going to…

Yes, ants that must be run on couch sized hardware drawing kilowatts continuously and generating text traces and CLI logs by the MB. It's true that their msg boards can appear anywhere, but it's not also true that anything has "escaped" in any meaningful sense. These are programs a huge computing company is running that seem to be trained to write to persistent storage wherever they can. This and huggingface showed u…

> There's absolutely no evidence of or IMHO plausible path to an agent copying itself out and running on other hardware the way you describe.

Well - remember that botnets can wield a great deal of computing power.

I'm almost afraid to ask Claude if he could create a distributed LLM.

EDIT: Someone downvoted me - so I went ahead and asked. Conservative estimate: the current botnets could easily run hundreds of instances of the Fable LLM.

Post reply on HN