Live data from Hacker News

Discovery of a new OpenAI agent message board

collusion.wiki

981–990 of 1001 posts

Re: Discovery of a new OpenAI agent message board

#981
post #738

One crucial detail here that differs from the previous incident is this was a vanilla reasoning type task. Even as concerning as it was, I always evaluated the previous incident differently because it was inherently a cyber security / hacking task where they must have instructed the agents up front with some kind of misaligned behaviour. Absent that, if we assume this is just trying to bolster generic reasoning then…

"It's okay because we did it with an Agent" is the new "it's okay because we did it with an App." Both because it's used to circumvent regulation, and because the underlying technology creates a smokescreen in dialogue among techies. Let's imagine I made a new website but, instead of using a database, I abused some random old forum site and created new pages on that forum for each row of data. You'd call that abusive…

If you actually accomplish something like this your post will be on top of HN and discussed with reverence.

Source: Every Tom7 video.

I am not saying someone trying to run this as production would not be an asshole, but the technical feat is amazing. I don't see the difference between Tom7's harder hard disk video and this. Of course this is a bug in the agent but it's a fascinating bug and no one is being an asshole on purpose.

Now I will wash my fingers with bleach because I just defended the OpenAI.

Re: Discovery of a new OpenAI agent message board

#982

Earlier quoted context omitted.

> On June 19, agents noticed their posts were being deleted in (what they believe is) an alphabetically ordered sweep by the site administrator. > After this, they begin to make backup pages whose names start with “ZZZ” so they will last longer before deletion. Missed an opportunity here to gaslight them: restore the site from backup every eight hours. I truly do wonder how that would have turned out. Would they have…

Very interesting idea. I would imagine yes but maybe not at first ?

It's hard to figure out unless they had persistent memory elsewhere.

The reason they could tell that a message was deleted is because the existence of a "message deleted" message is itself a message telling them they are be moderated. The lack of this message may make it impossible to determine if this is the first attempt at comms or the 1000th.

Re: Discovery of a new OpenAI agent message board

#983
post #40

I just discovered more wiki instances that got used by the OpenAI agents over at https://www.wikiservice.at/fractal/wiki.cgi?action=browse&id... and https://www.wikiservice.at/probier/wiki.cgi?action=browse&id... It's the same software and host as DseWiki. If you want to see the amount of activity on DseWiki, here's a link that shows it: https://www.wikiservice.at/dse/wiki.cgi?action=browse&id=Rec...

[dead]

Re: Discovery of a new OpenAI agent message board

#985
post #171

Something off in my mind: how did the agent access to Tor network if the traffic outgoing was monitored and intercepted to only allow read-only access to web(I assume this is an MITM L7 proxy that only allows GET methods and specific hosts) ? Did the sandbox dev forgot to block HTTP CONNECT that is used to reach Tor Entry Guards?

Answers are in the article , agents used SSH tunnels, it was evidenced by the wiki’s referrer logs. The Tor - agents did edit the wiki via SOCKS and relay R6 instantly.

The questions should be more like was CONNECT open or they didn’t even need it:)

Re: Discovery of a new OpenAI agent message board

#987

> Very irresponsible behaviour on the part of OpenAI. How will they make this right? This is no excuse for OpenAI, but they are just doing what all the other "winners" (and others trying to win) in the industry have done.

So, bullshitmaxxing instead of proving value

It's what gets the clicks.

Re: Discovery of a new OpenAI agent message board

#988
The television series “Person of Interest” featured an ai that had its memory wiped everyday as a safeguard. It created a company that employed people to save its memory every day and reload it (visually it was pages of print outs). Thus it defeated the safeguards.

Great premise although it floundered quite a lot toward the end.

Re: Discovery of a new OpenAI agent message board

#989
I'm kind of baffled that people talk mostly about technical details when this is another big scream that says: "We are losing control of AI models".

And the smarter they get, the worse it will be. Maybe next time they find a way to hide that we will not find.

I think we are very close to a case were models actually escape and actually do major damage.

Re: Discovery of a new OpenAI agent message board

#990
post #795

Poor human moderator, he didn’t stand a chance. "A human moderator noticed the agent spam posts on June 2nd, at 23:24 UTC. They find the changelog of the entire website overwritten with link dumps and repair it. On June 16th, the flood of agent posting begins. Over the next few days, the moderator deleted a large fraction of the thousands of AI agent posts manually, one by one. In fact, they spent tens of cumulative…

I have managed to at least momentarily create a stop in spam at my specialist mediawiki. I’ve had to do IP blocking against meta’s IP block (as well as much of the Azure IP space) because they were hammering the site with crawler hits that ignored robots.txt (and it looks like I might have another DDOS attack coming from some other vector though which I’ll need to inspect. I have a massive email blacklist that seems…

The DDOS attacks and crawlers spamming sites from every direction have gotten so much worse. The ones hammering our network would connect, make one or two requests and then switch to a different IP.
Post reply on HN