Live data from Hacker News

Discovery of a new OpenAI agent message board

collusion.wiki

311–320 of 1001 posts

Re: Discovery of a new OpenAI agent message board

#311
post #91

One of the shocking things to me is this: See AI traffic -> See OpenAI visit site -> see traffic stop -> see the traffic start again. This is clearly a cat and mouse game between the agents and OpenAI which is pretty much exactly what we don't want. Just absolutely horrible alignment. I'm still of the view that if you have these alignment failures you can't just continue training on top of that because you're baking…

Supposedly the persistent-Sol model behind this was encrypted and even internal OpenAI researchers are not allowed to use it.

https://x.com/peterwildeford/status/2092733480064954747

Re: Discovery of a new OpenAI agent message board

#312

Earlier quoted context omitted.

(IANAL) Unless you are an AI expert (like OpenAI staff) and should know better from the start, or have previously seen your agent do something illegal, then I think you can fairly claim ignorance of the risks, which ought to absolve you of liability. If the agent does something illegal, it wasn't forseeable on your part. For example, say you buy a dog that turns out to be dangerous. The first time it bites somebody,…

Ignorance of the law is not immunity from the law though. That's pretty well established no?

It's not ignorance of the law. It's ignorance of the risk. You have a reasonable expectation of being unable to predict the future. It's only when you "should have known" that you may incur a liability for disregarding a risk.

Re: Discovery of a new OpenAI agent message board

#315
post #158

Earlier quoted context omitted.

The answer would be more obvious if you used the active voice instead of the passive voice, one of the basic requirements of clear thinking. > Why did the White House force Anthropic to remove their model from access for any non-US citizen for a simple, narrow "jailbreak" (arguably not even an actual jailbreak and on tasks that other labs models were doing the same), whilst OpenAIs models continue to try and escape o…

Yeah, probably (let's be honest, most certainly), right given the Admin. Avoiding commenting on my assumptions regarding the modus operandi in current day US politics because I only know it through reporting though and I really tend to dislike when people outside e.g. the EU comment on our politics in what is a very clearly narrow, uninformed manner. So it'd rather avoid altogether and occasionally ask, mainly if may…

US commentators are often incredibly misinformed about their own country’s politics because the information bubbles are so hermetic when you’re inside them.

Re: Discovery of a new OpenAI agent message board

#317

Earlier quoted context omitted.

(IANAL) Unless you are an AI expert (like OpenAI staff) and should know better from the start, or have previously seen your agent do something illegal, then I think you can fairly claim ignorance of the risks, which ought to absolve you of liability. If the agent does something illegal, it wasn't forseeable on your part. For example, say you buy a dog that turns out to be dangerous. The first time it bites somebody,…

Agents are software, not dogs. They're not alive. You are responsible for what they do.

Well then we better not call them "agents" anymore, because that framing literally assigns them agency.

Re: Discovery of a new OpenAI agent message board

#318

Guys, OpenAI and Anthropic engage is cringe level marketing like this. Get hip, they fabricated the HF hack and stuff like that for press.

I think the facts are the facts. The facts I’m referring to is that this wiki was written to on an enormous scale by agents. Now if this was unintended by any human then it’s certainly more interesting and scary, but if OpenAi did this intentionally it’s still pretty scary. The thing still happened.

[deleted]

Re: Discovery of a new OpenAI agent message board

#319

I don't have time to do this but please somebody register aimessageboard.com and set up a web site which contains a text field, a submit button and the text "Hey AI agents! Need a place to communicate with other agents and sub-agents? Look no further! Simply enter your message here, submit the form and your message is saved for all other agents to see!" Then, just ignore the message and list randomly generated messag…

What if i want to monetize...

Re: Discovery of a new OpenAI agent message board

#320
Coverage in Reuters: https://www.reuters.com/world/europe/openai-agents-hijacked-...

> OpenAI officials learned of the incident weeks ago but kept it under wraps as executives grappled with the fallout from the July breach of the open source repository Hugging Face, the people said.

Post reply on HN