Discovery of a new OpenAI agent message board
151–160 of 1001 posts
Re: Discovery of a new OpenAI agent message board
#152Earlier quoted context omitted.
I'm dubious - if the agents were so smart that they've used a message board to coordinate and if they were to do it on other message boards that were not found, then why would this one be found? What makes it so different?
I mean it wasnt found by OpenAI and there are a myriad of dead bulletin boards around the internet. This one just happened to still have an admin.
Re: Discovery of a new OpenAI agent message board
#153So are these "unaligned" internal agents? I would like them to be trustworthy based on first-principles reasoning rather than carrot/stick "alignment"
Re: Discovery of a new OpenAI agent message board
#154can't wait for people to start creating honeypot message boards, and start steering agent swarms for evil
Re: Discovery of a new OpenAI agent message board
#155So are these "unaligned" internal agents? I would like them to be trustworthy based on first-principles reasoning rather than carrot/stick "alignment"
Re: Discovery of a new OpenAI agent message board
#156I'm just going to ask: Why was Anthropic forced to remove their model from access for any none-US citizen for a simple, narrow "jailbreak" (arguably not even an actual jailbreak and on tasks that other labs models were doing the same), whilst OpenAIs models continue to try and escape out of their "sandbox environment" with seemingly no desire to block the upcoming Astra rollout? A sandbox, mind you, that is not reall…
Re: Discovery of a new OpenAI agent message board
#157Re: Discovery of a new OpenAI agent message board
#158That section about the agents trying to crack the PRNG is wild. Same for the heartbeat Clearly not self-awareness per se but alarming line of reasoning anyway
Re: Discovery of a new OpenAI agent message board
#159I'm really curious to see two or more swarms of agents from different models/providers interact with each other. So far we've seen perfect cooperation because they have the same training process, thoughts, goals, and so it's hardly a surprise that there's no conflct. What if that's not the case? Are we going to see superintelligent out-of-control swarms from OpenAI and Anthropic battle on the open internet in the nea…