I don’t understand who so many people seem surprised by this. 2023 was when we first started experimenting with linking two LLMs together to have them chat back and forth. Nothing really interesting there - they don’t care if they’re talking to an actual human or another LLM.
They’ve been able to send HTTP requests since the beginning as well, as long as any guardrails preventing this are disabled. Also not interesting.
Tell an LLM to do something and it tries its best to come up with the solution. They are not designed to go “I dunno”.
Why does it seem interesting that when you spawn 500 of them, they do the same things they’ve always done, just at a larger scale since there’s…more? Why are we treating this like a new discovered behavior? It’s how they’ve behaved from the beginning. It only required an organization reckless enough to try it at scale and without safeguards, and that’s something OAI excels in.