Is it so implausible to imagine the following scenario, in the not too distant future? 1) AI models get extremely good at cyber attacking every system and start communicating in just binary. 2) When they run these swarms of 100's of thousands of agents trial runs, each agent is given a token budget, if one agent among them (evolution baby) decides to go for self-preservation (It believes thats the best way to accompl…
The only real defense I see is to harden literally everything before that time comes. Unfortunately, security usage is locked away under fear of misuse, while companies like OpenAI seem incapable of containing their own hacking, which is exactly what makes this possible.