Everyone who's thinking about the ramifications of prompt injection attacks now, please consider: This is really just a specific instance of the AI alignment problem. What about when the AI gets really smart, and tries to achieve certain goals in the world that are not what we want? How do make sure that these soon-to-be omnipresent models don't go off the rails when they have the power to make really big changes in…
AI as it is now is unverifiable. It's also organically behaving, and means it can be manipulated, be victim of social engineering, etc, like a human do. You cannot try to fool a single person a thousand time, but you can try to fool a thousand instance of AI.
Simple consider what you could accomplish with a phishing or scam email that works on 100% of the population.