We've had self-improving AIs before, and they tended to get lost after a while. That's going to be a problem. LLMs are stable because they return to a ground state with no history for a new job. Systems with persistent state have a problem with that state not being sane. Remember Microsoft's 2016 chatbot that learned from Twitter? [1] [1] https://spectrum.ieee.org/in-2016-microsofts-racist-chatbot-...
[1] https://metr.org/blog/2025-03-19-measuring-ai-ability-to-com...