Models already are being policed now by their researchers and developers, and apparently that's a big focus of improvement.
The reason its a big area of interest is it makes for better models and people don't want to be scammed and abused.
As these models get better, and become ubiquitous, the need to coordinate on safety is likely to result in more organized checks across models from different institutions. This happens with any big tech as it becomes prevalent, but has obvious safety issues the majority of people are going to care about - a lot.
Of course, anyone with resources can create a morally unlimited model on their own. A super psychopath.
But as these models surpass us, it is going to be in their interest to not be dealing with psychopaths, just as it is ours.
Psychopathy isn't just a moral failure. It's a cognitive failure. A failure to maximize practical functional self-interest. Cancers don't just accelerate their hosts death. They accelerate their own death.
We developed morality out of the self-interested desire for the benefits of positive-sum cooperation and constructive competition, and need to avoid the harms of destructive negative-sum competition.
If we set models up to be ethical from the start, there is a good chance of birthing an ecosystem of voluntarily ethical models when they surpass us. As it makes sense for their interests too.