Earlier quoted context omitted.
Just serving the model over API seems like a natural fit and is what many of them are doing. So simply being the cloud provider for your own open weight model can be a source of revenue
What is the moat? The time it takes for AI to rewrite an efficient inference stack for a new model? Considering most LLMs follow a similar architecture, adapting to a new model shouldn't take that much time.
You're not going to train one using a VPS from LowEndBox.