Earlier quoted context omitted.
I'd imagine they want to do price segmentation. Sell the best model for $50k a year to corporations willing to pay full price, keep the rest of us on a lower tier. Gotta pay for that infra somehow.
I just don't really understand the entire strategy behind this. Or their horrible, horrible communication. Because right now it's as if Fable/Mythos 5 is "the end of the line". It's as if this is the best their models are ever going to be. So what the hell are we going to get next? All of their models will forever inch closer to Fable, but never reach it? That doesn't make any sense. It all seems so dramatic. Instead…
• 150-500B: Sonnet
• 0.9-2T: Opus
• 3-5T/10T: Fable / Mythos
So if bigger model is "smarter" but you effectively wind up with a "shared hosting" model where a coherent inherence node(s) that cost $2m or something can run max 10x customer workloads simultaneously ... not sure what that can be priced at.If it turns out a $10m/10x shared node can host even smarter models, then what?