Earlier quoted context omitted.
You're unlikely to ever be priced out of tokens, at least if you'd be willing to settle for a model closer to Sonnet 4.5. That level of model certainly isn't as efficient as Fable 5, but it can crank out CRUD apps and other consulting mainstays quite well, with some supervision. To give you an example of model in this class, the DeepSeek V4 Flash preview is a 284B A13B model, with a native quantization mixing 4 bit a…
> You can easily run it on an RTX Pro 6000 Sure, and people can just build their own dropbox for like $250 too. How many people will bother though
Or you could just take your credit card and spend $20 on credits at https://openrouter.ai/deepseek/deepseek-v4-flash. I'm not sure that I could manage to spend even a $1/day at those rates.
My larger point is that while frontier tokens are a near-monopoly and who knows what they really cost, many real-world workflows can be run using commodity tokens, or even served in-house by anyone who can afford to hire US or EU programmers. And if you're willing to settle for what would have been a state-of-the-art coding model in October 2025, commodity tokens are close to free.