Earlier quoted context omitted.
Uhhh, I'm pretty sure DeepSeek shook the industry because of a 14x reduction in training cost, not inference cost. We also don't know the per-token cost for OpenAI and Anthropic models, but I would be highly surprised if it was significantly more expensive than open models anyone can use and run themselves. It's not like they're also not investing in inference research.
DeepSeek was trained with distillation. Any accurate estimate of training costs should include the training costs of the model that it was distilling.
Seriously, that claim was always completely disingenuous