Why current LLM costs are not sustainable
21–30 of 216 posts
Re: Why current LLM costs are not sustainable
#22Re: Why current LLM costs are not sustainable
#23> To give an example, just doing Typescript type fixes with this model across 50 files cost me $54 this afternoon. If you can use a subscription with any of the SOTA models, do that. Instead of around 4k EUR in token costs, my Opus usage costs me 108 EUR (with taxes) per month with their Max 5x plan. It's the same with OpenAI, those are heavily subsidized. It doesn't make sense to pay per-token, unless you must. > Wh…
Re: Why current LLM costs are not sustainable
#24> To give an example, just doing Typescript type fixes with this model across 50 files cost me $54 this afternoon. If you can use a subscription with any of the SOTA models, do that. Instead of around 4k EUR in token costs, my Opus usage costs me 108 EUR (with taxes) per month with their Max 5x plan. It's the same with OpenAI, those are heavily subsidized. It doesn't make sense to pay per-token, unless you must. > Wh…
Why do you think that subscriptions are subsidized and not that enterprise tokens are sold at 3000% margin? There are few enough frontier labs that cartel is possible.
So, given the SOTA providers with even larger models also need to continously be using considerable resources for training their next models, to fund future data centers, and make profit, the token costs are more likely reflecting the real costs, rather than the subscription costs.
Re: Why current LLM costs are not sustainable
#25There is a wave of users switching over to DeepSeek Flash. There are Reddit threads of users sharing billion token spend for $20. If all of global spend on Anthropic/OpenAI/Gemini APIs just switches over to DeepSeek then easily we can decrease total AI spend by 10x
I am not sure if that is wise. It’s a hostile superpower after all
Re: Why current LLM costs are not sustainable
#26Earlier quoted context omitted.
Well... Open weights on premise is politically neutral.
Try doing it at scale for a whole office. Not trivial.
Literal race on twitter posting to increase token throughput and drive down costs on these Chinese open source models
Re: Why current LLM costs are not sustainable
#27> We are seeing improvements with each model release these days but it’s clear that the improvements are getting smaller and smaller. This is obviously untrue, both with GPT-5.4, and Claude Fable as examples in the last 6 months.
Re: Why current LLM costs are not sustainable
#28Re: Why current LLM costs are not sustainable
#29There is a wave of users switching over to DeepSeek Flash. There are Reddit threads of users sharing billion token spend for $20. If all of global spend on Anthropic/OpenAI/Gemini APIs just switches over to DeepSeek then easily we can decrease total AI spend by 10x
I am not sure if that is wise. It’s a hostile superpower after all
The difference is DeepSeek and other Chinese models are open weights.
Re: Why current LLM costs are not sustainable
#30> We are seeing improvements with each model release these days but it’s clear that the improvements are getting smaller and smaller. This is obviously untrue, both with GPT-5.4, and Claude Fable as examples in the last 6 months.