I really believe that in the near-term future we will run our LLMs in hardware, not in software. Hardwire a capable model into a device the size of a graphics card, embed it into a laptop, and you have something that uses less power, does faster inference, doesn't require additional CPU or memory, doesn't cost a monthly fee, and will probably eventually be available for under a (few) hundred bucks.
Why current LLM costs are not sustainable
11–20 of 216 posts
Re: Why current LLM costs are not sustainable
#12There is a wave of users switching over to DeepSeek Flash. There are Reddit threads of users sharing billion token spend for $20. If all of global spend on Anthropic/OpenAI/Gemini APIs just switches over to DeepSeek then easily we can decrease total AI spend by 10x
I am not sure if that is wise. It’s a hostile superpower after all
Re: Why current LLM costs are not sustainable
#13I think companies will fire 5-10% of people and convert them to token budget. I also believe that before any real companies are running these models locally, they will already have some kind of agentic layer. With the current frontier model lab progress, i do not see any real company which makes real money, running local models. Running local models is easy for me, for sure not that easy for any company. Your DC need…
That's the only way I can see frontier labs charging high enough to sustain the cash flow needed to operate as racing to the bottom is not possible for them.
It is interesting to think whether this is another "Cambrian" era like the smartphone OSes when you had Symbian, Android, iOs, Windows Mobile and so many others competing.
Re: Why current LLM costs are not sustainable
#14Re: Why current LLM costs are not sustainable
#15There is a wave of users switching over to DeepSeek Flash. There are Reddit threads of users sharing billion token spend for $20. If all of global spend on Anthropic/OpenAI/Gemini APIs just switches over to DeepSeek then easily we can decrease total AI spend by 10x
Re: Why current LLM costs are not sustainable
#16Re: Why current LLM costs are not sustainable
#17Re: Why current LLM costs are not sustainable
#18There is a wave of users switching over to DeepSeek Flash. There are Reddit threads of users sharing billion token spend for $20. If all of global spend on Anthropic/OpenAI/Gemini APIs just switches over to DeepSeek then easily we can decrease total AI spend by 10x
I am not sure if that is wise. It’s a hostile superpower after all
Re: Why current LLM costs are not sustainable
#19This is obviously untrue, both with GPT-5.4, and Claude Fable as examples in the last 6 months.
Re: Why current LLM costs are not sustainable
#20Earlier quoted context omitted.
Well... Open weights on premise is politically neutral.
Try doing it at scale for a whole office. Not trivial.