Earlier quoted context omitted.
Deepseek isn't philanthropy, it's a hedgefund trying to short the western AI market by saying "hey we can do 90% of they can (arguably better at a density metric) for a 1/10th of the cost" it's my theory at least, the Hindenburg Research of AI
Widely held belief in investor circles is that the Chinese government has a goal to deflate the US AI bubble and Deepseek is part of the plan to achieve that goal.
The gap between open weights LLMs and closed source LLMs
211–220 of 264 posts
Re: The gap between open weights LLMs and closed source LLMs
#212IMHO, the biggest problem with the future of open weights models is that currently, open weights models are the result of philanthropy by some private org. (e.g. DeepSeek). The spigot can be turned off at any time. Until there's some sort of "community owned hardware", open weights models are always at risk of being discontinued.
I am the original author of the post - thanks for reading it! I think the future of open weights models will be similar to fabless chip design companies. There will be companies that can train models and they will licence those models to inference companies that manage the APIs. The inference companies need much less capital and the training companies dont need to divert resources from training to inference. Some of…
I think at some point, Deepseek or Z or other AI training companies will sell their models for fees. I can imagine buying an LLM model for $499 one-time payment for personal use. Maybe buying software and owning it will come back. Some will make you subscribe so you get the latest models as they release them.
Of course, they will also license their models to inference providers like you said.
Re: The gap between open weights LLMs and closed source LLMs
#213Re: The gap between open weights LLMs and closed source LLMs
#214Earlier quoted context omitted.
Yeah, but the biggest plus for open models is that they can never be taken away. In other words, whatever capabilities they reach (even if there will never be another model), those stay forever. That can't be said for API-based models where a provider can sunset models whenever they feel like (i.e. gpt5-mini will soon be gone, and replaced by a more expensive 5.4-mini, same for goog, etc). And there will always be in…
> they can never be taken away Your right to 3d print whatever you want is about to be taken away (in California). What software you can run on your computer can already be restricted. Absolutely everything can be taken away. The simplest way to remove open models is probably to declare them a tool that terrorists could use. Crazy? Yes, the world is totally crazy these days.
Re: The gap between open weights LLMs and closed source LLMs
#215Earlier quoted context omitted.
Believe me, if the government wants to stop you from having access to something like that, they could do it. Just give people some incentive to report you and make really harsh punishments and everyone will be thinking really hard about how bad they want have access.
Fun fact: Hacker News is canonically banned in China, but I'm still talking here. There are plenty of techs to work around region block. The incentive to report somebody is comically called '50w' (500k CNY) and no one gives a shit about it in real life.
Re: The gap between open weights LLMs and closed source LLMs
#216Earlier quoted context omitted.
Widely held belief in investor circles is that the Chinese government has a goal to deflate the US AI bubble and Deepseek is part of the plan to achieve that goal.
Why would China care about deflating the US AI bubble? Why do we think there is a bubble for sure in 2025/2026? Why doesn't China also worry about their own AI bubble inside the country?
To weaken the stature of the USA on the global stage relative to themselves. Perhaps decrease US investment in AI and slow creation of some general AI superweapon I suppose.
Because the goal is to show that cheap chinese AI can compete with expensive USA AI, it's nessisarially a low-cost attack relative to the "damage" it could create.
>Why do we think there is a bubble for sure in 2025/2026?
Well that's the position that these chinese firms are trying to convince us of, and they can convince us by undercutting proprietary models in price/performance/openness.
In other words, we can be sure there is a bubble to the extent that open-weight models can successfully demonstrate that there is no moat.
>Why doesn't China also worry about their own AI bubble inside the country?
Because they haven't bet the farm on AI like the USA has.
Re: The gap between open weights LLMs and closed source LLMs
#217> Now is probably a good time to liquidate your pension, fly to a remote island somewhere, and live out the remaining 6 months or so of civilization in peace. > So maybe the open source apocalypse won’t happen yet. Sorry I wasn't at the last doomer meeting, when did we decide good open source models are a harbinger for the apocalypse?
Re: The gap between open weights LLMs and closed source LLMs
#218> Now is probably a good time to liquidate your pension, fly to a remote island somewhere, and live out the remaining 6 months or so of civilization in peace. > So maybe the open source apocalypse won’t happen yet. Sorry I wasn't at the last doomer meeting, when did we decide good open source models are a harbinger for the apocalypse?
Doomerism is at all time high People becoming more and more neurotic by the day
Re: The gap between open weights LLMs and closed source LLMs
#219Earlier quoted context omitted.
It's really unclear how much innovation DeepSeek has actually done, vs training on frontier model conversations.
Then you have no understanding of what DeepSeek has actually done. They publish their work openly: go have a look! Their architectural improvements are fascinating.