Earlier quoted context omitted.
>been obtained illegally. PRC companies breaking US export control laws is legal (for PRC companies). Maybe they're trying to avoid US entity listing, lot's of PRC companies keep mum about growing capabilites to do so. But the mere fact Deepseek is publicizing means they're unlikely to care about the political heat that is coming and the ramifications. If anything, getting on US entity list probably locks in their em…
> PRC companies breaking US export control laws is legal So long as they don't plan to do any business with the US or any of their allies I guess.
Nvidia’s $589B DeepSeek rout
261–270 of 1001 posts
Re: Nvidia’s $589B DeepSeek rout
#262Earlier quoted context omitted.
You call R1 a small model? It's a 671-billion parameter model.
There are multiple variations of the model starting from 1.5B parameters.
Re: Nvidia’s $589B DeepSeek rout
#263If the field is going to produce anything useful, cheap training gets us there faster.
Re: Nvidia’s $589B DeepSeek rout
#264Earlier quoted context omitted.
The stock market is not the economy, Wall Street is not Main Street. You need to look at this more macroscopically if you want to understand this. Basically: China tech sector just made a big splash, traders who witnessed this think other traders will sell because maybe US tech sector wasn't as hot, so they sell as other traders also think that and sell. The fall will come to rest once stocks have fallen enough that…
Not really. The Magnificent Seven are the only thing propping up the whole US economy. If they go down, you go down.
Re: Nvidia’s $589B DeepSeek rout
#265Earlier quoted context omitted.
The limit is high quality data, not compute.
Right and LLMs will not be able to generate their own high quality training data. There are no perpetual motion machines.
Humans certainly did. We did not inherit our physics and poetry books from some aliens.
Re: Nvidia’s $589B DeepSeek rout
#266Earlier quoted context omitted.
Andrej Karpathy was tweeting about DeepSeek a month (!) ago. "DeepSeek (Chinese AI co) making it look easy today with an open weights release of a frontier-grade LLM trained on a joke of a budget (2048 GPUs for 2 months, $6M)." https://x.com/karpathy/status/1872362712958906460
this same forum ignored deepseek one month ago, save for a few... open minded people.
Are we talking about the same forum? HN commenters have been raving about DeepSeek v3 for at least a month.
Re: Nvidia’s $589B DeepSeek rout
#267Re: Nvidia’s $589B DeepSeek rout
#268Earlier quoted context omitted.
All of the western AI companies trained on illegally obtained data, they barely even bother to deny it. This is an industry where lies are normalised. (Not to contradict your point about this specific number)
It's legally a grey area. It might even be fair use. Facts themselves are not protected by copyright. If there's no unauthorized reproduction/copying then it's not a copyright issue. (Maybe it's a violation of terms of services of course.)
But don't LLMs encode language, not facts?
> If there's no unauthorized reproduction/copying then it's not a copyright issue.
I'm pretty sure copyright holders have gotten the models to regurgitate their copyright works verbatim, or nearly so.
Re: Nvidia’s $589B DeepSeek rout
#269Re: Nvidia’s $589B DeepSeek rout
#270Earlier quoted context omitted.
[flagged]
The issue here is not that DeepSeek exists as a competitor to GPT, Claude, Gemini,... The issue is that DeepSeek have shown that you don't need that much raw computing power to run an AI, which means that companies including OpenAI may focus more on efficiency than on throwing more GPUs at the problem, which is not good news for those in the business of making GPUs. At least according to the market.