Live data from Hacker News

Open-R1: an open reproduction of DeepSeek-R1

huggingface.co

71–80 of 246 posts

Re: Open-R1: an open reproduction of DeepSeek-R1

#71
post #9

Earlier quoted context omitted.

Oh yes, I am firmly on Team China here because US companies got too greedy. Meta is an exception here though and they also propelled AI development massively. DeepSeek is awesome. Any AI task yet implemented in our business can be run from my local PC with just the smaller models. And my PC is fairly crappy to begin with. OpenAI looks quite silly with their "we have to close everything".

The US companies got too greedy? How? They invented this entire space, literally. DeepSeek built their base models off Llama releases and OpenAI outputs (or so it’s thought), and while they added some optimizations on top, it seems like they’ve lied about the costs to produce their models by simply being vague about their base model and training data, and quoting the cost of their final training run. And then there’s…

> DeepSeek built their base models off Llama releases and OpenAI outputs

Those models are also trained on data that was ignoring licenses / copyrighted content.

Re: Open-R1: an open reproduction of DeepSeek-R1

#72
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

No. My memory of the advent of the WWW was a sidebar in PC Magazine in Nov 1993 with an FTP link to download the NCSA Mosaic browser. It was a wow! moment to visit the few sites that existed. But nothing like this. What we’re seeing now is generating vastly more interest and excitement. It’s more akin to the 1999 dotcom bubble, but with far more impact and reach.

Re: Open-R1: an open reproduction of DeepSeek-R1

#73
post #52

Exciting to see this being reproduced, loving the hyper-fast movement in open source! This is exactly why it is not “US vs China”, the battle is between heavily-capitalized Silicon Valley companies versus open source. Every believer in this tech owes DeepSeek some gratitude, but even they stand on shoulders of giants in the form of everyone else who pushed the frontier forward and chose to publish, rather than exploi…

Open source has pragmatic merits and I love the culture. But I don't like associating it with a moral high ground because it doesn't charge you money. By this standard, we should also ask Intel/AMD to open-source their CPUs, video game studios to open-source their code and artifacts, and Google/Amazon for their search engine and infrastructure. Not all business sectors can afford to sustain with the Open Source model…

> By this standard, we should also ask Intel/AMD to open-source their CPUs, video game studios to open-source their code and artifacts, and Google/Amazon for their search engine and infrastructure.

You aren't? You're weird.

Re: Open-R1: an open reproduction of DeepSeek-R1

#75
post #64
post #56

Earlier quoted context omitted.

Anthropic was the only company in that list not to have paid for their CEO or founder to attend the inaugeration of the current ruler of the exectuive branch of the US a week ago.

That makes sense, but why trust them more than DeepSeek?

Because, from my non-expert but reasonably well informed understanding of world affairs, there's a very high chance of a Chinese company in an important area like AI having to bend the knee to Xi Jinping pretty much exactly as the American companies are doing with Trump.

Re: Open-R1: an open reproduction of DeepSeek-R1

#76
post #61

Earlier quoted context omitted.

Deep mind has one set of censorship, OpenAI another, anything musk does a third It’s all “massaged”

I don’t think “Taiwan is China” is the same kind of massaging as not telling people how to make napalm… What a weird thing to equate, though!

It is though. Western AI tries to hide information like that with the justification of safety as well as things that might be offensive to current popular beliefs. Chinese AI presumably says Taiwan is China to help get more people on side for a possible future invasion. Propaganda does work - look at how many people think Donbas is still Ukraine and Israel is still Palestine.

Re: Open-R1: an open reproduction of DeepSeek-R1

#77
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

the web was fascinating every second. You could click on a link without having ANY idea what you would land on. The overall quality was very poor, but it was thrilling.

A bit like indie cinema.

Re: Open-R1: an open reproduction of DeepSeek-R1

#78
post #63

Earlier quoted context omitted.

maybe it's a generational difference ? I personally feel burned out by all the generetive AI stuff, the internet was already ruined with bots, and now generitive AI took the garbage to the next level. very far from exciting or even "right"

It’s still better that blockchain-this, crypto-that…

Crypto is a disease of the skin; AI is a disease of the heart. – Chiang Kai Shek

Re: Open-R1: an open reproduction of DeepSeek-R1

#79

Earlier quoted context omitted.

I don’t think “Taiwan is China” is the same kind of massaging as not telling people how to make napalm… What a weird thing to equate, though!

It is though. Western AI tries to hide information like that with the justification of safety as well as things that might be offensive to current popular beliefs. Chinese AI presumably says Taiwan is China to help get more people on side for a possible future invasion. Propaganda does work - look at how many people think Donbas is still Ukraine and Israel is still Palestine.

The difference is that in China the info isn’t available without use of Western content, due to the totalitarian control over media, whereas in the West, information is pretty trivially available, even if the big companies keep it off of their platforms.

And sure ignorance is prevalent, but even GPT4 will tell me Donbas is still Ukraine, for instance. What a strange example to use, though!

Re: Open-R1: an open reproduction of DeepSeek-R1

#80
post #17

Given DeepSeek's open philosophy I wonder what their response is to simply being asked for access to the code and data that this project intends to recreate?

No company ever will disclose data due it would open endless liability.

Exactly. Meta won't do it for the same reason. Liability alone, imagine all the copyright lawsuits...

Secondly the dataset for now has a lot of competitive advantage.

In a way it seems like a good thing that AI giants compete on methodology now.

Post reply on HN