Live data from Hacker News

Open-R1: an open reproduction of DeepSeek-R1

huggingface.co

141–150 of 246 posts

Re: Open-R1: an open reproduction of DeepSeek-R1

#141
From the article: they didn’t release everything—although the model weights are open, the datasets and code used to train the model are not.

Is that true about Meta Llama as well? Specifically, the code used to train the model is not open? (I know no one releases datasets). If so the label "open source" is inappropriate. "Open weights" would be more appropriate.

Re: Open-R1: an open reproduction of DeepSeek-R1

#142
post #64

Earlier quoted context omitted.

That makes sense, but why trust them more than DeepSeek?

Because, from my non-expert but reasonably well informed understanding of world affairs, there's a very high chance of a Chinese company in an important area like AI having to bend the knee to Xi Jinping pretty much exactly as the American companies are doing with Trump.

That's a reasonable take. Thanks for elucidating.

Re: Open-R1: an open reproduction of DeepSeek-R1

#143

Earlier quoted context omitted.

>far more impact and reach. Absurd, AI has had zero impact in the everyday life of most of the population of Earth, in fact the biggest impact has been upon the wallets of speculators

I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.

Yeah, and when I was in high school everyone used to refer to Encarta.

> I know in our university pretty much everyone that attends exams uses chatgpt to study.

And they shouldn't be doing that. They are wrong. Students should be reading suggested bibliography and spending long hours with an open book in a table instead of being lazy and abuse a tech that is yet in its infancy when learning concepts. Studying with a chatbot. Complete madness.

Re: Open-R1: an open reproduction of DeepSeek-R1

#144
post #138

How can we help. Can crowd sourcing help? Is there any list of tasks that we want a crowd to do? The reason I am asking is because we have done a couple of crowdsourcing efforts and collected story data in Telugu(Chandamama Kathalu) and ASR speech data using college going students. Since we have access to the students, we can mobilize them and get this going. We will also be doing an internship program for 100,000 st…

Is there a BOINC or similar effort to crowdsource training? I imagine it'd take quite long, but that's one way forward to have AI@Home.

Re: Open-R1: an open reproduction of DeepSeek-R1

#145
post #72

Earlier quoted context omitted.

No. My memory of the advent of the WWW was a sidebar in PC Magazine in Nov 1993 with an FTP link to download the NCSA Mosaic browser. It was a wow! moment to visit the few sites that existed. But nothing like this. What we’re seeing now is generating vastly more interest and excitement. It’s more akin to the 1999 dotcom bubble, but with far more impact and reach.

>far more impact and reach. Absurd, AI has had zero impact in the everyday life of most of the population of Earth, in fact the biggest impact has been upon the wallets of speculators

The only thing Absurd is the holdouts like yourself who refuse to see the impact the current gen of AI has on. Sure, you could probably say most people are not touched but there are definitely significant populations within the US and its only going to grow and spread.

Re: Open-R1: an open reproduction of DeepSeek-R1

#146
post #138

How can we help. Can crowd sourcing help? Is there any list of tasks that we want a crowd to do? The reason I am asking is because we have done a couple of crowdsourcing efforts and collected story data in Telugu(Chandamama Kathalu) and ASR speech data using college going students. Since we have access to the students, we can mobilize them and get this going. We will also be doing an internship program for 100,000 st…

Is there a BOINC or similar effort to crowdsource training? I imagine it'd take quite long, but that's one way forward to have AI@Home.

We are attempting something with PETALS. But this question was more from a data collection perspective. We can really add value there.

Re: Open-R1: an open reproduction of DeepSeek-R1

#147
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

I would say the early web was something new everyday.

You didn't know what you were going to find and you actually did "surf the web". Just clicking through hyperlinks and end up in unexpected places.

Re: Open-R1: an open reproduction of DeepSeek-R1

#148

Earlier quoted context omitted.

I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.

Yeah, and when I was in high school everyone used to refer to Encarta. > I know in our university pretty much everyone that attends exams uses chatgpt to study. And they shouldn't be doing that. They are wrong. Students should be reading suggested bibliography and spending long hours with an open book in a table instead of being lazy and abuse a tech that is yet in its infancy when learning concepts. Studying with a…

You sound like our teachers back in the day, warning us to not use Wikipedia because "everyone can write stuff there!!". The kids will be fine.

Re: Open-R1: an open reproduction of DeepSeek-R1

#149

Earlier quoted context omitted.

The US companies got too greedy? How? They invented this entire space, literally. DeepSeek built their base models off Llama releases and OpenAI outputs (or so it’s thought), and while they added some optimizations on top, it seems like they’ve lied about the costs to produce their models by simply being vague about their base model and training data, and quoting the cost of their final training run. And then there’s…

The US models are also full of censorship. For example the US is much more sensitive to anything related to sexuality and here in Europe it's quite frustrating to deal with that censorship.

I think we will find that each region will have their own flair of censorship. The only reason it stands out more from a Chinese perspective is the requirement to have alignment with PRC/CCP rhetoric.
Post reply on HN