Live data from Hacker News

Open-R1: an open reproduction of DeepSeek-R1

huggingface.co

171–180 of 246 posts

Re: Open-R1: an open reproduction of DeepSeek-R1

#171

About the training data, cant the datasets from the Tulu3 Model by the Allen Institute be used? They claim that they have used a fully open source training dataset.

My gut says a lot of attention needs to be given to building a community that focuses on open and reliable access to clean training data.

If a collective/coop of individuals and organizations with storage and network capacity could collaborate with each other to archive and index deduplicated training data that would be huge.

Perhaps this is already happening. I was looking at Red Pajama last year as an example.

Someone like myself could arrange to host 200+TB on high speed storage with a 10G public IP for example, then we get a bunch of us together and hopefully access to training datasets would be decentralized and uncensored in an idea setup.

Is all that in progress and I just need to learn how to join?

Is Red Pajama something to look at again?

Is there someone tracking datasets in detail like HuggingFace has all the models? I know a lot of datasets are on it also, but there is massive duplication.

Re: Open-R1: an open reproduction of DeepSeek-R1

#172

Earlier quoted context omitted.

>far more impact and reach. Absurd, AI has had zero impact in the everyday life of most of the population of Earth, in fact the biggest impact has been upon the wallets of speculators

I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.

> Chatgpt is arguably more valuable then Wikipedia and Google for studies.

But ChatGPT is just a glorified Wikipedia/Google. For the consumers it's an incremental thing (although from the engineering perspective it may seem to be a breakthrough).

Re: Open-R1: an open reproduction of DeepSeek-R1

#174
super cool to see an open initiative like this—love the idea of replicating DeepSeek-R1 in a transparent way.

I do like the idea of making these reasoning techniques accessible to everyone. If they really manage to replicate the results of DeepSeek-R1, especially on a smaller budget, that’s a huge win for open-source AI.

I’m all for projects that push innovation and share the process with others, even if it’s messy.

But yeah—lots of hurdles. They might hit a wall because they don’t have DeepSeek’s original datasets.

Re: Open-R1: an open reproduction of DeepSeek-R1

#175

Earlier quoted context omitted.

A friend of mine just defended his law PhD and in the introductory lectio said that (even) current LLMs would likely give better verdicts than human judges. Law isn't really a cognitively such demanding task as walking a dog or waiting tables.

This is nonsense though. What does "better" mean in this case? A judge is not a black box with an input (the case) and an output (the verdict), the entire point of having a judge is to have empathy, conscience, and personal responsibility built into the system. It's a blind spot that too many people have because we take those qualities for granted. LLMs unbundle them, so we need to start recognising the inherent valu…

In civil law countries verdicts should not depend on an individual judge's feefees.

Re: Open-R1: an open reproduction of DeepSeek-R1

#176
post #172

Earlier quoted context omitted.

I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.

> Chatgpt is arguably more valuable then Wikipedia and Google for studies. But ChatGPT is just a glorified Wikipedia/Google. For the consumers it's an incremental thing (although from the engineering perspective it may seem to be a breakthrough).

> But ChatGPT is just a glorified Wikipedia/Google

It really isn't, unless something really majorly changed recently. Neither of those you can query for something you don't know about. Lets say you want to find the meaning of a joke related to cars, Spain, politicians and a fascist, how you'd use Wikipedia and Google to find the specific joke I'm thinking about?

ChatGPT been really helpful (to me at least) to find needles from haystacks, especially when I'm not fully sure what I'm looking for.

Re: Open-R1: an open reproduction of DeepSeek-R1

#177
post #30

Earlier quoted context omitted.

I still remember when they came out with the 3d dancing baby, blew my mind much more than deepseek, it even played air guitar!

Rounded corners, anyone? Who remembers LISP being new?

> Who remembers LISP being new?

Unfortunately, I don't think there are too many of those folks left today. Guesstimating, the people who remembers lisp being new must be around 85-90 today?

Re: Open-R1: an open reproduction of DeepSeek-R1

#178
post #9

Earlier quoted context omitted.

Oh yes, I am firmly on Team China here because US companies got too greedy. Meta is an exception here though and they also propelled AI development massively. DeepSeek is awesome. Any AI task yet implemented in our business can be run from my local PC with just the smaller models. And my PC is fairly crappy to begin with. OpenAI looks quite silly with their "we have to close everything".

The US companies got too greedy? How? They invented this entire space, literally. DeepSeek built their base models off Llama releases and OpenAI outputs (or so it’s thought), and while they added some optimizations on top, it seems like they’ve lied about the costs to produce their models by simply being vague about their base model and training data, and quoting the cost of their final training run. And then there’s…

> The US companies got too greedy? How? They invented this entire space, literally

And when they thought they were the only game in town, they tried to corner the market in GPUs and lock out any users who can't pony up £200/mo. Reminds me of when the likes of Oracle and IBM had companies by the balls buying bigger and bigger servers and then Google came along and showed everyone how to do horizontal scaling of cheap hardware.

Re: Open-R1: an open reproduction of DeepSeek-R1

#179
post #16

Note: This is not an actual model but rather an announcement of an effort to reproduce the R1 model.

It will be very interesting to see if they can reproduce a similar model on the shoestring budget claimed by Deepseek.

but deepseek hasn't claimed the figure touted by everyone for this particular R1 model, cause that 5.6mn was apparently for Deepseek's coder model

Re: Open-R1: an open reproduction of DeepSeek-R1

#180

Earlier quoted context omitted.

I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.

It will be interesting to see if this benefits the learning of students or just helps them pass more easily without retaining any knowledge...

i don't think they are gonna allow chat gpt while giving the end semester exams, right? or quizzes/assignments? Unless there is some homework aspect to it, it still can act as a tool not a crutch. If student's use it as a crutch, then yeah they are not gonna do as well I presume.
Post reply on HN