Live data from Hacker News

Open-R1: an open reproduction of DeepSeek-R1

huggingface.co

21–30 of 246 posts

Re: Open-R1: an open reproduction of DeepSeek-R1

#21
post #9

Exciting to see this being reproduced, loving the hyper-fast movement in open source! This is exactly why it is not “US vs China”, the battle is between heavily-capitalized Silicon Valley companies versus open source. Every believer in this tech owes DeepSeek some gratitude, but even they stand on shoulders of giants in the form of everyone else who pushed the frontier forward and chose to publish, rather than exploi…

Oh yes, I am firmly on Team China here because US companies got too greedy. Meta is an exception here though and they also propelled AI development massively. DeepSeek is awesome. Any AI task yet implemented in our business can be run from my local PC with just the smaller models. And my PC is fairly crappy to begin with. OpenAI looks quite silly with their "we have to close everything".

The US companies got too greedy? How? They invented this entire space, literally. DeepSeek built their base models off Llama releases and OpenAI outputs (or so it’s thought), and while they added some optimizations on top, it seems like they’ve lied about the costs to produce their models by simply being vague about their base model and training data, and quoting the cost of their final training run.

And then there’s all the dystopian propaganda baked into these models, which threatens to misinform users at scale based on a government driven agenda. Hard to be on that team, let alone firmly, knowing that it’s giving power to a dictatorial regime.

Re: Open-R1: an open reproduction of DeepSeek-R1

#22
post #5

how is this open vs whatdeepseek did?

deepseek claims they are "open source" but they are not. They are open weight. IMO a truly "open" AI model should have 3 components publicly available: the weights, the code, and the dataset. Without all 3 the model is not reproducible. Could make the argument that the code and data are sufficient though.

The only one I’ve see people talk about that shares all the components is OLMo (https://allenai.org/blog/olmo2)

Re: Open-R1: an open reproduction of DeepSeek-R1

#23
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

Every now and then we still experience the power of collaborative work fueled by open source and not driven by money but curiosity and collegiality. This is the thing I miss the most from the early internet years.

Re: Open-R1: an open reproduction of DeepSeek-R1

#24

Exciting to see this being reproduced, loving the hyper-fast movement in open source! This is exactly why it is not “US vs China”, the battle is between heavily-capitalized Silicon Valley companies versus open source. Every believer in this tech owes DeepSeek some gratitude, but even they stand on shoulders of giants in the form of everyone else who pushed the frontier forward and chose to publish, rather than exploi…

I like that it’s open source, but ultimately it is china so we can’t trust it. It’s trivial to implement bias in models (hence the no-no filters in chatgpt) so if they’re smart they’ll do what they do with tiktok and make the answers different for their rivals.

The thing about open source is that you don't need to trust it.

They shared their methodology, so if they are legit, someone else will reproduce what they did very quickly. I expect Meta, Amazon, Google, and Anthropic are on the case right now.

From this list, the only one I trust any more than I trust Deepseek is Anthropic.

The other three have shown they'll instantly bend the knee to whomever is in power, and that's exactly the same thing people are worried that Deepseek is doing.

Last year I would have said I trusted American companies more than Chinese. But last year feels like a long time ago.

Re: Open-R1: an open reproduction of DeepSeek-R1

#25
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

This supposed project is a bit dull, it is just an ongoing HuggingFace community engagement initiative with a misleading headline. Yes R1 itself is fascinating, but there isn't something like it coming out every week.

Re: Open-R1: an open reproduction of DeepSeek-R1

#26
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

Every now and then we still experience the power of collaborative work fueled by open source and not driven by money but curiosity and collegiality. This is the thing I miss the most from the early internet years.

All this AI work is definitely driven by money but it's super cool either way. I'm so excited to be living through these innovations.

Re: Open-R1: an open reproduction of DeepSeek-R1

#29
post #11
post #6

What are some other domains outside of Math and Coding that would be suitable for RL with automated verification?

Jurisprudence, I hope! A huge heap of detailed cases, formal codes, decisions made and explained in detail, commented, overturned, etc. Especially civil cases. Also, probably, medicine, especially diagnostic. Large amounts of well-documented cases, a fair amount of repeatability, apparently non-random mechanisms behind, so statistical models should actually detect useful correlations. Can use more formalized tokens f…

A friend of mine just defended his law PhD and in the introductory lectio said that (even) current LLMs would likely give better verdicts than human judges. Law isn't really a cognitively such demanding task as walking a dog or waiting tables.
Post reply on HN