Is this what the Web was like in the beginning? Something exciting and fascinating every week?
Open-R1: an open reproduction of DeepSeek-R1
51–60 of 246 posts
Re: Open-R1: an open reproduction of DeepSeek-R1
#52Exciting to see this being reproduced, loving the hyper-fast movement in open source! This is exactly why it is not “US vs China”, the battle is between heavily-capitalized Silicon Valley companies versus open source. Every believer in this tech owes DeepSeek some gratitude, but even they stand on shoulders of giants in the form of everyone else who pushed the frontier forward and chose to publish, rather than exploi…
Re: Open-R1: an open reproduction of DeepSeek-R1
#53Earlier quoted context omitted.
deepseek claims they are "open source" but they are not. They are open weight. IMO a truly "open" AI model should have 3 components publicly available: the weights, the code, and the dataset. Without all 3 the model is not reproducible. Could make the argument that the code and data are sufficient though.
No one will release the dataset because we all know it is gathered through dodgy means.
Re: Open-R1: an open reproduction of DeepSeek-R1
#54Earlier quoted context omitted.
Just a decade and a half as it turns out! (though things definitely felt dizzyingly fast back then - think Google was launched just 5 years after HTML) - HTML first released in 1993 - AJAX in 1999 - Websocket first proposed in 2008 https://en.wikipedia.org/wiki/HTML https://en.wikipedia.org/wiki/Ajax_(programming) https://en.wikipedia.org/wiki/WebSocket
ActiveX XMLHTTP might have been released in 99, but it didn’t see any sort of real wider usage until 2004, 2005. I’d suggest its usage was really kickstarted when jQuery 1.0 launched in 2006 and standardised the interface to a simple API.
Re: Open-R1: an open reproduction of DeepSeek-R1
#55What are some other domains outside of Math and Coding that would be suitable for RL with automated verification?
Re: Open-R1: an open reproduction of DeepSeek-R1
#56Earlier quoted context omitted.
The thing about open source is that you don't need to trust it. They shared their methodology, so if they are legit, someone else will reproduce what they did very quickly. I expect Meta, Amazon, Google, and Anthropic are on the case right now. From this list, the only one I trust any more than I trust Deepseek is Anthropic. The other three have shown they'll instantly bend the knee to whomever is in power, and that'…
> the only one I trust any more than I trust Deepseek is Anthropic Will you kindly explain why? Thank you.
Re: Open-R1: an open reproduction of DeepSeek-R1
#57Earlier quoted context omitted.
Oh yes, I am firmly on Team China here because US companies got too greedy. Meta is an exception here though and they also propelled AI development massively. DeepSeek is awesome. Any AI task yet implemented in our business can be run from my local PC with just the smaller models. And my PC is fairly crappy to begin with. OpenAI looks quite silly with their "we have to close everything".
The US companies got too greedy? How? They invented this entire space, literally. DeepSeek built their base models off Llama releases and OpenAI outputs (or so it’s thought), and while they added some optimizations on top, it seems like they’ve lied about the costs to produce their models by simply being vague about their base model and training data, and quoting the cost of their final training run. And then there’s…
Re: Open-R1: an open reproduction of DeepSeek-R1
#58Earlier quoted context omitted.
I still remember when they came out with the 3d dancing baby, blew my mind much more than deepseek, it even played air guitar!
Rounded corners, anyone? Who remembers LISP being new?
Also have SMALLTALK-80 book with it's railtrack diagrams of syntax on the inside covers.
What's really interesting about the AI mania we're in is that no one has shown that what we have now will get to AGI and how. We have great models that simulate reasoning, but how close are they?
How do we measure their quality? Benchmarks? Tooling?
Re: Open-R1: an open reproduction of DeepSeek-R1
#59Earlier quoted context omitted.
This supposed project is a bit dull, it is just an ongoing HuggingFace community engagement initiative with a misleading headline. Yes R1 itself is fascinating, but there isn't something like it coming out every week.
You mean not like in the last 1-2 months
Re: Open-R1: an open reproduction of DeepSeek-R1
#60how is this open vs whatdeepseek did?
deepseek claims they are "open source" but they are not. They are open weight. IMO a truly "open" AI model should have 3 components publicly available: the weights, the code, and the dataset. Without all 3 the model is not reproducible. Could make the argument that the code and data are sufficient though.
DeepSeek's gifts to the world of its open weights, public research and OSS code of its SOTA models are all any reasonable person should expect given no organization is going to release their dataset and open themselves up to criticism and legal exposure.
You shouldn't expect to any to see datasets behind any SOTA models until they're able to be synthetically generated from larger models. Models only trained on sanctioned "public" datasets are not going to perform as well which makes them a lot less interesting and practically useful.
Yes it would be great for their to be open models containing original datasets and a working pipeline to recreate models from scratch. But when few people would even have the resources to train the models and the huge training costs just result in worse performing models, it's only academically interesting to a few research labs.
Open model releases should be celebrated, not criticized with unreasonable nitpicking and expectations that serves no useful purpose other than discouraging future open releases. When the norm is for Open Models to include their datasets, we can start criticizing those that don't, but until then be gracious that they're contributing anything at all.