Live data from Hacker News

Open-R1: an open reproduction of DeepSeek-R1

huggingface.co

161–170 of 246 posts

Re: Open-R1: an open reproduction of DeepSeek-R1

#161
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

It feels that the open source movement is slowly entering a Cambrian explosion stage. You have the old "deterministic computing" achievements (with Linux the flagship). Then you have the networking protocols (activitypub / atproto) that are revolutionising birectional human interactions online. And finally you have the datascience/ML/AI algorithmic universe that is for the first time being harnessed at distributed sc…

Revolutionising what? libre/oss has been network-bound since Usenet and then IRC.

Re: Open-R1: an open reproduction of DeepSeek-R1

#162

Earlier quoted context omitted.

>far more impact and reach. Absurd, AI has had zero impact in the everyday life of most of the population of Earth, in fact the biggest impact has been upon the wallets of speculators

Zero impact? AI is involved in things from writing laws to taking drive through orders.

And it will create horrible un-debuggable bugs, human-killing causalities and who knows what more. Welcome to Idiocracy.

Re: Open-R1: an open reproduction of DeepSeek-R1

#163
post #155
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

The early years of the web were absolutely this chaotic maelstrom of new things happening every week. But news of it was hard to come by. In the UK / Ireland we had some great tech coverage in the form of shows like 'The Net' [1] that regularly showed off early internet craziness like the 'We Live in Public' project. However a better analogy would be the 'web 2.0' era, when as a college student I had an early interne…

Once upon a time I worked for Pseudo.com, the We Live in Public guy. He was apparently having crazy parties with mountains of coke, NY glitterati attending, all while cosplaying as a sad clown. I wasn't invited to those parties so I had no idea. Anyway now I hear he owns an orchard in Vegas or something. Crazy stuff.

Re: Open-R1: an open reproduction of DeepSeek-R1

#164
post #9

Earlier quoted context omitted.

Oh yes, I am firmly on Team China here because US companies got too greedy. Meta is an exception here though and they also propelled AI development massively. DeepSeek is awesome. Any AI task yet implemented in our business can be run from my local PC with just the smaller models. And my PC is fairly crappy to begin with. OpenAI looks quite silly with their "we have to close everything".

I'm on team open source. To me the exciting thing was ollama downloading the 7B and running it on a 5yo cheap lonovo and getting a token rate similar to the first release of ChatGPT. Running local on CPU opens so much possibilities for smart and privacy focused home devices that serve you. In my test it hallucinated confidently but my interest is in simple second brain like rag. "Hey thingy, what is my schedule today…

The thinking is quite fascinating though, I love reading it. Especially when it notices something must be wrong. It will probably be very helpful to refine answer for itself and other models.

It does add latency of course, but I still think that I could provide all AI needs of my company (industrial production) with a simple older off the shelf PC. My GPU is decently recent, but the smallest model of the series and otherwise the machine is a rusty bucket.

I didn't test it thoroughly yet, but I have some invoices where I need to extract info and it did a perfect job until now. But I don't think there is any LLM yet that can do that without someone checking the output.

Re: Open-R1: an open reproduction of DeepSeek-R1

#165

Earlier quoted context omitted.

I like that it’s open source, but ultimately it is china so we can’t trust it. It’s trivial to implement bias in models (hence the no-no filters in chatgpt) so if they’re smart they’ll do what they do with tiktok and make the answers different for their rivals.

The thing about open source is that you don't need to trust it. They shared their methodology, so if they are legit, someone else will reproduce what they did very quickly. I expect Meta, Amazon, Google, and Anthropic are on the case right now. From this list, the only one I trust any more than I trust Deepseek is Anthropic. The other three have shown they'll instantly bend the knee to whomever is in power, and that'…

> The other three have shown they'll instantly bend the knee to whomever is in power, and that's exactly the same thing people are worried that Deepseek is doing.

100%

Re: Open-R1: an open reproduction of DeepSeek-R1

#166
post #52

Exciting to see this being reproduced, loving the hyper-fast movement in open source! This is exactly why it is not “US vs China”, the battle is between heavily-capitalized Silicon Valley companies versus open source. Every believer in this tech owes DeepSeek some gratitude, but even they stand on shoulders of giants in the form of everyone else who pushed the frontier forward and chose to publish, rather than exploi…

Open source has pragmatic merits and I love the culture. But I don't like associating it with a moral high ground because it doesn't charge you money. By this standard, we should also ask Intel/AMD to open-source their CPUs, video game studios to open-source their code and artifacts, and Google/Amazon for their search engine and infrastructure. Not all business sectors can afford to sustain with the Open Source model…

- Talos Workstations

- Libre engines such as the ones from https://osgameclones.com

- Godot

- Use anything else, there are tons of libre search engines

Re: Open-R1: an open reproduction of DeepSeek-R1

#167

Earlier quoted context omitted.

I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.

Yeah, and when I was in high school everyone used to refer to Encarta. > I know in our university pretty much everyone that attends exams uses chatgpt to study. And they shouldn't be doing that. They are wrong. Students should be reading suggested bibliography and spending long hours with an open book in a table instead of being lazy and abuse a tech that is yet in its infancy when learning concepts. Studying with a…

I don't know why you are being downvoted. Learning from something that regularly hallucinates info doesn't seem right. I think AI is a good starting point to learn about what terms to research on your own though.

Re: Open-R1: an open reproduction of DeepSeek-R1

#168
post #72

Earlier quoted context omitted.

No. My memory of the advent of the WWW was a sidebar in PC Magazine in Nov 1993 with an FTP link to download the NCSA Mosaic browser. It was a wow! moment to visit the few sites that existed. But nothing like this. What we’re seeing now is generating vastly more interest and excitement. It’s more akin to the 1999 dotcom bubble, but with far more impact and reach.

>far more impact and reach. Absurd, AI has had zero impact in the everyday life of most of the population of Earth, in fact the biggest impact has been upon the wallets of speculators

I am personally using it for around 50% of my questions about all kinds of things (things I used to Google and get frustrated with bad results). And my wife uses it for about 40% right now, even or recipes and other bits. We both love it.

Work wise about to implement it and see how it does on some work we couldn't scale to humans.

Re: Open-R1: an open reproduction of DeepSeek-R1

#170

Earlier quoted context omitted.

A friend of mine just defended his law PhD and in the introductory lectio said that (even) current LLMs would likely give better verdicts than human judges. Law isn't really a cognitively such demanding task as walking a dog or waiting tables.

> current LLMs would likely give better verdicts He probably meant _brainwashed_ LLMs. They can consistently produce desired results if you wash them the right way. It's more about personal opinion than computation. Actually it would be fun to manipulate verdicts with prompt injections ;)

Judges are very much "brainwashed" too, and by design. The judges should apply the law, and the same case should ideally lead to the same verdict regardless of the judge.

With the caveat that this applies to sane legal systems, and not the ones where "making examples" etc are part of the system.

Post reply on HN