Live data from Hacker News

Open-R1: an open reproduction of DeepSeek-R1

huggingface.co

151–160 of 246 posts

Re: Open-R1: an open reproduction of DeepSeek-R1

#151
post #6

What are some other domains outside of Math and Coding that would be suitable for RL with automated verification?

These LLMs are already very helpful when studying scientific fields. If you're reading a scientific paper and come across an equation you don't know how to derive, LLMs can often correctly derive it from first principles. It's not 100% reliable, but when it works, it's incredibly helpful.

I can't really think of something that involves learning that these tools wouldn't be helpful with.

Some people talk though as if the books on my bookshelf spontaneously combust if a language model helps me with anything.

Not to mention that I had many professors in college that were so full of shit they put the hallucinations of chatgpt3.5 to shame.

Re: Open-R1: an open reproduction of DeepSeek-R1

#152

Now that things are really getting wild in the LLM space and people are just running anything that come it seems I did a quick search on the thead model of hosting you own LLM. I didn't find much, starting with llama.ccp which is just reminding you to sandbox and isolate everything if running untrusted models. I feel we are back in the Windows 95 / early Internet era when people would just run anything without caring…

[deleted]

Re: Open-R1: an open reproduction of DeepSeek-R1

#153

Earlier quoted context omitted.

Yeah, and when I was in high school everyone used to refer to Encarta. > I know in our university pretty much everyone that attends exams uses chatgpt to study. And they shouldn't be doing that. They are wrong. Students should be reading suggested bibliography and spending long hours with an open book in a table instead of being lazy and abuse a tech that is yet in its infancy when learning concepts. Studying with a…

You sound like our teachers back in the day, warning us to not use Wikipedia because "everyone can write stuff there!!". The kids will be fine.

Wikipedia furnishes sources, which is what you actually use.

Re: Open-R1: an open reproduction of DeepSeek-R1

#154

Now that things are really getting wild in the LLM space and people are just running anything that come it seems I did a quick search on the thead model of hosting you own LLM. I didn't find much, starting with llama.ccp which is just reminding you to sandbox and isolate everything if running untrusted models. I feel we are back in the Windows 95 / early Internet era when people would just run anything without caring…

Anyone caring about security would be left behind in the race.

[deleted]

Re: Open-R1: an open reproduction of DeepSeek-R1

#155
post #10

Is this what the Web was like in the beginning? Something exciting and fascinating every week?

The early years of the web were absolutely this chaotic maelstrom of new things happening every week. But news of it was hard to come by. In the UK / Ireland we had some great tech coverage in the form of shows like 'The Net' [1] that regularly showed off early internet craziness like the 'We Live in Public' project.

However a better analogy would be the 'web 2.0' era, when as a college student I had an early internet politics / technology podcast [3]. It seemed like every week there was a huge new development either in technology or surveillance. From the first location based social networks [4] to the birth of Youtube. People were podcasting for the first time, and internet video was becoming economically feasible at low to no cost. It was really a radical time, with broadcasters freaking out about how they would adapt, and a whole generation of people becoming whats now known as 'content creators'.

[1] https://en.wikipedia.org/wiki/The_Net_(British_TV_series)

[2] https://en.wikipedia.org/wiki/We_Live_in_Public

[3] https://archive.org/search?query=technolotics

[4]https://en.wikipedia.org/wiki/Jaiku

Re: Open-R1: an open reproduction of DeepSeek-R1

#156

Now that things are really getting wild in the LLM space and people are just running anything that come it seems I did a quick search on the thead model of hosting you own LLM. I didn't find much, starting with llama.ccp which is just reminding you to sandbox and isolate everything if running untrusted models. I feel we are back in the Windows 95 / early Internet era when people would just run anything without caring…

Anyone caring about security would be left behind in the race.

somebody should pentest all your stuff

Re: Open-R1: an open reproduction of DeepSeek-R1

#157
post #61

Earlier quoted context omitted.

Deep mind has one set of censorship, OpenAI another, anything musk does a third It’s all “massaged”

I don’t think “Taiwan is China” is the same kind of massaging as not telling people how to make napalm… What a weird thing to equate, though!

But Taiwan is China. Even the USA acknowledge that.

Re: Open-R1: an open reproduction of DeepSeek-R1

#158
post #18
post #9

Earlier quoted context omitted.

Oh yes, I am firmly on Team China here because US companies got too greedy. Meta is an exception here though and they also propelled AI development massively. DeepSeek is awesome. Any AI task yet implemented in our business can be run from my local PC with just the smaller models. And my PC is fairly crappy to begin with. OpenAI looks quite silly with their "we have to close everything".

Can you elaborate which models you are using? I‘m running an R1 distilled Qwen coder with 32B Q4, and while it’s giving useful answers, it‘s quite slow on my M1 Max. Slow enough that I keep reaching for cloud models.

Not on my machine currently, I use the 14b Q4 model I think, which delivers very good answers. I run a 4060 with 16gb memory and performance is quite good. I used the largest model that was recommended with this amount of VRAM, I think it was the 14b one.

I do have some applications that process images, text and pdf files and I use smaller models for extracting embeddings. I think my system wouldn't be able to handle it with decent speed otherwise.

I do run LLM on a M1 16gb macbook air and performance is surprisingly good. Not for image synthesis though and a PC with a dedicated GPU is still significantly faster with LLM responses as well. Haven't tried to run deepseek on the macbook yet.

Re: Open-R1: an open reproduction of DeepSeek-R1

#159
post #9

Earlier quoted context omitted.

Oh yes, I am firmly on Team China here because US companies got too greedy. Meta is an exception here though and they also propelled AI development massively. DeepSeek is awesome. Any AI task yet implemented in our business can be run from my local PC with just the smaller models. And my PC is fairly crappy to begin with. OpenAI looks quite silly with their "we have to close everything".

The US companies got too greedy? How? They invented this entire space, literally. DeepSeek built their base models off Llama releases and OpenAI outputs (or so it’s thought), and while they added some optimizations on top, it seems like they’ve lied about the costs to produce their models by simply being vague about their base model and training data, and quoting the cost of their final training run. And then there’s…

That was perhaps a bit too general, but aside from meta and Google they didn't share their research and tried to sell AI products as fast as possible and tried to lobby legislation to keep their head start. I would also include nvidia here, that has some moat through software integrations.

I haven't tested deepseek for censorship yet, but they shared their release and even their input data. And in this case you could correct its shortcomings, so propaganda would be difficult.

Re: Open-R1: an open reproduction of DeepSeek-R1

#160
post #138

How can we help. Can crowd sourcing help? Is there any list of tasks that we want a crowd to do? The reason I am asking is because we have done a couple of crowdsourcing efforts and collected story data in Telugu(Chandamama Kathalu) and ASR speech data using college going students. Since we have access to the students, we can mobilize them and get this going. We will also be doing an internship program for 100,000 st…

We don’t need your help anymore. We don’t need anyone’s help. We have Deepseek now to do it. God save us all
Post reply on HN