Live data from Hacker News

Open-R1: an open reproduction of DeepSeek-R1

huggingface.co

111–120 of 246 posts

Re: Open-R1: an open reproduction of DeepSeek-R1

#111
Now that things are really getting wild in the LLM space and people are just running anything that come it seems I did a quick search on the thead model of hosting you own LLM.

I didn't find much, starting with llama.ccp which is just reminding you to sandbox and isolate everything if running untrusted models.

I feel we are back in the Windows 95 / early Internet era when people would just run anything without caring about security.

Re: Open-R1: an open reproduction of DeepSeek-R1

#113
post #72

Earlier quoted context omitted.

No. My memory of the advent of the WWW was a sidebar in PC Magazine in Nov 1993 with an FTP link to download the NCSA Mosaic browser. It was a wow! moment to visit the few sites that existed. But nothing like this. What we’re seeing now is generating vastly more interest and excitement. It’s more akin to the 1999 dotcom bubble, but with far more impact and reach.

>far more impact and reach. Absurd, AI has had zero impact in the everyday life of most of the population of Earth, in fact the biggest impact has been upon the wallets of speculators

chatgpt has >1b arr, so for comparison that’s about the size of Notion.

Re: Open-R1: an open reproduction of DeepSeek-R1

#114

Now that things are really getting wild in the LLM space and people are just running anything that come it seems I did a quick search on the thead model of hosting you own LLM. I didn't find much, starting with llama.ccp which is just reminding you to sandbox and isolate everything if running untrusted models. I feel we are back in the Windows 95 / early Internet era when people would just run anything without caring…

Anyone caring about security would be left behind in the race.

Re: Open-R1: an open reproduction of DeepSeek-R1

#115

Now that things are really getting wild in the LLM space and people are just running anything that come it seems I did a quick search on the thead model of hosting you own LLM. I didn't find much, starting with llama.ccp which is just reminding you to sandbox and isolate everything if running untrusted models. I feel we are back in the Windows 95 / early Internet era when people would just run anything without caring…

You’ll want to use something trusted like Ollama to run the model. The model itself is just data though, like a video file. That doesn’t mean it can’t be crafted to use a bug in Ollama to launch an exploit but it’s a lot safer than you make it sound.

Re: Open-R1: an open reproduction of DeepSeek-R1

#116

Now that things are really getting wild in the LLM space and people are just running anything that come it seems I did a quick search on the thead model of hosting you own LLM. I didn't find much, starting with llama.ccp which is just reminding you to sandbox and isolate everything if running untrusted models. I feel we are back in the Windows 95 / early Internet era when people would just run anything without caring…

You’ll want to use something trusted like Ollama to run the model. The model itself is just data though, like a video file. That doesn’t mean it can’t be crafted to use a bug in Ollama to launch an exploit but it’s a lot safer than you make it sound.

If used as an agent, given access to execute code, search web, use other form of tools, it could do potentially much more. And most productive usecases require access to such tools. If you want to automate things and get most of the modeel, you will have to give it ability to use tools.

E.g. it could have been trained to launch a delayed attack if context indicates it has access to execute code and given certain conditions, e.g. date, or other type of codeword that is input to it.

So if a malicious actor gets to a certain stage with an LLM where they are confident it will be able to reliably run this attack, all they have to do is open source it, wait for enough adoption and then use some of those methods to launch such attack. No one would be able to identify it since the weights are unreadable, but really somewhere in the weights this attack is just hiding and waiting to happen given correct pathway triggered.

Re: Open-R1: an open reproduction of DeepSeek-R1

#118
post #72

Earlier quoted context omitted.

No. My memory of the advent of the WWW was a sidebar in PC Magazine in Nov 1993 with an FTP link to download the NCSA Mosaic browser. It was a wow! moment to visit the few sites that existed. But nothing like this. What we’re seeing now is generating vastly more interest and excitement. It’s more akin to the 1999 dotcom bubble, but with far more impact and reach.

>far more impact and reach. Absurd, AI has had zero impact in the everyday life of most of the population of Earth, in fact the biggest impact has been upon the wallets of speculators

I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.

Re: Open-R1: an open reproduction of DeepSeek-R1

#119
post #58

Earlier quoted context omitted.

Rounded corners, anyone? Who remembers LISP being new?

I've got an original Apple ][ reference manual (red cover) with the hand annotated ROM listing. Also have SMALLTALK-80 book with it's railtrack diagrams of syntax on the inside covers. What's really interesting about the AI mania we're in is that no one has shown that what we have now will get to AGI and how. We have great models that simulate reasoning, but how close are they? How do we measure their quality? Benchm…

A different point of view on AGI is that we humans do not achieve AGI. Our brains aren’t capable of it. We get close enough to trick the other humans we compete against for resources. How would we prove that’s not true? Something like IQ tests? We don’t have good tests or benchmarks or tooling for this in ourselves, let alone the reproduction in machines. No one knows definitively what AGI actually is so, depending on where you set that bar, we might already be there.

Re: Open-R1: an open reproduction of DeepSeek-R1

#120
post #71

Earlier quoted context omitted.

The US companies got too greedy? How? They invented this entire space, literally. DeepSeek built their base models off Llama releases and OpenAI outputs (or so it’s thought), and while they added some optimizations on top, it seems like they’ve lied about the costs to produce their models by simply being vague about their base model and training data, and quoting the cost of their final training run. And then there’s…

> DeepSeek built their base models off Llama releases and OpenAI outputs Those models are also trained on data that was ignoring licenses / copyrighted content.

That's a problem for sure. But why would that argument play in favour china which would be even less constrained by licenses / copyright ?
Post reply on HN