Live data from Hacker News

Open-R1: an open reproduction of DeepSeek-R1

huggingface.co

201–210 of 246 posts

Re: Open-R1: an open reproduction of DeepSeek-R1

#201
post #199

Earlier quoted context omitted.

I ran it 10 times with the extra information, and each time got a different result. I don't know if any of them were the specific joke you were after, I get the feeling it was just making them up on the spot. None of them are even funny

Here is an example of Sonnet finding the right joke after two messages: https://i.imgur.com/nKvS2cW.png It seems to be censored with US puritan morality (like most US models), but I think that's besides the point (just like if the joke is "even funny" or not), as it did find the correct joke at least.

I just got a load of responses like "Sure, here’s a joke that combines cars, Spain, politicians, and a fascist with a touch of space humor: Why did the Spanish politician, the fascist, and the car mechanic get together to start a space program? Because the politician wanted to go "far-right," the mechanic said he could "fix" anything, and the fascist just wanted to take the car to the moon... so they could all escape when things got "too hot" here on Earth!"

Re: Open-R1: an open reproduction of DeepSeek-R1

#202
post #77

Earlier quoted context omitted.

the web was fascinating every second. You could click on a link without having ANY idea what you would land on. The overall quality was very poor, but it was thrilling. A bit like indie cinema.

My first search: cocktails

The Webtender is still going strong!

https://www.webtender.com/

Re: Open-R1: an open reproduction of DeepSeek-R1

#203
post #199

Earlier quoted context omitted.

Here is an example of Sonnet finding the right joke after two messages: https://i.imgur.com/nKvS2cW.png It seems to be censored with US puritan morality (like most US models), but I think that's besides the point (just like if the joke is "even funny" or not), as it did find the correct joke at least.

I just got a load of responses like "Sure, here’s a joke that combines cars, Spain, politicians, and a fascist with a touch of space humor: Why did the Spanish politician, the fascist, and the car mechanic get together to start a space program? Because the politician wanted to go "far-right," the mechanic said he could "fix" anything, and the fascist just wanted to take the car to the moon... so they could all escape…

Ok, that's cool. So because you were unable to find a needle in this case, your conclusion is that it's impossible that other people to use LLMs for this, and LLMs truly are just glorified Wikipedia/Google?

Re: Open-R1: an open reproduction of DeepSeek-R1

#204
post #203

Earlier quoted context omitted.

I just got a load of responses like "Sure, here’s a joke that combines cars, Spain, politicians, and a fascist with a touch of space humor: Why did the Spanish politician, the fascist, and the car mechanic get together to start a space program? Because the politician wanted to go "far-right," the mechanic said he could "fix" anything, and the fascist just wanted to take the car to the moon... so they could all escape…

Ok, that's cool. So because you were unable to find a needle in this case, your conclusion is that it's impossible that other people to use LLMs for this, and LLMs truly are just glorified Wikipedia/Google?

No, I don't think that LLMs are glorified Wikipedia/Google. I think they're a glorified version of pressing the middle button on your phone's autocomplete repeatedly

Re: Open-R1: an open reproduction of DeepSeek-R1

#205

Earlier quoted context omitted.

It is though. Western AI tries to hide information like that with the justification of safety as well as things that might be offensive to current popular beliefs. Chinese AI presumably says Taiwan is China to help get more people on side for a possible future invasion. Propaganda does work - look at how many people think Donbas is still Ukraine and Israel is still Palestine.

The difference is that in China the info isn’t available without use of Western content, due to the totalitarian control over media, whereas in the West, information is pretty trivially available, even if the big companies keep it off of their platforms. And sure ignorance is prevalent, but even GPT4 will tell me Donbas is still Ukraine, for instance. What a strange example to use, though!

If western governments are so tolerant and permissive with information, I wonder why can't I access RT in Europe?

Re: Open-R1: an open reproduction of DeepSeek-R1

#206

Earlier quoted context omitted.

A delayed attack is a bit of a stretch because that’s a stateful thing but a reminder of the Ken Thompson hack does feel very relevant.

But if it's specifically trained to react to a date in its context, it seems very doable. Or to a combination of otherwise seemingly innocent words or even a statement or topic. E.g. a malicious actor could make some certain notion go viral and agentic LLMs integrated with news headlines might react to that. It seems like it would be very arbitrary to train it to behave like this. Most agentic systems would provide a…

> And China has a lot of other tech available to it already that they could do much more harm (phones, robot vacuums)

True. But here it is more about the computing power they would be able to access.

If only Bitcoin or Ethereum were stilled mined using GPU, that would be a great cryptojacking opportunity

Re: Open-R1: an open reproduction of DeepSeek-R1

#207
post #172

Earlier quoted context omitted.

I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.

> Chatgpt is arguably more valuable then Wikipedia and Google for studies. But ChatGPT is just a glorified Wikipedia/Google. For the consumers it's an incremental thing (although from the engineering perspective it may seem to be a breakthrough).

Go try to learn a college level mathematics concept from Wikipedia, then try to learn it from ChatGPT. The wiki article may as well be written in a foreign language

Re: Open-R1: an open reproduction of DeepSeek-R1

#208
post #110

Earlier quoted context omitted.

Terminology exists for a reason. Doubly so for well-established terms of art that pertain to licensing and contract law. They could have used "open wights" which would have conveyed the company's desired intent just as well as "open source", but without the ambiguity. They deliberately chose to misuse a well established term instead. I applaud and thank deepseek for opening their weights, but i absolutely condemn the…

> i absolutely condemn them and others (e.g Facebook) for their deliberate and continued misuse of the term This is the kind of inconsequential nitpicking diatribe I'm referring to. When has "open data" ever meant Open Source? > They deliberately chose to misuse a well established term instead. Their model weights as well as their repositories containing their technical papers and any source code are published under…

We're talking past each other at this point. I believe both our positions have been adequately presented. Cheers.

Re: Open-R1: an open reproduction of DeepSeek-R1

#209
post #203

Earlier quoted context omitted.

Ok, that's cool. So because you were unable to find a needle in this case, your conclusion is that it's impossible that other people to use LLMs for this, and LLMs truly are just glorified Wikipedia/Google?

No, I don't think that LLMs are glorified Wikipedia/Google. I think they're a glorified version of pressing the middle button on your phone's autocomplete repeatedly

So you didn't enter the conversation to follow along with the existing discussion, but to share your grievance about how LLMs work regardless? Useful

Re: Open-R1: an open reproduction of DeepSeek-R1

#210
post #161

Earlier quoted context omitted.

It feels that the open source movement is slowly entering a Cambrian explosion stage. You have the old "deterministic computing" achievements (with Linux the flagship). Then you have the networking protocols (activitypub / atproto) that are revolutionising birectional human interactions online. And finally you have the datascience/ML/AI algorithmic universe that is for the first time being harnessed at distributed sc…

Revolutionising what? libre/oss has been network-bound since Usenet and then IRC.

how many people ever used Usenet versus the billions who think the "internet" is facebook or tiktok. Techies living in their own universe detached from human reality is actually a factor why libre/oss is not as widely adopted as it could be.
Post reply on HN