Earlier quoted context omitted.
I ran it 10 times with the extra information, and each time got a different result. I don't know if any of them were the specific joke you were after, I get the feeling it was just making them up on the spot. None of them are even funny
Here is an example of Sonnet finding the right joke after two messages: https://i.imgur.com/nKvS2cW.png It seems to be censored with US puritan morality (like most US models), but I think that's besides the point (just like if the joke is "even funny" or not), as it did find the correct joke at least.
Open-R1: an open reproduction of DeepSeek-R1
201–210 of 246 posts
Re: Open-R1: an open reproduction of DeepSeek-R1
#202Earlier quoted context omitted.
the web was fascinating every second. You could click on a link without having ANY idea what you would land on. The overall quality was very poor, but it was thrilling. A bit like indie cinema.
My first search: cocktails
Re: Open-R1: an open reproduction of DeepSeek-R1
#203Earlier quoted context omitted.
Here is an example of Sonnet finding the right joke after two messages: https://i.imgur.com/nKvS2cW.png It seems to be censored with US puritan morality (like most US models), but I think that's besides the point (just like if the joke is "even funny" or not), as it did find the correct joke at least.
I just got a load of responses like "Sure, here’s a joke that combines cars, Spain, politicians, and a fascist with a touch of space humor: Why did the Spanish politician, the fascist, and the car mechanic get together to start a space program? Because the politician wanted to go "far-right," the mechanic said he could "fix" anything, and the fascist just wanted to take the car to the moon... so they could all escape…
Re: Open-R1: an open reproduction of DeepSeek-R1
#204Earlier quoted context omitted.
I just got a load of responses like "Sure, here’s a joke that combines cars, Spain, politicians, and a fascist with a touch of space humor: Why did the Spanish politician, the fascist, and the car mechanic get together to start a space program? Because the politician wanted to go "far-right," the mechanic said he could "fix" anything, and the fascist just wanted to take the car to the moon... so they could all escape…
Ok, that's cool. So because you were unable to find a needle in this case, your conclusion is that it's impossible that other people to use LLMs for this, and LLMs truly are just glorified Wikipedia/Google?
Re: Open-R1: an open reproduction of DeepSeek-R1
#205Earlier quoted context omitted.
It is though. Western AI tries to hide information like that with the justification of safety as well as things that might be offensive to current popular beliefs. Chinese AI presumably says Taiwan is China to help get more people on side for a possible future invasion. Propaganda does work - look at how many people think Donbas is still Ukraine and Israel is still Palestine.
The difference is that in China the info isn’t available without use of Western content, due to the totalitarian control over media, whereas in the West, information is pretty trivially available, even if the big companies keep it off of their platforms. And sure ignorance is prevalent, but even GPT4 will tell me Donbas is still Ukraine, for instance. What a strange example to use, though!
Re: Open-R1: an open reproduction of DeepSeek-R1
#206Earlier quoted context omitted.
A delayed attack is a bit of a stretch because that’s a stateful thing but a reminder of the Ken Thompson hack does feel very relevant.
But if it's specifically trained to react to a date in its context, it seems very doable. Or to a combination of otherwise seemingly innocent words or even a statement or topic. E.g. a malicious actor could make some certain notion go viral and agentic LLMs integrated with news headlines might react to that. It seems like it would be very arbitrary to train it to behave like this. Most agentic systems would provide a…
True. But here it is more about the computing power they would be able to access.
If only Bitcoin or Ethereum were stilled mined using GPU, that would be a great cryptojacking opportunity
Re: Open-R1: an open reproduction of DeepSeek-R1
#207Earlier quoted context omitted.
I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.
> Chatgpt is arguably more valuable then Wikipedia and Google for studies. But ChatGPT is just a glorified Wikipedia/Google. For the consumers it's an incremental thing (although from the engineering perspective it may seem to be a breakthrough).
Re: Open-R1: an open reproduction of DeepSeek-R1
#208Earlier quoted context omitted.
Terminology exists for a reason. Doubly so for well-established terms of art that pertain to licensing and contract law. They could have used "open wights" which would have conveyed the company's desired intent just as well as "open source", but without the ambiguity. They deliberately chose to misuse a well established term instead. I applaud and thank deepseek for opening their weights, but i absolutely condemn the…
> i absolutely condemn them and others (e.g Facebook) for their deliberate and continued misuse of the term This is the kind of inconsequential nitpicking diatribe I'm referring to. When has "open data" ever meant Open Source? > They deliberately chose to misuse a well established term instead. Their model weights as well as their repositories containing their technical papers and any source code are published under…
Re: Open-R1: an open reproduction of DeepSeek-R1
#209Earlier quoted context omitted.
Ok, that's cool. So because you were unable to find a needle in this case, your conclusion is that it's impossible that other people to use LLMs for this, and LLMs truly are just glorified Wikipedia/Google?
No, I don't think that LLMs are glorified Wikipedia/Google. I think they're a glorified version of pressing the middle button on your phone's autocomplete repeatedly
Re: Open-R1: an open reproduction of DeepSeek-R1
#210Earlier quoted context omitted.
It feels that the open source movement is slowly entering a Cambrian explosion stage. You have the old "deterministic computing" achievements (with Linux the flagship). Then you have the networking protocols (activitypub / atproto) that are revolutionising birectional human interactions online. And finally you have the datascience/ML/AI algorithmic universe that is for the first time being harnessed at distributed sc…
Revolutionising what? libre/oss has been network-bound since Usenet and then IRC.