I've been using https://chat.deepseek.com/ over My ChatGPT Pro subscription because being able to read the thinking in the way they present it is just much much easier to "debug" - also I can see when it's bending it's reply to something, often softening it or pandering to me - I can just say "I saw in your thinking you should give this type of reply, don't do that". If it stays free and gets better that's going to b…
I tried signing up, but it gave me some bullshit "this email domain isn't supported in your region." I guess they insist on a GMail account or something? Regardless I don't even trust US-based LLM products to protect my privacy, let alone China-based. Remember kids: If it's free, you're the product. I'll give it a while longer before I can run something competitive on my own hardware. I don't mind giving it a few yea…
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
511–520 of 1001 posts
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#512For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…
I don't get the hype at all?
What am I doing wrong?
And of course if you ask it anything related to the CCP it will suddenly turn into a Pinokkio simulator.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#513Earlier quoted context omitted.
It should be. I think AMD has left a lot on the table with respect to competing in the space (probably to the point of executive negligence) and the new US laws will help create several new Chinese competitors. NVIDIA probably has a bit of time left as the market leader, but it's really due mostly to luck.
As we have seen here it won't be a Western company that saves us from the dominant monopoly. Xi Jinping, you're our only hope.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#514For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…
There has never been much secret sauce in the model itself. The secret sauce or competitive advantage has always been in the engineering that goes into the data collection, model training infrastructure, and lifecycle/debugging management of model training. As well as in the access to GPUs. Yeah, with Deepseek the barrier to entry has become significantly lower now. That's good, and hopefully more competition will co…
In my opinion there is something qualitatively better about Deepseek in spite of its small size, even compared to o1-pro, that suggests a door has been opened.
GPUs are needed to rapidly iterate on ideas, train, evaluate, etc., but Deepseek has shown us that we are not yet in the phase where hardware CapEx guarantees victory. Imagine if Deeepseek hadn't been open sourced!
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#515Earlier quoted context omitted.
From my casual read, right now everyone is on reputation tarnishing tirade, like spamming “Chinese stealing data! Definitely lying about everything! API can’t be this cheap!”. If that doesn’t go through well, I’m assuming lobbyism will start for import controls, which is very stupid. I have no idea how they can recover from it, if DeepSeek’s product is what they’re advertising.
Funny, everything I see (not actively looking for DeepSeek related content) is absolutely raving about it and talking about it destroying OpenAI (random YouTube thumbnails, most comments in this thread, even CNBC headlines). If DeepSeek's claims are accurate, then they themselves will be obsolete within a year, because the cost to develop models like this has dropped dramatically. There are going to be a lot of teams…
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#516Earlier quoted context omitted.
There has never been much secret sauce in the model itself. The secret sauce or competitive advantage has always been in the engineering that goes into the data collection, model training infrastructure, and lifecycle/debugging management of model training. As well as in the access to GPUs. Yeah, with Deepseek the barrier to entry has become significantly lower now. That's good, and hopefully more competition will co…
The word you're looking for is copyright enfrignment. That's the secret sause that every good model uses.
It is at this point hard to imagine a model that is good at reasoning that doesn't also have vast implicit "knowledge".
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#517I've been using https://chat.deepseek.com/ over My ChatGPT Pro subscription because being able to read the thinking in the way they present it is just much much easier to "debug" - also I can see when it's bending it's reply to something, often softening it or pandering to me - I can just say "I saw in your thinking you should give this type of reply, don't do that". If it stays free and gets better that's going to b…
If you ask it about the Tienanmen Square Massacre its "thought process" is very interesting.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#518For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…
I must be missing something, but I tried Deepseek R1 via Kagi assistant and IMO it doesn't even come close to Claude? I don't get the hype at all? What am I doing wrong? And of course if you ask it anything related to the CCP it will suddenly turn into a Pinokkio simulator.
All models at this point have various politically motivated filters. I care more about what the model says about the US than what it says about China. Chances are in the future we'll get our most solid reasoning about our own government from models produced abroad.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#519Earlier quoted context omitted.
There has never been much secret sauce in the model itself. The secret sauce or competitive advantage has always been in the engineering that goes into the data collection, model training infrastructure, and lifecycle/debugging management of model training. As well as in the access to GPUs. Yeah, with Deepseek the barrier to entry has become significantly lower now. That's good, and hopefully more competition will co…
I don't disagree, but the important point is that Deepseek showed that it's not just about CapEx, which is what the US firms were/are lining up to battle with. In my opinion there is something qualitatively better about Deepseek in spite of its small size, even compared to o1-pro, that suggests a door has been opened. GPUs are needed to rapidly iterate on ideas, train, evaluate, etc., but Deepseek has shown us that w…
With R1 as inspiration/imperative, many new US startups will emerge who will be very strong. Can you feel a bunch of talent in limbo startups pivoting/re-energized now?
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#520Earlier quoted context omitted.
There has never been much secret sauce in the model itself. The secret sauce or competitive advantage has always been in the engineering that goes into the data collection, model training infrastructure, and lifecycle/debugging management of model training. As well as in the access to GPUs. Yeah, with Deepseek the barrier to entry has become significantly lower now. That's good, and hopefully more competition will co…
The word you're looking for is copyright enfrignment. That's the secret sause that every good model uses.
I personally hope that countries recognize copyright and patents for what they really are and abolish them. Countries that refuse to do so can play catch up.