For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
881–890 of 1001 posts
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#882[deleted]
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#883Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#884For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…
It’s not better than o1. And given that OpenAI is on the verge of releasing o3, has some “o4” in the pipeline, and Deepseek could only build this because of o1, I don’t think there’s as much competition as people seem to imply. I’m excited to see models become open, but given the curve of progress we’ve seen, even being “a little” behind is a gap that grows exponentially every day.
This theory has yet to be demonstrated. As yet, it seems open source just stays behind by about 6-10 months consistently.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#885For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…
This suggests r1 is indeed better at reasoning but its coding is holding it back, which checks out given the large corpus of coding tasks and much less rich corpus for reasoning.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#886Earlier quoted context omitted.
As we have seen here it won't be a Western company that saves us from the dominant monopoly. Xi Jinping, you're our only hope.
If China really released a GPU competitive with the current generation of nvidia you can bet it'd be banned in the US like BYD and DJI.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#887Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#888Aside from the usual Tiananmen Square censorship, there's also some other propaganda baked-in: https://prnt.sc/HaSc4XZ89skA (from reddit)
Try asking ChatGPT about the genocide Israel is committing. Then you'll see what censorship looks like.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#889Earlier quoted context omitted.
yes, this is all ollamas fault
ollama is stating there's a difference: https://ollama.com/library/deepseek-r1 "including six dense models distilled from DeepSeek-R1 based on Llama and Qwen. " people just don't read? not sure there's reason to criticize ollama here.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#890Earlier quoted context omitted.
I haven't tried kagi assistant, but try it at deepseek.com. All models at this point have various politically motivated filters. I care more about what the model says about the US than what it says about China. Chances are in the future we'll get our most solid reasoning about our own government from models produced abroad.
False equivalency. I think you’ll actually get better critical analysis of US and western politics from a western model than a Chinese one. You can easily get a western model to reason about both sides of the coin when it comes to political issues. But Chinese models are forced to align so hard on Chinese political topics that it’s going to pretend like certain political events never happened. E.g try getting them to…