Live data from Hacker News

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

arxiv.org

881–890 of 1001 posts

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#881

For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…

I think you mean American EV competition. China has a very large and primarily-unknown-to-the-average-American large EV industry. It's not just Tesla.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#884

For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…

It’s not better than o1. And given that OpenAI is on the verge of releasing o3, has some “o4” in the pipeline, and Deepseek could only build this because of o1, I don’t think there’s as much competition as people seem to imply. I’m excited to see models become open, but given the curve of progress we’ve seen, even being “a little” behind is a gap that grows exponentially every day.

> even being “a little” behind is a gap that grows exponentially every day

This theory has yet to be demonstrated. As yet, it seems open source just stays behind by about 6-10 months consistently.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#885

For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…

The aider benchmarks that swyx posted below suggest o1 is still better than r1 (though an oom more expensive). Interestingly r1+sonnet (architect/editor) wins though.

This suggests r1 is indeed better at reasoning but its coding is holding it back, which checks out given the large corpus of coding tasks and much less rich corpus for reasoning.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#886

Earlier quoted context omitted.

As we have seen here it won't be a Western company that saves us from the dominant monopoly. Xi Jinping, you're our only hope.

If China really released a GPU competitive with the current generation of nvidia you can bet it'd be banned in the US like BYD and DJI.

DJI isn't banned in the US?

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#887

Earlier quoted context omitted.

yes, they are not r1

Can you explain what you mean by this?

For example, the model named "deepseek-r1:8b" by ollama is not a deepseek r1 model. It is actually a fine tune of Meta's Llama 8b, fine tuned on data generated by deepseek r1.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#888
post #558
post #151

Aside from the usual Tiananmen Square censorship, there's also some other propaganda baked-in: https://prnt.sc/HaSc4XZ89skA (from reddit)

Try asking ChatGPT about the genocide Israel is committing. Then you'll see what censorship looks like.

Well, I just tried this, and I didn't see any censorship?

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#889

Earlier quoted context omitted.

yes, this is all ollamas fault

ollama is stating there's a difference: https://ollama.com/library/deepseek-r1 "including six dense models distilled from DeepSeek-R1 based on Llama and Qwen. " people just don't read? not sure there's reason to criticize ollama here.

i’ve seen so many people make this misunderstanding, huggingface clearly differentiates the model, and from the cli that isn’t visible

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#890
post #605

Earlier quoted context omitted.

I haven't tried kagi assistant, but try it at deepseek.com. All models at this point have various politically motivated filters. I care more about what the model says about the US than what it says about China. Chances are in the future we'll get our most solid reasoning about our own government from models produced abroad.

False equivalency. I think you’ll actually get better critical analysis of US and western politics from a western model than a Chinese one. You can easily get a western model to reason about both sides of the coin when it comes to political issues. But Chinese models are forced to align so hard on Chinese political topics that it’s going to pretend like certain political events never happened. E.g try getting them to…

[deleted]
Post reply on HN