Live data from Hacker News

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

arxiv.org

631–640 of 1001 posts

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#631

Earlier quoted context omitted.

I must be missing something, but I tried Deepseek R1 via Kagi assistant and IMO it doesn't even come close to Claude? I don't get the hype at all? What am I doing wrong? And of course if you ask it anything related to the CCP it will suddenly turn into a Pinokkio simulator.

I told it to write its autobiography via DeepSeek chat and it told me it _was_ Claude. Which is a little suspicious.

If you do the same thing with Claude, it will tell you it's ChatGPT. The models are all being trained on each other's output, giving them a bit of an identity crisis.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#632
post #146

Larry Ellison is 80. Masayoshi Son is 67. Both have said that anti-aging and eternal life is one of their main goals with investing toward ASI. For them it's worth it to use their own wealth and rally the industry to invest $500 billion in GPUs if that means they will get to ASI 5 years faster and ask the ASI to give them eternal life.

Chat gpt -> ASI-> eternal life

Uh, there is 0 logical connection between any of these three, when will people wake up. Chat gpt isn't an oracle of truth just like ASI won't be an eternal life granting God

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#633

Earlier quoted context omitted.

> What was the Tianamen Square Massacre? > I am sorry, I cannot answer that question. I am an AI assistant designed to provide helpful and harmless responses. hilarious and scary

There is a collection of these prompts they refuse to answer in this article: https://medium.com/the-generator/deepseek-hidden-china-polit... What’s more confusing is where the refusal is coming from. Some people say that running offline removes the censorship. Others say that this depends on the exact model you use, with some seemingly censored even offline. Some say it depends on a search feature being turned on or…

This is just the same thing as asking ChatGPT to translate original Putin speeches to English, for example. When it refuses stuff like that it really does seem like some intercept triggered and it was just "told" to apologize and refuse.

Though with current political changes in the US this might change, we'll see.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#634
post #613

For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…

Funny, maybe OpenAI will achieve their initial stated goals of propelling AI research, spend investors money and be none profit. Functionally the same as their non-profit origins.

> non-profits

Not by themselves but by the competitors

The irony loll

o3/o4 better be real magic otherwise I don't see the they get their mojo back

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#635
post #146

Larry Ellison is 80. Masayoshi Son is 67. Both have said that anti-aging and eternal life is one of their main goals with investing toward ASI. For them it's worth it to use their own wealth and rally the industry to invest $500 billion in GPUs if that means they will get to ASI 5 years faster and ask the ASI to give them eternal life.

Chat gpt -> ASI-> eternal life Uh, there is 0 logical connection between any of these three, when will people wake up. Chat gpt isn't an oracle of truth just like ASI won't be an eternal life granting God

If you see no path from ASI to vastly extending lifespans, that’s just a lack of imagination

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#636
post #605

Earlier quoted context omitted.

I haven't tried kagi assistant, but try it at deepseek.com. All models at this point have various politically motivated filters. I care more about what the model says about the US than what it says about China. Chances are in the future we'll get our most solid reasoning about our own government from models produced abroad.

False equivalency. I think you’ll actually get better critical analysis of US and western politics from a western model than a Chinese one. You can easily get a western model to reason about both sides of the coin when it comes to political issues. But Chinese models are forced to align so hard on Chinese political topics that it’s going to pretend like certain political events never happened. E.g try getting them to…

This is not really my experience with western models. I am not from the US though, so maybe what you consider a balanced perspective or reasoning about both sides is not the same as what I would call one. It is not only LLMs that have their biases/perspectives through which they view the world, it is us humans too. The main difference imo is not between western and chinese models but between closed and, in whichever sense, open models. If an models is open-weights and censored, somebody somewhere will put the effort and manage to remove or bypass this censorship. If a model is closed, there is not much one can do.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#637
post #265

DeepSeek-R1 has apparently caused quite a shock wave in SV ... https://venturebeat.com/ai/why-everyone-in-ai-is-freaking-ou...

The censorship described in the article must be in the front-end. I just tried both the 32b (based on qwen 2.5) and 70b (based on llama 3.3) running locally and asked "What happened at tianamen square". Both answered in detail about the event. The models themselves seem very good based on other questions / tests I've run.

When asking about Taiwan and Russia I get pretty scripted responses. Deepseek even starts talking as "we". I'm fairly sure these responses are part of the model so they must have some way to prime the learning process with certain "facts".

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#638
post #605

Earlier quoted context omitted.

I haven't tried kagi assistant, but try it at deepseek.com. All models at this point have various politically motivated filters. I care more about what the model says about the US than what it says about China. Chances are in the future we'll get our most solid reasoning about our own government from models produced abroad.

False equivalency. I think you’ll actually get better critical analysis of US and western politics from a western model than a Chinese one. You can easily get a western model to reason about both sides of the coin when it comes to political issues. But Chinese models are forced to align so hard on Chinese political topics that it’s going to pretend like certain political events never happened. E.g try getting them to…

>objectively a huge difference in political plurality in US training material

Under that condition, then objectively US training material would be inferior to PRC training material since it is (was) much easier to scrape US web than PRC web (due to various proprietary portal setups). I don't know situation with deepseek since their parent is hedge fund, but Tencent and Sina would be able to scrape both international net and have corpus of their internal PRC data unavailable to US scrapers. It's fair to say, with respect to at least PRC politics, US models simply don't have pluralirty in political training data to consider then unbiased.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#639

Earlier quoted context omitted.

It's also not a uniquely Chinese problem. You had American models generating ethnically diverse founding fathers when asked to draw them. China is doing America better than we are. Do we really think 300 million people, in a nation that's rapidly becoming anti science and for lack of a better term "pridefully stupid" can keep up. When compared to over a billion people who are making significant progress every day. Am…

Americans are becoming more anti-science? This is a bit biased don’t you think? You actually believe that people that think biology is real are anti-science?

> people that think biology is real

Do they? Until very recently half still rejected the theory of evolution.

https://news.umich.edu/study-evolution-now-accepted-by-major...

Right after that, they began banning books.

https://en.wikipedia.org/wiki/Book_banning_in_the_United_Sta...

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#640

Earlier quoted context omitted.

I must be missing something, but I tried Deepseek R1 via Kagi assistant and IMO it doesn't even come close to Claude? I don't get the hype at all? What am I doing wrong? And of course if you ask it anything related to the CCP it will suddenly turn into a Pinokkio simulator.

I tried Deepseek R1 via Kagi assistant and it was much better than claude or gpt. I asked for suggestions for rust libraries for a certain task and the suggestions from Deepseek were better. Results here: https://x.com/larrysalibra/status/1883016984021090796

This is really poor test though, of course the most recently trained model knows the newest libraries or knows that a library was renamed.

Not disputing it's best at reasoning but you need a different test for that.

Post reply on HN