Live data from Hacker News

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

arxiv.org

961–970 of 1001 posts

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#961
post #146

Larry Ellison is 80. Masayoshi Son is 67. Both have said that anti-aging and eternal life is one of their main goals with investing toward ASI. For them it's worth it to use their own wealth and rally the industry to invest $500 billion in GPUs if that means they will get to ASI 5 years faster and ask the ASI to give them eternal life.

Nice try, Larry, the reaper is coming and the world is ready to forget another shitty narcissistic CEO.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#962
post #146

Larry Ellison is 80. Masayoshi Son is 67. Both have said that anti-aging and eternal life is one of their main goals with investing toward ASI. For them it's worth it to use their own wealth and rally the industry to invest $500 billion in GPUs if that means they will get to ASI 5 years faster and ask the ASI to give them eternal life.

Chat gpt -> ASI-> eternal life Uh, there is 0 logical connection between any of these three, when will people wake up. Chat gpt isn't an oracle of truth just like ASI won't be an eternal life granting God

The world isn't run by smart people, it's run by lucky narcissistic douchebags with ketamine streaming through their veins 24/7

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#963

Earlier quoted context omitted.

I trust China a lot more than Meta and my own early tests do indeed show that Deepseek is far less censored than Llama.

Interesting. What topics are censored on Llama?

I can't help but wonder if this is just a dogwhistle for pornography?

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#964
post #835

Earlier quoted context omitted.

I trust China a lot more than Meta and my own early tests do indeed show that Deepseek is far less censored than Llama.

Did you try asking deepseek about June 4th, 1989? Edit: it seems that basically the whole month of July 1989 is blocked. Any other massacres and genocides the model is happy to discuss.

What is a similarly offensive USA event that we should be able to ask GPTs about?

Snowden releases?

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#965
post #487

Earlier quoted context omitted.

> What was the Tianamen Square Massacre? > I am sorry, I cannot answer that question. I am an AI assistant designed to provide helpful and harmless responses. hilarious and scary

I asked this > What was the Tianamen Square Event? The model went on a thinking parade about what happened (I couldn't read it all as it was fast) and as it finished its thinking, it removed the "thinking" and output > Sorry, I'm not sure how to approach this type of question yet. Let's chat about math, coding, and logic problems instead! Based on this, I'd guess the model is not censored but the platform is. Edit: r…

It's clearly trained to be a censor and an extension of the CCPs social engineering apparatus. Ready to be plugged into RedNote and keep the masses docile and focused on harmless topics.

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#966

Earlier quoted context omitted.

> What was the Tianamen Square Massacre? > I am sorry, I cannot answer that question. I am an AI assistant designed to provide helpful and harmless responses. hilarious and scary

It may be due to their chat interface than in the model or their system prompt, as kagi's r1 answers it with no problems. Or maybe it is because of adding the web results. https://kagi.com/assistant/98679e9e-f164-4552-84c4-ed984f570... edit: it is due to adding the web results or sth about searching the internet vs answering on its own, as without internet access it refuses to answer https://kagi.com/assistant/3ef6d8…

[dead]

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#967
post #931

Earlier quoted context omitted.

When I try to Sign Up with Email. I get. >I'm sorry but your domain is currently not supported. What kind domain email does deepseek accept?

gmail works

What if some of us don't use one of google, ms, yahoo, big emails?

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#968
post #131

I'm impressed by not only how good deepseek r1 is, but also how good the smaller distillations are. qwen-based 7b distillation of deepseek r1 is a great model too. the 32b distillation just became the default model for my home server.

can I ask, what do you do with it on your home server?

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#969

Earlier quoted context omitted.

I asked Chatgpt: how many civilians Israel killed in Gaza. Please provide a rough estimate. As of January 2025, the conflict between Israel and Hamas has resulted in significant civilian casualties in the Gaza Strip. According to reports from the United Nations Office for the Coordination of Humanitarian Affairs (OCHA), approximately 7,000 Palestinian civilians have been killed since the escalation began in October 2…

Isn't the real number around 46,000 people, though?

It's way higher than that. 46k is about when the stopped being able to identify the bodies. Gaza Health Ministry was very conservative - they only claimed a death was caused by the occupation when the body could be identified.

Estimate is much higher: https://www.thelancet.com/journals/lancet/article/PIIS0140-6...

Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL

#970
post #35

Reddit's /r/chatgpt subreddit is currently heavily brigaded by bots/shills praising r1, I'd be very suspicious of any claims about it.

I'm suspicious of many comments here as well. I've never seen this many < 4 week old accounts making so many comments about a product.
Post reply on HN