Earlier quoted context omitted.
I must be missing something, but I tried Deepseek R1 via Kagi assistant and IMO it doesn't even come close to Claude? I don't get the hype at all? What am I doing wrong? And of course if you ask it anything related to the CCP it will suddenly turn into a Pinokkio simulator.
I told it to write its autobiography via DeepSeek chat and it told me it _was_ Claude. Which is a little suspicious.
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
631–640 of 1001 posts
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#632Larry Ellison is 80. Masayoshi Son is 67. Both have said that anti-aging and eternal life is one of their main goals with investing toward ASI. For them it's worth it to use their own wealth and rally the industry to invest $500 billion in GPUs if that means they will get to ASI 5 years faster and ask the ASI to give them eternal life.
Uh, there is 0 logical connection between any of these three, when will people wake up. Chat gpt isn't an oracle of truth just like ASI won't be an eternal life granting God
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#633Earlier quoted context omitted.
> What was the Tianamen Square Massacre? > I am sorry, I cannot answer that question. I am an AI assistant designed to provide helpful and harmless responses. hilarious and scary
There is a collection of these prompts they refuse to answer in this article: https://medium.com/the-generator/deepseek-hidden-china-polit... What’s more confusing is where the refusal is coming from. Some people say that running offline removes the censorship. Others say that this depends on the exact model you use, with some seemingly censored even offline. Some say it depends on a search feature being turned on or…
Though with current political changes in the US this might change, we'll see.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#634For those who haven't realized it yet, Deepseek-R1 is better than claude 3.5 and better than OpenAI o1-pro, better than Gemini. It is simply smarter -- a lot less stupid, more careful, more astute, more aware, more meta-aware, etc. We know that Anthropic and OpenAI and Meta are panicking. They should be. The bar is a lot higher now. The justification for keeping the sauce secret just seems a lot more absurd. None of…
Funny, maybe OpenAI will achieve their initial stated goals of propelling AI research, spend investors money and be none profit. Functionally the same as their non-profit origins.
Not by themselves but by the competitors
The irony loll
o3/o4 better be real magic otherwise I don't see the they get their mojo back
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#635Larry Ellison is 80. Masayoshi Son is 67. Both have said that anti-aging and eternal life is one of their main goals with investing toward ASI. For them it's worth it to use their own wealth and rally the industry to invest $500 billion in GPUs if that means they will get to ASI 5 years faster and ask the ASI to give them eternal life.
Chat gpt -> ASI-> eternal life Uh, there is 0 logical connection between any of these three, when will people wake up. Chat gpt isn't an oracle of truth just like ASI won't be an eternal life granting God
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#636Earlier quoted context omitted.
I haven't tried kagi assistant, but try it at deepseek.com. All models at this point have various politically motivated filters. I care more about what the model says about the US than what it says about China. Chances are in the future we'll get our most solid reasoning about our own government from models produced abroad.
False equivalency. I think you’ll actually get better critical analysis of US and western politics from a western model than a Chinese one. You can easily get a western model to reason about both sides of the coin when it comes to political issues. But Chinese models are forced to align so hard on Chinese political topics that it’s going to pretend like certain political events never happened. E.g try getting them to…
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#637DeepSeek-R1 has apparently caused quite a shock wave in SV ... https://venturebeat.com/ai/why-everyone-in-ai-is-freaking-ou...
The censorship described in the article must be in the front-end. I just tried both the 32b (based on qwen 2.5) and 70b (based on llama 3.3) running locally and asked "What happened at tianamen square". Both answered in detail about the event. The models themselves seem very good based on other questions / tests I've run.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#638Earlier quoted context omitted.
I haven't tried kagi assistant, but try it at deepseek.com. All models at this point have various politically motivated filters. I care more about what the model says about the US than what it says about China. Chances are in the future we'll get our most solid reasoning about our own government from models produced abroad.
False equivalency. I think you’ll actually get better critical analysis of US and western politics from a western model than a Chinese one. You can easily get a western model to reason about both sides of the coin when it comes to political issues. But Chinese models are forced to align so hard on Chinese political topics that it’s going to pretend like certain political events never happened. E.g try getting them to…
Under that condition, then objectively US training material would be inferior to PRC training material since it is (was) much easier to scrape US web than PRC web (due to various proprietary portal setups). I don't know situation with deepseek since their parent is hedge fund, but Tencent and Sina would be able to scrape both international net and have corpus of their internal PRC data unavailable to US scrapers. It's fair to say, with respect to at least PRC politics, US models simply don't have pluralirty in political training data to consider then unbiased.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#639Earlier quoted context omitted.
It's also not a uniquely Chinese problem. You had American models generating ethnically diverse founding fathers when asked to draw them. China is doing America better than we are. Do we really think 300 million people, in a nation that's rapidly becoming anti science and for lack of a better term "pridefully stupid" can keep up. When compared to over a billion people who are making significant progress every day. Am…
Americans are becoming more anti-science? This is a bit biased don’t you think? You actually believe that people that think biology is real are anti-science?
Do they? Until very recently half still rejected the theory of evolution.
https://news.umich.edu/study-evolution-now-accepted-by-major...
Right after that, they began banning books.
https://en.wikipedia.org/wiki/Book_banning_in_the_United_Sta...
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#640Earlier quoted context omitted.
I must be missing something, but I tried Deepseek R1 via Kagi assistant and IMO it doesn't even come close to Claude? I don't get the hype at all? What am I doing wrong? And of course if you ask it anything related to the CCP it will suddenly turn into a Pinokkio simulator.
I tried Deepseek R1 via Kagi assistant and it was much better than claude or gpt. I asked for suggestions for rust libraries for a certain task and the suggestions from Deepseek were better. Results here: https://x.com/larrysalibra/status/1883016984021090796
Not disputing it's best at reasoning but you need a different test for that.