Earlier quoted context omitted.
Agreed. I am no fan of the CCP but I have no issue with using DeepSeek since I only need to use it for coding which it does quite well. I still believe Sonnet is better. DeepSeek also struggles when the context window gets big. This might be hardware though. Having said that, DeepSeek is 10 times cheaper than Sonnet and better than GPT-4o for my use cases. Models are a commodity product and it is easy enough to add a…
Curious why you have to qualify this with a “no fan of the CCP” prefix. From the outset, this is just a private organization and its links to CCP aren’t any different than, say, Foxconn’s or DJI’s or any of the countless Chinese manufacturers and businesses You don’t invoke “I’m no fan of the CCP” before opening TikTok or buying a DJI drone or a BYD car. Then why this, because I’ve seen the same line repeated everywh…
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
351–360 of 1001 posts
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#352Earlier quoted context omitted.
Great as long as you’re not interested in Tiananmen Square or the Uighurs.
try asking US models about the influence of Israeli diaspora on funding genocide in Gaza then come back
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#353Earlier quoted context omitted.
Great as long as you’re not interested in Tiananmen Square or the Uighurs.
Have you even tried it out locally and asked about those things?
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#354DeepSeek-R1 has apparently caused quite a shock wave in SV ... https://venturebeat.com/ai/why-everyone-in-ai-is-freaking-ou...
Correct me if I'm wrong but if Chinese can produce the same quality at %99 discount, then the supposed $500B investment is actually worth $5B. Isn't that the kind wrong investment that can break nations? Edit: Just to clarify, I don't imply that this is public money to be spent. It will commission $500B worth of human and material resources for 5 years that can be much more productive if used for something else - i.e…
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#355DeepSeek-R1 has apparently caused quite a shock wave in SV ... https://venturebeat.com/ai/why-everyone-in-ai-is-freaking-ou...
Correct me if I'm wrong but if Chinese can produce the same quality at %99 discount, then the supposed $500B investment is actually worth $5B. Isn't that the kind wrong investment that can break nations? Edit: Just to clarify, I don't imply that this is public money to be spent. It will commission $500B worth of human and material resources for 5 years that can be much more productive if used for something else - i.e…
It's such a weird question. You made it sound like 1) the $500B is already spent and wasted. 2) infrastructure can't be repurposed.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#356Larry Ellison is 80. Masayoshi Son is 67. Both have said that anti-aging and eternal life is one of their main goals with investing toward ASI. For them it's worth it to use their own wealth and rally the industry to invest $500 billion in GPUs if that means they will get to ASI 5 years faster and ask the ASI to give them eternal life.
Side note: I’ve read enough sci-fi to know that letting rich people live much longer than not rich is a recipe for a dystopian disaster. The world needs incompetent heirs to waste most of their inheritance, otherwise the civilization collapses to some kind of feudal nightmare.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#357DeepSeek V3 came in the perfect time, precisely when Claude Sonnet turned into crap and barely allows me to complete something without me hitting some unexpected constraints. Idk, what their plans is and if their strategy is to undercut the competitors but for me, this is a huge benefit. I received 10$ free credits and have been using Deepseeks api a lot, yet, I have barely burned a single dollar, their pricing are t…
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#358Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#359Afaict they’ve hidden them primarily to stifle the competition… which doesn’t seem to matter at present!