...and China is two years behind in AI. Right ?
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
171–180 of 1001 posts
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#172"Reasoning" will be disproven for this again within a few days I guess. Context: o1 does not reason, it pattern matches. If you rename variables, suddenly it fails to solve the request.
These models can and do work okay with variable names that have never occurred in the training data. Though sure, choice of variable names can have an impact on the performance of the model.
That's also true for humans, go fill a codebase with misleading variable names and watch human programmers flail. Of course, the LLM's failure modes are sometimes pretty inhuman, -- it's not a human after all.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#173Larry Ellison is 80. Masayoshi Son is 67. Both have said that anti-aging and eternal life is one of their main goals with investing toward ASI. For them it's worth it to use their own wealth and rally the industry to invest $500 billion in GPUs if that means they will get to ASI 5 years faster and ask the ASI to give them eternal life.
Probably shouldn't be firing their blood boys just yet ... According to Musk, SoftBank only has $10B available for this atm.
He says stuff that’s wrong all the time with extreme certainty.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#174Earlier quoted context omitted.
It’s not just the economy that is vulnerable, but global geopolitics. It’s definitely worrying to see this type of technology in the hands of an authoritarian dictatorship, especially considering the evidence of censorship. See this article for a collected set of prompts and responses from DeepSeek highlighting the propaganda: https://medium.com/the-generator/deepseek-hidden-china-polit... But also the claimed cost i…
have you tried asking chatgpt something even slightly controversial? chatgpt censors much more than deepseek does. also deepseek is open-weights. there is nothing preventing you from doing a finetune that removes the censorship. they did that with llama2 back in the day.
This is an outrageous claim with no evidence, as if there was any equivalence between government enforced propaganda and anything else. Look at the system prompts for DeepSeek and it’s even more clear.
Also: fine tuning is not relevant when what is deployed at scale brainwashes the masses through false and misleading responses.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#175DeepSeek-R1 has apparently caused quite a shock wave in SV ... https://venturebeat.com/ai/why-everyone-in-ai-is-freaking-ou...
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#176Aside from the usual Tiananmen Square censorship, there's also some other propaganda baked-in: https://prnt.sc/HaSc4XZ89skA (from reddit)
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#177Earlier quoted context omitted.
It does if the spend drives GPU prices so high that more researchers can't afford to use them. And DS demonstrated what a small team of researchers can do with a moderate amount of GPUs.
The DS team themselves suggest large amounts of compute are still required
GPU prices could be a lot lower and still give the manufacturer a more "normal" 50% gross margin and the average researcher could afford more compute. A 90% gross margin, for example, would imply that price is 5x the level that that would give a 50% margin.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#178I've always been leery about outrageous GPU investments, at some point I'll dig through and find my prior comments where I've said as much to that effect. The CEOs, upper management, and governments derive their importance on how much money they can spend - AI gave them the opportunity for them to confidently say that if you give me $X I can deliver Y and they turn around and give that money to NVidia. The problem wa…
I think you are underestimating the fear of being beaten (for many people making these decisions, "again") by a competitor that does "dumb scaling".
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#179...and China is two years behind in AI. Right ?
Now maybe 4? It's hard to say.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#180Aside from the usual Tiananmen Square censorship, there's also some other propaganda baked-in: https://prnt.sc/HaSc4XZ89skA (from reddit)
Apparently the censorship isn't baked-in to the model itself, but rather is overlayed in the public chat interface. If you run it yourself, it is significantly less censored [0] [0] https://thezvi.substack.com/p/on-deepseeks-r1?open=false#%C2...