DeepSeek-R1 has apparently caused quite a shock wave in SV ... https://venturebeat.com/ai/why-everyone-in-ai-is-freaking-ou...
Meta is in full panic last I heard. They have amassed a collection of pseudo experts there to collect their checks. Yet, Zuck wants to keep burning money on mediocrity. I’ve yet to see anything of value in terms products out of Meta.
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
211–220 of 1001 posts
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#212DeepSeek-R1 has apparently caused quite a shock wave in SV ... https://venturebeat.com/ai/why-everyone-in-ai-is-freaking-ou...
Meta is in full panic last I heard. They have amassed a collection of pseudo experts there to collect their checks. Yet, Zuck wants to keep burning money on mediocrity. I’ve yet to see anything of value in terms products out of Meta.
Llama models are also still best in class for specific tasks that require local data processing. They also maintain positions in the top 25 of the lmarena leaderboard (for what that's worth these days with suspected gaming of the platform), which places them in competition with some of the best models in the world.
But, going back to my first point, Llama set the stage for almost all open weights models after. They spent millions on training runs whose artifacts will never see the light of day, testing theories that are too expensive for smaller players to contemplate exploring.
Pegging Llama as mediocre, or a waste of money (as implied elsewhere), feels incredibly myopic.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#213Earlier quoted context omitted.
I don't say that at all. Money spent on BS still sucks resources, no matter who spends that money. They are not going to make the GPU's from 500 billion dollar banknotes, they will pay people $500B to work on this stuff which means people won't be working on other stuff that can actually produce value worth more than the $500B. I guess the power plants are salvageable.
Deepseek didn't train the model on sheets of paper, there are still infrastructure costs.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#214Aside from the usual Tiananmen Square censorship, there's also some other propaganda baked-in: https://prnt.sc/HaSc4XZ89skA (from reddit)
I ask O1 how to download a YouTube music playlist as a premium subscriber, and it tells me it can't help.
Deepseek has no problem.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#215I'm impressed by not only how good deepseek r1 is, but also how good the smaller distillations are. qwen-based 7b distillation of deepseek r1 is a great model too. the 32b distillation just became the default model for my home server.
tried the 7b, it switched to chinese mid-response
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#216DeepSeek-R1 has apparently caused quite a shock wave in SV ... https://venturebeat.com/ai/why-everyone-in-ai-is-freaking-ou...
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#217...and China is two years behind in AI. Right ?
And (some people here are saying that)* if they are up-to-date is because they're cheating. The copium itt is astounding.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#218Something like: collect some thoughts about this input; review the thoughts you created; create more thoughts if needed or provide a final answer; ...
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#219Earlier quoted context omitted.
Correct me if I'm wrong but if Chinese can produce the same quality at %99 discount, then the supposed $500B investment is actually worth $5B. Isn't that the kind wrong investment that can break nations? Edit: Just to clarify, I don't imply that this is public money to be spent. It will commission $500B worth of human and material resources for 5 years that can be much more productive if used for something else - i.e…
$500 billion is $500 billion. If new technology means we can get more for a dollar spent, then $500 billion gets more, not less.
Re: DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL
#220How can openai justify their $200/mo subscriptions if a model like this exists at an incredibly low price point? Operator? I've been impressed in my brief personal testing and the model ranks very highly across most benchmarks (when controlled for style it's tied number one on lmarena). It's also hilarious that openai explicitly prevented users from seeing the CoT tokens on the o1 model (which you still pay for btw)…
From my casual read, right now everyone is on reputation tarnishing tirade, like spamming “Chinese stealing data! Definitely lying about everything! API can’t be this cheap!”. If that doesn’t go through well, I’m assuming lobbyism will start for import controls, which is very stupid. I have no idea how they can recover from it, if DeepSeek’s product is what they’re advertising.
Somehow I doubt it.