Live data from Hacker News

DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

arxiv.org

1–10 of 37 posts

Re: DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

#2
Not jus being impressed that every paper coming out is SOTA, but also leads the way in being Open-Source in the pure definition of OSS, even with permissible licensing.

Let's not confuse the company with the country by over-fitting a narrative. Popular media is reenforcing hatred or anything that sponsors them, especially to weaker groups. Less repercussions and more clicks/money to be made I guess.

While Politicians may hate each other, Scientists love to work with other aspiring Scientists who have similar ambitions and the only competition is in achieving measurable success and the reward it means to the greater public.

Without any bias, but it's genuinely admirable when companies release their sources to enable faster scientific progress cycles. It's ironic that this company is dedicated to finance, yet shares their progress, while non-profits and companies dedicated purely to AI are locking all knowledge about their findings from access.

Are there other companies like DeepSeek that you know of that commonly release great papers? I am following Mistral already, but I'd love to enrich my sources of publications that I consume. Highly appreciated!

Re: DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

#3
DeepSeek R1 is by far the best at writing prose of any model, including Grok-3, GPT-4o, o1-pro, o3, claude, etc.

Paste in a snippet from a book and ask the model to continue the story in the style of the snippet. It's surprising how bad most of the models are.

Grok-3 comes in a close second, likely because it is actually DeepSeek R1 with a few mods behind the scenes.

Re: DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

#5
post #2

Not jus being impressed that every paper coming out is SOTA, but also leads the way in being Open-Source in the pure definition of OSS, even with permissible licensing. Let's not confuse the company with the country by over-fitting a narrative. Popular media is reenforcing hatred or anything that sponsors them, especially to weaker groups. Less repercussions and more clicks/money to be made I guess. While Politicians…

When OpenAI surged ahead Meta ended up giving away its incredibly expensive to make llama model to reduce the OpenAI valuations.

Is DeepSeeks openness in part to reduce the big American tech companies?

Re: DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

#6
post #3

DeepSeek R1 is by far the best at writing prose of any model, including Grok-3, GPT-4o, o1-pro, o3, claude, etc. Paste in a snippet from a book and ask the model to continue the story in the style of the snippet. It's surprising how bad most of the models are. Grok-3 comes in a close second, likely because it is actually DeepSeek R1 with a few mods behind the scenes.

why do you think that grok 3 is deepseek, out of curiosity?

Re: DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

#7
post #2

Not jus being impressed that every paper coming out is SOTA, but also leads the way in being Open-Source in the pure definition of OSS, even with permissible licensing. Let's not confuse the company with the country by over-fitting a narrative. Popular media is reenforcing hatred or anything that sponsors them, especially to weaker groups. Less repercussions and more clicks/money to be made I guess. While Politicians…

When OpenAI surged ahead Meta ended up giving away its incredibly expensive to make llama model to reduce the OpenAI valuations. Is DeepSeeks openness in part to reduce the big American tech companies?

Correlation isn't causation, I hate to say this, but here's really applicable. Facebook aka Meta has always been very opensource. Let's not talk about the license though. :)

Why do you imply malice in OSS companies? Or for profit companies opensourcing their models and sourcecode?

Re: DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

#8
post #2

Not jus being impressed that every paper coming out is SOTA, but also leads the way in being Open-Source in the pure definition of OSS, even with permissible licensing. Let's not confuse the company with the country by over-fitting a narrative. Popular media is reenforcing hatred or anything that sponsors them, especially to weaker groups. Less repercussions and more clicks/money to be made I guess. While Politicians…

> Let's not confuse the company with the country

What's wrong with China? They're wonderful in the OSS ecosystem.

Re: DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

#9
post #2

Not jus being impressed that every paper coming out is SOTA, but also leads the way in being Open-Source in the pure definition of OSS, even with permissible licensing. Let's not confuse the company with the country by over-fitting a narrative. Popular media is reenforcing hatred or anything that sponsors them, especially to weaker groups. Less repercussions and more clicks/money to be made I guess. While Politicians…

When OpenAI surged ahead Meta ended up giving away its incredibly expensive to make llama model to reduce the OpenAI valuations. Is DeepSeeks openness in part to reduce the big American tech companies?

If only totalitarian nation states used their subjects' money to undermine the dominance of US-based software vendors by releasing open-source alternatives created with slave labour... Oh wait, it can't work because software patents are here to the rescue again ... Wait, open source is communism? Always has been. /s

Re: DeepSeek: Inference-Time Scaling for Generalist Reward Modeling

#10
post #7

Earlier quoted context omitted.

When OpenAI surged ahead Meta ended up giving away its incredibly expensive to make llama model to reduce the OpenAI valuations. Is DeepSeeks openness in part to reduce the big American tech companies?

Correlation isn't causation, I hate to say this, but here's really applicable. Facebook aka Meta has always been very opensource. Let's not talk about the license though. :) Why do you imply malice in OSS companies? Or for profit companies opensourcing their models and sourcecode?

Personally I don't impute any malice whatsoever -- these are soulless corporate entities -- but a for-profit company with fiduciary duty to shareholders releasing expensive, in-house-developed intellectual property for free certainly deserves some scrutiny.

I tend to believe this is a "commoditize your complement" strategy on Meta's part, myself. No idea what Deepseek's motivation is, but it wouldn't surprise me if it was a similar strategy.

Post reply on HN