Live data from Hacker News

DSpark: Speculative decoding accelerates LLM inference [pdf]

github.com

331–340 of 393 posts

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#331
post #144

Earlier quoted context omitted.

The question is also what game they're playing. Deepseek came out of a hedge fund. I think it's no coincidence that their publications tend to have a large impact on AI stock prices. Destroying the growth story of overvalued stocks is an interesting investment strategy. It's not even new. Shortsellers understandably get terrible rep from execs, but their actions are more often in the public interest than you'd think.…

[flagged]

And OpenAI has the audacity to call themselves a non profit. All leading US models are closed source for one purpose, money.

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#332
post #30

Earlier quoted context omitted.

Chinese labs are also still behind, so they’re incentivized to collaborate and have no reason to do it in private. I suspect their tune will change if they ever take the lead..

Which is a good thing. Self-serving motives are more reliable than altruistic ones.

Agreed

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#333
post #22

Earlier quoted context omitted.

Publishing by necessity I wonder? American labs on the cutting edge pioneering the way forward, so Deepseek open sourcing what they’ve got is to help even the playing field. Hopefully the experts here can offer insight. The above is just my hunch and I’m not a specialist in this field.

I'm afraid I'm even balking at the word "pioneering" in context with US frontier labs. They are probably doing a few new things, right, but they are not blazing any trails for others to follow along, the Chinese are.

Or if the US labs are innovating, they're not talking about specifics.

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#334

Earlier quoted context omitted.

> The question is also what game they're playing. Deepseek came out of a hedge fund. I think it's no coincidence that their publications tend to have a large impact on AI stock prices. Its revealing that they always seem to publish after some big announcement by American AI companies. But regardless, this is one of the benefits of a duopoly.

No more revealing than OpenAI, Anthropic and Google always having some new model that just so happens to be waiting in the wings whenever their competitors announce their own model bump.

That's because OpenAI, Anthropic and Google work on many models in parallel which work cooperatively from the user's POV. So GPT-5.6 is just a checkpoint of their multi-model development.

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#335
post #63

Earlier quoted context omitted.

Projection is a funny thing. It causes people to misread situations all the time. Southern slaveowners feared violent retribution from freed slaves, for example [1]. It was pure projection and said more about the South than it did the slaves. The reality was there was no violent retribution. It was the opposite where the former slaveowners continued to inflict violence on the formerly enslaved. I say this because we…

It's even worse than that. China publishes stacks upon stacks of policy documents in which they explain clearly what they will do and why. This includes why they do poverty alleviation and why they believe big monopolies that own everything are bad. But almost no western observers care to read those documents. Instead, western observers, including HN, speculate endlessly about China's intentions, and "it would be nai…

Aspiration for the masses and policy reality are different things. They will say monopolies are bad, but what they mean is only if the state doesn't control them. The CCP has huge state run monopolies and there's nothing you can do about it. You can't understand their policy without understanding Marxism-Leninism. When pesky reality gets in the way of policy, it's called corruption.

In communist or post-communist countries like China and Russia, the percentage of government workers is extreme. To them, all social action is political action, and that includes economic. Since there is only one party allowed, any economic action (and thus political action) which threatens the CCP is unacceptable. They leave other private companies alone.

The CCP is bad and this conclusion is justified not only in theory as a distant observation of ideological concepts, but from their behaviors around the world which echo some of the causes of World War 2. That doesn't mean Chinese AI companies are bad, but the CCP will certainly find ways of using it for its purposes. For now, they're quite far behind on AI, but they deserve credit for optimizing for some use cases which masks poorer generalization.

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#336

Earlier quoted context omitted.

Please explain how distillation == innovation Especially since your 5-day-old account is sus, and thus likely not yet proven not to be a Chinese bot You can't lead by following the actual leader LOL The only real innovation I've seen from Deepseek is the out-loud reasoning thing in R1

You call yourself a philosopher in your profile.

I am also a fan of giving credit where it is due, and not giving it where it is not due. China is not an innovator. Perhaps they are only beginning to be, but historically, this is simply not the case, and yet distillation still falls squarely under "not innovating".

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#337
post #144

Earlier quoted context omitted.

The question is also what game they're playing. Deepseek came out of a hedge fund. I think it's no coincidence that their publications tend to have a large impact on AI stock prices. Destroying the growth story of overvalued stocks is an interesting investment strategy. It's not even new. Shortsellers understandably get terrible rep from execs, but their actions are more often in the public interest than you'd think.…

> The question is also what game they're playing. Deepseek came out of a hedge fund. I think it's no coincidence that their publications tend to have a large impact on AI stock prices. Its revealing that they always seem to publish after some big announcement by American AI companies. But regardless, this is one of the benefits of a duopoly.

[deleted]

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#338
post #63

Earlier quoted context omitted.

Projection is a funny thing. It causes people to misread situations all the time. Southern slaveowners feared violent retribution from freed slaves, for example [1]. It was pure projection and said more about the South than it did the slaves. The reality was there was no violent retribution. It was the opposite where the former slaveowners continued to inflict violence on the formerly enslaved. I say this because we…

It's even worse than that. China publishes stacks upon stacks of policy documents in which they explain clearly what they will do and why. This includes why they do poverty alleviation and why they believe big monopolies that own everything are bad. But almost no western observers care to read those documents. Instead, western observers, including HN, speculate endlessly about China's intentions, and "it would be nai…

> And they seem mostly happy with sticking to speculations that match preconceived notions

...you've almost achieved enlightenment on the nature of the majority of HN comments: vibes-based, off the cuff braindumps, where an idea is examined as it is being typed. It's great for tech and software discussions where many commenters have good knowledge or even mastery, but on "exotic" topics, an incorrect take can be voted to the top because it sounds right by affirming the biases of the majority. If you're a practitioner in a non-tech field and a topic in your field comes up for discussion on HN, be ready to be disappointed - and be ready to question all the other correct-sounding comments in areas unfamiliar to you.

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#339

Earlier quoted context omitted.

Hasn't that been the mantra of open source for 40 years. Armies of companies, trillions of valuation, or even just Wayland, suggest that isn't always the case.

And yet, Linux runs approximately every ounce of computing substrate on earth

The Linux Foundation was bankrolled by the US government (via grants and code donations) to undermine the EU Operating System industry. Symbian was going to be amazing, until Microsoft - an American company with government links - nuked it /s

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#340
post #63

Earlier quoted context omitted.

Projection is a funny thing. It causes people to misread situations all the time. Southern slaveowners feared violent retribution from freed slaves, for example [1]. It was pure projection and said more about the South than it did the slaves. The reality was there was no violent retribution. It was the opposite where the former slaveowners continued to inflict violence on the formerly enslaved. I say this because we…

> I say this because we see the same thing used as an argument against China. "If they overtake us, they'll do imperialism (like us)." Again, it says more about us than them. Or because they're human and that's what humans have always done. If the US is no longer a check on China, what will happen to Taiwan? Frankly, you seem to be arguing that the US is somehow uniquely bad when, in actuality, The US has been hegemo…

https://en.wikipedia.org/wiki/United_States_involvement_in_r...
Post reply on HN