Live data from Hacker News

DSpark: Speculative decoding accelerates LLM inference [pdf]

github.com

181–190 of 393 posts

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#181

Would love to see these numbers reproduced on consumer GPUs, not just A100s.

This is an efficiency improvement that significantly lowers the amount of RAM you have to look at, on average, during decode.

It should improve performance on most hardware because most LLMs are memory bandwidth bound during decode.

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#182
Title is bad, it's the first line of the abstract instead of the paper title. Speculative decoding for LLM inference was published in 2022: https://arxiv.org/abs/2211.17192

This paper seems to be an improvement to speculative decoding but I haven't read it yet.

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#184
post #2

Nice. Guessing the timing isn't accidental. Demonstrated openness vs harsh regulation

China = Open. US = Harsh Regulation Strange timeline, though this only works because it’s aligned with Xi’s goals.

Yeah can definitely see a world where china pivots and we're stuck with closed/closed

Mistral...don't fumble this

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#185

DeepSeek continues to not only push the boundaries but also publish these incredible papers explaining how they achieved their gains - something the American labs no longer do unfortunately. Chinese labs are doing the most interesting work in AI right now.

Thank you so much to everyone at DeepSeek who is working on this and who have the courage and generosity to open source this for humanity.

We in the United States will never forget!

For all the harm Trump does to the US at least he is helping China!

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#186
post #30

Earlier quoted context omitted.

Chinese labs are also still behind, so they’re incentivized to collaborate and have no reason to do it in private. I suspect their tune will change if they ever take the lead..

Which is a good thing. Self-serving motives are more reliable than altruistic ones.

[deleted]

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#187
post #94
post #87

Earlier quoted context omitted.

They are self financed, the company that makes DeepSeek is a finance company that trades on the markets.

The CCP's approach has historically been to subsidize their companies far more than other countries do. Why would LLMs be any different? https://www.oecd.org/en/data/dashboards/magic-database-indus...

According to EU statistics, yeah

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#188

DeepSeek is, as I feel currently, the sole AI company which is actually trying to innovate rather than top mere benchmarks. Others like OpenAI, Anthropic and Google are mostly just competeing with each rather than keep innovating around the clock.

Besides the founder, the only real external investor for DeepSeek is Chinese govt. there are literally zero revenue pressure compare to O, A & G.

To compete in that direction, USG needs to learn from CCP to "seize the means of production", which they are sort of doing, but in such an incompetent way that I'm afraid we will probably end up mixing the worst of both communism and capitalism.

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#189

Earlier quoted context omitted.

Probably because American AI companies are on the hook for quite a lot of investment money. I think they are trying to find the magical moat to justify their valuation. Revealing optimizations similar to these would pretty much reduce their competitive position.

Chinese labs are also still behind, so they’re incentivized to collaborate and have no reason to do it in private. I suspect their tune will change if they ever take the lead..

True!

Re: DSpark: Speculative decoding accelerates LLM inference [pdf]

#190
post #132
post #126

Earlier quoted context omitted.

Not everyone is motivated by greed

What do you think is the underlaying motivation?

You ask me what I think. So far deepseek has been very consistently trying to advance state of the art research in a transplant and public way by writing papers and publishing working code. They are also not at the mercy of the stock market in the same way many Americans companies are. Before anyone assumes too much, I live in Europe.
Post reply on HN