Live data from Hacker News

Ultrafast machine learning on FPGAs via Kolmogorov-Arnold Networks

aarushgupta.io

11–20 of 58 posts

Re: Ultrafast machine learning on FPGAs via Kolmogorov-Arnold Networks

#11
post #10
post #7

Earlier quoted context omitted.

Yes, this work is focused on accelerating very small models, typically for real-time systems that require extremely low power or low latency. One primary application of this work is in high-energy physics ( https://home.cern/smarter-decisions-at-the-speed-of-collisio... ). Ultrafast and real-time learning is also very applicable for problems in quantum computing, plasma control, etc. ( https://arxiv.org/pdf/2602.0200…

I'm not in HFT, but I assume this is also an interesting applicable domain?

Yes, definitely: this type of work is applicable in domains where software run on general-purpose processors cannot meet latency or power requirements.

Re: Ultrafast machine learning on FPGAs via Kolmogorov-Arnold Networks

#12
post #8

Archive link, as it looks like the original post was taken down: https://web.archive.org/web/20260609200156/https://aarushgup...

Hmm the post is still up for me?

For us too, but we'll put the archive link in the toptext since these things seem to vary a lot by region.

p.s. Thanks for posting this and welcome to HN!

Re: Ultrafast machine learning on FPGAs via Kolmogorov-Arnold Networks

#14
So for people wondering if it can be used to accelerate LLM inference, sadly not.

I've been trying to hit 100,000tokens/s with a 3.28m dumb model, and even this is an order of magnitude too large to benefit.

It appears to be focussed more on latency, than throughput. Happy to be corrected?

Re: Ultrafast machine learning on FPGAs via Kolmogorov-Arnold Networks

#15

So for people wondering if it can be used to accelerate LLM inference, sadly not. I've been trying to hit 100,000tokens/s with a 3.28m dumb model, and even this is an order of magnitude too large to benefit. It appears to be focussed more on latency, than throughput. Happy to be corrected?

You're correct that this work is not very applicable for LLMs and that the focus here is primarily on latency.

Re: Ultrafast machine learning on FPGAs via Kolmogorov-Arnold Networks

#17
post #10
post #7

Earlier quoted context omitted.

Yes, this work is focused on accelerating very small models, typically for real-time systems that require extremely low power or low latency. One primary application of this work is in high-energy physics ( https://home.cern/smarter-decisions-at-the-speed-of-collisio... ). Ultrafast and real-time learning is also very applicable for problems in quantum computing, plasma control, etc. ( https://arxiv.org/pdf/2602.0200…

I'm not in HFT, but I assume this is also an interesting applicable domain?

The author actually works at Jane Street.
Post reply on HN