Live data from Hacker News

Kolmogorov-Arnold Networks

github.com

1–10 of 149 posts

Re: Kolmogorov-Arnold Networks

#4
post #3

Interesting! Would this approach (with non-linear learning) still be able to utilize GPUs to speed up training?

Seconded. I’m guessing you could create an implementation that is able to do that and then write optimised triton/cuda kernels to accelerate them but need to investigate further

Re: Kolmogorov-Arnold Networks

#7
I can't assess this, but I do worry that overnight some algorithmic advance will enhance LLMs by orders of magnitude and the next big model to get trained is suddenly 10,000x better than GPT-4 and nobody's ready for it.

Re: Kolmogorov-Arnold Networks

#9

I can't assess this, but I do worry that overnight some algorithmic advance will enhance LLMs by orders of magnitude and the next big model to get trained is suddenly 10,000x better than GPT-4 and nobody's ready for it.

>some algorithmic advance will enhance LLMs by orders of magnitude

I would worry if I'd own Nvidia shares.

Re: Kolmogorov-Arnold Networks

#10

I can't assess this, but I do worry that overnight some algorithmic advance will enhance LLMs by orders of magnitude and the next big model to get trained is suddenly 10,000x better than GPT-4 and nobody's ready for it.

What to be worried about? Technical progress will happen, sometimes by sudden jumps. Some company will become a leader, competitors will catch up after a while.
Post reply on HN