Live data from Hacker News

TurboQuant: Redefining AI efficiency with extreme compression

research.google

201–202 of 202 posts

Re: TurboQuant: Redefining AI efficiency with extreme compression

#201

I feel like I’m not the only who feel excited about the whole “compression” tricks while maintaining fidelity in our AI era. In a way, it has a vibe similar to the early 2000s when digital music became popular and the need for lossless compression was paramount. Sort of a pied piper moment for us now . Someone please make a Weisseman score for this stuff.

[deleted]

Re: TurboQuant: Redefining AI efficiency with extreme compression

#202

This is a great development for KV cache compression. I did notice a missing citation in the related works regarding the core mathematical mechanism, though. The foundational technique of applying a geometric rotation prior to extreme quantization, specifically for managing the high-dimensional geometry and enabling proper bias correction, was introduced in our NeurIPS 2021 paper, "DRIVE" ( https://proceedings.neurip…

Check out the most recent comment about the paper on OpenReview. This doesn't seem like isolated behavior:

https://openreview.net/forum?id=tO3ASKZlok

Post reply on HN