Earlier quoted context omitted.
If models become more efficient we will move more of the work to local devices instead of using SaaS models. We’re still in the mainframe era of LLM.
We moved from the mainframe era to desktops and smaller servers because computers got fast enough to do what we needed them to do locally. Centralized computing resources are still vastly more powerful than what's under your desk or in a laptop, but it doesn't matter because people generally don't need that much power for their daily tasks. The problem with AI is that it's not obvious what the upper limit of capabili…
What if AI doesn't need more RAM but better math?
71–80 of 111 posts
Re: What if AI doesn't need more RAM but better math?
#72Re: What if AI doesn't need more RAM but better math?
#73Can we please start talking about this in that context? We already know what TurboQuant will do to DRAM demand. We already know what it will do to context windows. There is no need to speculate. There is no need to panic sell stocks.
Re: What if AI doesn't need more RAM but better math?
#74Earlier quoted context omitted.
And when that happens people STILL won’t be able to afford the hardware.
> And when that happens people STILL won’t be able to afford the hardware. Of course they will - if that happens all these AI token providers won't have a use for all that hardware they bought. You'll be buying used H100s and H200s off eBay for pennies on the dollar.
Re: What if AI doesn't need more RAM but better math?
#75The drop in memory stocks seems counterintuitive to me. The demand for memory isn't going to go down, we'll just be able to do more with the same amount of memory.
Re: What if AI doesn't need more RAM but better math?
#76Earlier quoted context omitted.
If models become more efficient we will move more of the work to local devices instead of using SaaS models. We’re still in the mainframe era of LLM.
I don't see how we'll ever get to widespread local LLM. The power efficiency alone is a strong enough pressure to use centralized model providers. My 3090 running 24b or 32b models is fun, but I know I'm paying way more per token in electricity, on top of lower quality tokens. It's fun to run them locally, but for anything actually useful it's cheaper to just pay API prices currently.
Re: What if AI doesn't need more RAM but better math?
#77I mean, since GPT-4, I believe the RAM is no longer creating the miracle that the LLM performance scales directly with the model size. At least ChatGPT itself convinced me that any decent-sized company can create a GPT4 equivalent in terms of model size, but limited by service options, like memory cache and hallucination handling. Companies buy RAM simply to ride the stock hype. I am no expert, so this is a shallow t…
the models are still very stupid atm something needs to change
Re: What if AI doesn't need more RAM but better math?
#78Re: What if AI doesn't need more RAM but better math?
#79> applying this compression algorithm at scale may significantly relax the memory bottleneck issue. I don’t think they’re going to downsize though, I think the big players are just going to use the freed up memory for more workflows or larger models because the big players want to scale up. It’s a cat and mouse race for the best models.
It will also help with local inference, making AI without big players possible.