Is it possible that after 3-4 years of performance optimizations, both algorithmic and in hardware efficiency, it will turn out that we didn’t really need all of the nuclear plants we’re currently in the process of setting up to satisfy the power demands of AI data centers?
Congrats, you have independently reinvented the Hardware Overhang hypothesis: that early AGI could be very inefficient, undergo several optimization passes, and go from needing a datacenter of compute to, say, a single video game console's worth: https://www.lesswrong.com/posts/75dnjiD8kv2khe9eQ/measuring-... In that scenario, you can go from 0 independent artificial intelligences to tens of millions of them, very qu…
New LLM optimization technique slashes memory costs
61–70 of 227 posts
Re: New LLM optimization technique slashes memory costs
#62Is it possible that after 3-4 years of performance optimizations, both algorithmic and in hardware efficiency, it will turn out that we didn’t really need all of the nuclear plants we’re currently in the process of setting up to satisfy the power demands of AI data centers?
Congrats, you have independently reinvented the Hardware Overhang hypothesis: that early AGI could be very inefficient, undergo several optimization passes, and go from needing a datacenter of compute to, say, a single video game console's worth: https://www.lesswrong.com/posts/75dnjiD8kv2khe9eQ/measuring-... In that scenario, you can go from 0 independent artificial intelligences to tens of millions of them, very qu…
Re: New LLM optimization technique slashes memory costs
#63Earlier quoted context omitted.
True. Microsoft's all in, Apple's all in, Nvidia is selling shovels, insurance companies are all in, police & military are all in, education is all in, office management is all in. Who is left to pump line up?
no one is successfully using LLMs for anything other than customer service related things and text generation(coding, writing)
Re: New LLM optimization technique slashes memory costs
#64Re: New LLM optimization technique slashes memory costs
#65Is it possible that after 3-4 years of performance optimizations, both algorithmic and in hardware efficiency, it will turn out that we didn’t really need all of the nuclear plants we’re currently in the process of setting up to satisfy the power demands of AI data centers?
Nobody is building nuclear power plants for data centres. A few people have signed some paperwork saying that they would buy electricity from new nuclear plants if they could deliver it at a certain price, a price mind you that has not been done before. Others are trying to restart an existing reactor at three mile island (a thing that has never been done before, and likely won't be done now since the reactor was shu…
https://www.theguardian.com/technology/2024/sep/15/data-cent...
Re: New LLM optimization technique slashes memory costs
#66Earlier quoted context omitted.
We are putting lots of optimisation efforts into lots of worthwhile endeavours.
I dunno, software seems to be getting worse, hardware is getting more expensive and both Microsoft and Apple are distracted by AI, not to mention NVIDIA who seem to have bet the farm on Deus Ex Shovel
Though I had thought you were talking about stuff like eg producing more corn on a given piece of land, or making more furniture from less wood or so. Or even just making better batteries and solar cells.
Re: New LLM optimization technique slashes memory costs
#67Is it possible that after 3-4 years of performance optimizations, both algorithmic and in hardware efficiency, it will turn out that we didn’t really need all of the nuclear plants we’re currently in the process of setting up to satisfy the power demands of AI data centers?
Example:
1. To decrease total gas consumption, more fuel efficient vehicles are invented.
2. Instead of using less gas, people drive more miles. They take longer road trips, commute farther for work, and more people can now afford to drive.
3. This increased driving leads to higher overall gasoline consumption, despite each car using gas more efficiently.
Re: New LLM optimization technique slashes memory costs
#68Is it possible that after 3-4 years of performance optimizations, both algorithmic and in hardware efficiency, it will turn out that we didn’t really need all of the nuclear plants we’re currently in the process of setting up to satisfy the power demands of AI data centers?
If we run 7B now, why wouldn't we run 700b with memory optimizations?
Re: New LLM optimization technique slashes memory costs
#69Earlier quoted context omitted.
Like what?
Oh, I don’t know, how about reducing the search space/accelerating the search speed for potential room temperature superconductors? Or how about the same for viable battery chemistries?
and what if it's a dead end?
Re: New LLM optimization technique slashes memory costs
#70Is it possible that after 3-4 years of performance optimizations, both algorithmic and in hardware efficiency, it will turn out that we didn’t really need all of the nuclear plants we’re currently in the process of setting up to satisfy the power demands of AI data centers?
Nobody is building nuclear power plants for data centres. A few people have signed some paperwork saying that they would buy electricity from new nuclear plants if they could deliver it at a certain price, a price mind you that has not been done before. Others are trying to restart an existing reactor at three mile island (a thing that has never been done before, and likely won't be done now since the reactor was shu…