Prediction - we are going to figure out SOTA AI performance without requiring 1TB of memory within a year or so. Of course CXMT, Micron, and family will still be profitable, but maybe not 'surge 470% from IPO' profitable.
How would we do that? There's no historic precedent for that. Its fundamentally an information theory thing: what's the max amount of intelligence you can get out of 1 KB/MB/GB? There has to be a limit and I'm not convinced that it's far off.
LLMs haven’t been on the scene for very long. There’s still lots of room for efficiency discoveries. Plus, I keep hearing quantum is going to be a big deal in the next few years.