Earlier quoted context omitted.
This is exactly backwards: Dennard scaling stopped. Moore’s Law has continued and it’s what made training and running inference on these models practical at interactive timescales.
You are technically correct. The best kind of correct. However, most people don't know the difference between the proper Moore's Law scaling (the cost of a transistor halves every 2 years) which is still continuing (sort of) and the colloquial version (the speed of a transistor doubles every 2 years) which got broken when Dennard scaling ran out. To them, Moore's Law just broke. Nevertheless, you are reinforcing my p…
And in fact, the agentic looped LLMs are executing much better than that today. They could stop advancing right now and still be revolutionary.