Earlier quoted context omitted.
I think his point is hand wavy at best. It presupposes infinite scaling and ignores all the algorithmic efficiency wins that are being discovered. Ironically, many of which are being discovered with autoresearch style workflows, using the very LLMs that his company builds. The #1 post on HN right now[1] is full of people jubilating about how they can run Qwen 3.8 27B on their > 5 year old GPUs. If that isn't democrat…
Which percentage of people have GPUs capable of running Qwen3.8 27B? I am one of those, and for my job I am still resorting to hyperscalers because tasks are completed faster and more accurately that way. Even if we assume that models will no longer improve and we reach a point where everyone can run Fable in their laptop, surely running 1000x Fable agents would give you an advantage. I think access to compute will m…
I'm not sure whether that's low or high, or how it compares to a general audience.