Earlier quoted context omitted.
It's probably not long till frontier AI companies automate AI research. Then we get recursive self-improvement and eventually superintelligence. The singularity is near. Only a few years perhaps.
Forgot the /s
Autoresearch: Agents researching on single-GPU nanochat training automatically
31–40 of 66 posts
Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#32Earlier quoted context omitted.
this is very far from hyperparameter tuning in at least three important ways: - it can modify code arbitrarily, the notion of a "hyperparameter" dissolves - there is no need to run "sweeps" - this is the standard parallel process that wastes compute. because LLM agents are sequential, they can do more efficient versions such as binary search to narrow in on the right setting very quickly (usually many parameters will…
How about the very last "Kept Improvement" in the plot? It's titled "random seed 42 -> 137". I do think this project is quite conceptually interesting, but the model literally choosing a different random seed to achieve lower loss feels pretty far removed from the flowery sci-fi writing at the top of the readme.
Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#33Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#34Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#35As ai improves, most tasks will become something like this. Environments setup where the model learns through trial and error Any human endeavor that can be objectively verified in some environment like this can be completely automated
Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#36As ai improves, most tasks will become something like this. Environments setup where the model learns through trial and error Any human endeavor that can be objectively verified in some environment like this can be completely automated
People make fun of prompt engineering, but I think "AI ops" will eventually become a real role at most if not all software companies. Harness Engineers and Agent Reliability Engineers will be just as important as something like DevOps is now.
Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#37Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#38Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#39As ai improves, most tasks will become something like this. Environments setup where the model learns through trial and error Any human endeavor that can be objectively verified in some environment like this can be completely automated
So much this. People make fun of prompt engineering, but I think "AI ops" will eventually become a real role at most if not all software companies. Harness Engineers and Agent Reliability Engineers will be just as important as something like DevOps is now.
Re: Autoresearch: Agents researching on single-GPU nanochat training automatically
#40Earlier quoted context omitted.
It's probably not long till frontier AI companies automate AI research. Then we get recursive self-improvement and eventually superintelligence. The singularity is near. Only a few years perhaps.
Forgot the /s
I think the most disappointing thing will be that even we do achieve ASI, everything will carry on as business as usual for a while before it starts making an economic impact because of how resistant to change we have made society.