Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
81–90 of 126 posts
Re: Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
#82The most surprising part: the agent had access to both H100s and H200s. Without being told, it noticed H200s scored better and started screening ideas on H100s, then promoting winners to H200s for validation. That strategy emerged entirely on its own.
Why do we think this emerged “on its own”? Surely this technique has been discussed in research papers that are in the training set.
Re: Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
#83The US governement should 'autoresearch' a way to charge this man for his crimes as head of autopilot at tesla.
Re: Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
#84The US governement should 'autoresearch' a way to charge this man for his crimes as head of autopilot at tesla.
Can you elaborate?
i don't know a great deal about the guy. i know: he worked at tesla, led autopilot there. if we ignore the character defects required to work at tesla, he's responsible for designing systems that would certainly kill people because they decided lidar was too expensive.
Re: Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
#85Earlier quoted context omitted.
There is no burden of proof on me, because I'm not asserting that AI has invented something on its own. I haven't told you what my view is or whether I ever have a view. The problem with the reasoning of the person I was responding to is that it's assuming "if X is in the training set and LLM outputs X, then it did so because X is in the training set". That does not follow. Conceivably it's possible that X is in the…
Your pure logic is probably right; I do not have the time or interest to dissect it. But you’re missing the context and implication: “doing new stuff” is the major achievement we’re looking for next from LLMs. Seeing something that is “new” and is not in the training set is interesting in a way that something contained in the training set is not. We cannot introspect LLMs meaningfully yet, so the difference between “…
A few examples: Axiom's proof of Fel’s open conjecture on syzygies of numerical semigroups: https://x.com/axiommathai/status/2019449659807219884
Erdos 457: https://www.erdosproblems.com/457
The stronger form of Erdos 650: https://www.erdosproblems.com/650
Re: Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
#86Re: Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
#87Earlier quoted context omitted.
Can you elaborate?
are you familiar with tesla? i'm not super, but am aware of their public things. they introduced fake marketing products called full self driving and autopilot that don't do those things. apparently this person karpathy was in charge of computer vision there. he led the team who is responsible for these systems that occupy our roads which can't navigate due to such outstanding occurrences as sunlight, precipitation,…
Re: Scaling Karpathy's Autoresearch: What Happens When the Agent Gets a GPU Cluster
#88Earlier quoted context omitted.
are you familiar with tesla? i'm not super, but am aware of their public things. they introduced fake marketing products called full self driving and autopilot that don't do those things. apparently this person karpathy was in charge of computer vision there. he led the team who is responsible for these systems that occupy our roads which can't navigate due to such outstanding occurrences as sunlight, precipitation,…
Building a tech and falsely advertising it to be something else that what it is (e.g. self driving instead of driving assistance) can typically done by different people. Lacking specific evidence, it's reckless to accuse this person.