Live data from Hacker News

François Chollet: The Arc Prize and How We Get to AGI [video]

youtube.com

71–80 of 230 posts

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#71

The Arc prize/benchmark is a terrible judge of whether we got to AGI. If we assume that humans have "general intelligence", we would assume all humans could ace Arc... but they can't. Try asking your average person, i.e. supermarket workers, gas station attendants etc to do the Arc puzzles, they will do poorly, especially on the newer ones, but AI has to do perfectly to prove they have general intelligence? (not tryi…

This is what is called "spikey" intelligence, where a model might be able to crack phd physics problems and solve byzantine pattern matching games at the 90th percentile, but also can't figure out how to look up a company and copy their address on the "customer" line of an invoice.

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#72

Let's not. Seriously. I absolutely love François and have used his work extensively. But looking around me at the social impact of AI I am really not convinced that this is what the world needs right now and that if we can stave off the turning point for another decade or two that humanity will likely benefit from that. The last thing we need is to inject yet another instability into a planet that is already fighting…

It doesn't matter what should or should not happen. Technology will continue to race forward at breakneck speed while everyone involved pats each other on the back for making a bunch of money before the consequences hit

technology doesn't just advance itself

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#73
post #72

Earlier quoted context omitted.

It doesn't matter what should or should not happen. Technology will continue to race forward at breakneck speed while everyone involved pats each other on the back for making a bunch of money before the consequences hit

technology doesn't just advance itself

This is true. We have a choice...in principle.

But in practice, it's like stopping an arms race.

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#74
post #4

I feel like I'm the only one who isn't convinced getting a high score on the ARC eval test means we have AGI. It's mostly about pattern matching (and some of it ambiguous even for humans what the actual true response aught to be). It's like how in humans there's lots of different 'types' of intelligence, and just overfitting on IQ tests doesn't in my mind convince me a person is actually that smart.

I agree with you but I'll go a step further - these benchmarks are a good example of how far we are from AGI. A good base test would be to give a manager a mixed team of remote workers, half being human and half being AI, and seeing if the manager or any of the coworkers would be able to tell the difference. We wouldn't be able to say that AI that passed that test would necessarily be AGI, since we would have to test…

The AGI test I think makes sense is to put it in a robot body and let it navigate the world. Can I take the robot to my back yard and have it weed my vegetable garden? Can I show it how to fold my laundry? Can I take it to the grocery store and tell it "go pick up 4 yellow bananas and two avocados that will be ready to eat in the next day or two, and then meet me in dairy"? Can I ask it to dice an onion for me during meal prep?

These are all things my kids would do when they were pretty young.

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#75
post #72

Earlier quoted context omitted.

It doesn't matter what should or should not happen. Technology will continue to race forward at breakneck speed while everyone involved pats each other on the back for making a bunch of money before the consequences hit

technology doesn't just advance itself

No, but one thing is certain, in large human systems you can only redirect greed, you can't stop it.

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#76

Earlier quoted context omitted.

I'd say we're not far off. Looking at the human side, it takes a while to actually learn something. If you've recently read something it remains in your "context window". You need to dream about it, to think about, to revisit and repeat until you actually learn it and "update your internal model". We need a mechanism for continuous weight updating. Goal-generation is pretty much covered by your body constantly drip-f…

> I'd say we're not far off. How are we not far off? How can LLMs generate goals and based on what?

You just train it on the goal. Then it has that goal.

Alternately, you can train it on following a goal and then you have a system where you can specify a goal.

At sufficient scale, a model will already contain goal-following algorithms because those help predict the next token when the model is basetrained on goal-following entities, ie. humans. Goal-driven RL then brings those algorithms to prominence.

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#77
post #4

I feel like I'm the only one who isn't convinced getting a high score on the ARC eval test means we have AGI. It's mostly about pattern matching (and some of it ambiguous even for humans what the actual true response aught to be). It's like how in humans there's lots of different 'types' of intelligence, and just overfitting on IQ tests doesn't in my mind convince me a person is actually that smart.

I agree with you but I'll go a step further - these benchmarks are a good example of how far we are from AGI. A good base test would be to give a manager a mixed team of remote workers, half being human and half being AI, and seeing if the manager or any of the coworkers would be able to tell the difference. We wouldn't be able to say that AI that passed that test would necessarily be AGI, since we would have to test…

The problem with "spot the difference" tests, imho, is that I would expect an AGI to be easily spotted. There's going to be a speed of calculation difference, at the very least. If nothing else, typing speed would be completely different unless the AGI is supposed to be deceptive. Who knows what it's personality would be like. I'd say it's a simple enough test just to see if an AGI could be hired as, for example, an entry level software developer and keep it's job based on the same criteria base-level humans have to meet.

I agree that current AI is nowhere near that level yet. If AI isn't even trying to extract meaning from the words it smiths or the pictures it diffuses then it's nothing more than a cute (albeit useful) parlor trick.

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#78
ARC-AGI 3 remindes me of PuzzleScript games: https://www.puzzlescript.net/Gallery/index.html

There are dozens of ready-made, well-designed, and very creative games there. All are tile-based and solved with only arrow keys and a single action button. Maybe someone should make a PuzzleScript AGI benchmark?

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#79
post #47

The Arc prize/benchmark is a terrible judge of whether we got to AGI. If we assume that humans have "general intelligence", we would assume all humans could ace Arc... but they can't. Try asking your average person, i.e. supermarket workers, gas station attendants etc to do the Arc puzzles, they will do poorly, especially on the newer ones, but AI has to do perfectly to prove they have general intelligence? (not tryi…

It is not a judge of whether we got to AGI. And literally no one except straw-manning critics are trying to claim it is . The point is, an AGI should easily be able to pass it. But it can obviously be passed without getting to AGI (as . It's a necessary but not sufficient criteria. If something can't pass a test as simple as AGI (which no AI currently can ) then it's definitely not AGI. Anyone claiming AGI should be…

So do you agree that a human that CANNOT solve ARC doesn't have general intelligence?

If we think humans have "GI" then I think we have AIs right now with "GI" too. Just like humans do, AIs spike in various directions. They are amazing at some things and weak at visual/IQ test type problems like ARC.

Re: François Chollet: The Arc Prize and How We Get to AGI [video]

#80

Earlier quoted context omitted.

I agree with you but I'll go a step further - these benchmarks are a good example of how far we are from AGI. A good base test would be to give a manager a mixed team of remote workers, half being human and half being AI, and seeing if the manager or any of the coworkers would be able to tell the difference. We wouldn't be able to say that AI that passed that test would necessarily be AGI, since we would have to test…

The AGI test I think makes sense is to put it in a robot body and let it navigate the world. Can I take the robot to my back yard and have it weed my vegetable garden? Can I show it how to fold my laundry? Can I take it to the grocery store and tell it "go pick up 4 yellow bananas and two avocados that will be ready to eat in the next day or two, and then meet me in dairy"? Can I ask it to dice an onion for me during…

I agree, I think of that as the next level beyond the digital assistant test - a physical assistant test. Once there are sufficiently capable robots, hook one up to the AI. Tell it to mow your lawn, drive your car to the mechanic and have the mechanic to get checked, box up an item, take it to the post office, and have it shiped, pick up your dry cleaning, buy ingredients from a grocery store, cook dinner, etc. Basic tasks an low-skilled worker would do as someone's assistant.
Post reply on HN