Is there something special about these questions that makes them resistant to memorization? Or is it more just the fact that there are 100 secret tasks?
ARC Prize – a $1M+ competition towards open AGI progress
11–20 of 351 posts
Re: ARC Prize – a $1M+ competition towards open AGI progress
#12I'd also urge you to use a different platform for communicating with the public because x.com links are now inaccessible without creating an account.
Re: ARC Prize – a $1M+ competition towards open AGI progress
#13Re: ARC Prize – a $1M+ competition towards open AGI progress
#14In it they question the ease of Chollet's tests: "One limitation on ARC’s usefulness for AI research is that it might be too challenging. Many of the tasks in Chollet’s corpus are difficult even for humans, and the corpus as a whole might be sufficiently difficult for machines that it does not reveal real progress on machine acquisition of core knowledge."
ConceptARC is designed to be easier, but then also has to filter ~15% of its own test takers for "[failing] at solving two or more minimal tasks... or they provided empty or nonsensical explanations for their solutions"
After this filtering, ConceptARC finds another 10-15% failure rate amongst humans on the main corpus questions, so they're seeing maybe 25-30% unable to solve these simpler questions meant to test for "AGI".
ConceptARC's main results show CG4 scoring well below the filtered humans, which would agree with a [Mensa] test result that its IQ=85.
Chollet and Mitchell could instead stratify their human groups to estimate IQ then compare with the Mensa measures and see if e.g. Claude3@IQ=100 compares with their ARC scores for their average human
[ConceptArc]https://arxiv.org/pdf/2305.07141 [Mensa]https://www.maximumtruth.org/p/ais-ranked-by-iq-ai-passes-10...
Re: ARC Prize – a $1M+ competition towards open AGI progress
#15I really like the idea of ARC. But to me the problems seem like they require a lot of spatial world knowledge, more than they require abstract reasoning. Shapes overlapping each other, containing each other, slicing up and reassembling pieces, denoising regular geometric shapes, you can call them "core knowledge" but to me it seems like they are more like "things that are intuitive to human visual processing". Would…
This is the wrong way to think about it IMO. Spatial relationships are just another type of logical relationship and we should expect AGI to be able to analyze relationships and generate algorithms on the fly to solve problems.
Just because humans can be biased in various ways doesn’t mean these biases are inherent to all intelligences.
Re: ARC Prize – a $1M+ competition towards open AGI progress
#16Re: ARC Prize – a $1M+ competition towards open AGI progress
#17I really like the idea of ARC. But to me the problems seem like they require a lot of spatial world knowledge, more than they require abstract reasoning. Shapes overlapping each other, containing each other, slicing up and reassembling pieces, denoising regular geometric shapes, you can call them "core knowledge" but to me it seems like they are more like "things that are intuitive to human visual processing". Would…
“Would an intelligent but blind human be able to solve these problems?” This is the wrong way to think about it IMO. Spatial relationships are just another type of logical relationship and we should expect AGI to be able to analyze relationships and generate algorithms on the fly to solve problems. Just because humans can be biased in various ways doesn’t mean these biases are inherent to all intelligences.
It’s similar to how chess problems are technically reasoning problems but they are not representative of general reasoning.
Re: ARC Prize – a $1M+ competition towards open AGI progress
#18Re: ARC Prize – a $1M+ competition towards open AGI progress
#19I really like the idea of ARC. But to me the problems seem like they require a lot of spatial world knowledge, more than they require abstract reasoning. Shapes overlapping each other, containing each other, slicing up and reassembling pieces, denoising regular geometric shapes, you can call them "core knowledge" but to me it seems like they are more like "things that are intuitive to human visual processing". Would…
“Would an intelligent but blind human be able to solve these problems?” This is the wrong way to think about it IMO. Spatial relationships are just another type of logical relationship and we should expect AGI to be able to analyze relationships and generate algorithms on the fly to solve problems. Just because humans can be biased in various ways doesn’t mean these biases are inherent to all intelligences.
Not really. By that reasoning, 5-dimensional spatial reasoning is "just another type of logical relationship" and yet humans mostly can't do that at all.
It's clear that we have incredibly specialized capabilities for dealing with two- and three-dimensional spatiality that don't have much of anything to do with general logical intelligence at all.
Re: ARC Prize – a $1M+ competition towards open AGI progress
#20This is treating “intelligence” like some abstract, platonic thing divorced from reality. Whatever else solving these puzzles is indicative of, it’s not intelligence.