Our human-ability to abstract things is underrated.
Arc-AGI-2 and ARC Prize 2025
21–30 of 103 posts
Re: Arc-AGI-2 and ARC Prize 2025
#22This is self-referential, the benchmark pinpointed the time when AI went from memorization to problem solving, because the benchmark requires problem solving to complete. How do we know it requires problem solving skills? Because memorization-only LLMs can't do it but humans can.
I think ARC are producing some great benchmarks, and I think they probably are pushing forward the state of the art, however I don't think they identified anything particular with o3, at least they don't seem to have proven a step change.
Re: Arc-AGI-2 and ARC Prize 2025
#23Earlier quoted context omitted.
ARC 3 is still spatially 2D, but it adds a time dimension, and it's interactive.
I think a lot of people got discouraged, seeing how openai solved arc agi 1 by what seems like brute forcing and throwing money at it. Do you believe arc was solved in the "spirit" of the challenge? Also all the open sourced solutions seem super specific to solving arc. Is this really leading us to human level AI at open ended tasks?
Re: Arc-AGI-2 and ARC Prize 2025
#24Have you had any neurologists utilize your dataset? My own reaction after solving a few of the puzzles was "Why is this so intuitive for me, but not for an LLM?". Our human-ability to abstract things is underrated.
Re: Arc-AGI-2 and ARC Prize 2025
#25Re: Arc-AGI-2 and ARC Prize 2025
#26Oh boy! Some of these tasks are not hard, but require full attention and a lot of counting just to get things right! ARC3 will go 3D perhaps? JK Congrats on launch, lets see how long it'll take to get saturated
ARC 3 is still spatially 2D, but it adds a time dimension, and it's interactive.
Re: Arc-AGI-2 and ARC Prize 2025
#27> and was the only benchmark to pinpoint the exact moment in late 2024 when AI moved beyond pure memorization This is self-referential, the benchmark pinpointed the time when AI went from memorization to problem solving, because the benchmark requires problem solving to complete. How do we know it requires problem solving skills? Because memorization-only LLMs can't do it but humans can. I think ARC are producing som…
ARC 1 was released long before in-context learning was identified in LLMs (and designed before Transformer-based LLMs existed), so the fact that LLMs can't do ARC was never a design consideration. It just turned out this way, which confirmed our initial assumption.
Re: Arc-AGI-2 and ARC Prize 2025
#28Oh boy! Some of these tasks are not hard, but require full attention and a lot of counting just to get things right! ARC3 will go 3D perhaps? JK Congrats on launch, lets see how long it'll take to get saturated
Re: Arc-AGI-2 and ARC Prize 2025
#29Oh boy! Some of these tasks are not hard, but require full attention and a lot of counting just to get things right! ARC3 will go 3D perhaps? JK Congrats on launch, lets see how long it'll take to get saturated
ARC 3 is still spatially 2D, but it adds a time dimension, and it's interactive.