Live data from Hacker News

OpenAI O3 breakthrough high score on ARC-AGI-PUB

arcprize.org

911–920 of 1001 posts

Re: OpenAI O3 breakthrough high score on ARC-AGI-PUB

#911

Earlier quoted context omitted.

Not sure how much that matters - I'm not an AI expert, but I did some intro courses where we had to train a classifier to recognize digits. How it worked basically was that we fed each pixel of the 2d grid of the image into an input of the network, essentially flattening it in a similar fashion. It worked just fine, and that was a tiny network.

The classifier was likely a convolutional network, so the assumption of the image being a 2D grid was baked into the architecture itself - it didn't have to be represented via the shape of the input for the network to use it.

I don't think so - convolutional neural networks also operate over 1D flat vectors - the spatial relationship of pixels is only learned from the training data.

Re: OpenAI O3 breakthrough high score on ARC-AGI-PUB

#912

It sucks that I would love to be excited about this... but I mostly feel anxiety and sadness.

We’re enabling a huge swath of humanity being put out of work so a handful of billionaires can become trillionaires.

This is the same boring alarmist argument we’ve heard since the Industrial Revolution. Humans have always turned extra output provided by technological advancement to increase overall productivity.

Re: OpenAI O3 breakthrough high score on ARC-AGI-PUB

#914

Incredibly impressive. Still can't really shake the feeling that this is o3 gaming the system more than it is actually being able to reason. If the reasoning capabilities are there, there should be no reason why it achieves 90% on one version and 30% on the next. If a human maintains the same performance across the two versions, an AI with reason should too.

Yes, if a system has actually achieved AGI, it is likely to not reveal that information

Re: OpenAI O3 breakthrough high score on ARC-AGI-PUB

#915

Incredibly impressive. Still can't really shake the feeling that this is o3 gaming the system more than it is actually being able to reason. If the reasoning capabilities are there, there should be no reason why it achieves 90% on one version and 30% on the next. If a human maintains the same performance across the two versions, an AI with reason should too.

Yes, if a system has actually achieved AGI, it is likely to not reveal that information

AGI is a spectrum, not a binary quality.

Re: OpenAI O3 breakthrough high score on ARC-AGI-PUB

#916

Incredibly impressive. Still can't really shake the feeling that this is o3 gaming the system more than it is actually being able to reason. If the reasoning capabilities are there, there should be no reason why it achieves 90% on one version and 30% on the next. If a human maintains the same performance across the two versions, an AI with reason should too.

But does it matter if it "really, really" reasons in the human sense, if it's able to prove some famous math theorem or come up with a novel result in theoretical physics?

While beyond current motels, that would be the final test of AGI capability.

Re: OpenAI O3 breakthrough high score on ARC-AGI-PUB

#917
I just noticed this bit:

>> Second, you need the ability to recombine these functions into a brand new program when facing a new task – a program that models the task at hand. Program synthesis.

"Program synthesis" is here used in an entirely idiosyncratic manner, to mean "combining programs". Everyone else in CS and AI for the last many decades has used "Program Synthesis" to mean "generating a program that satisfies a specification".

Note that "synthesis" can legitimately be used to mean "combining". In Greek it translates literally to "putting [things] together": "Syn" (plus) "thesis" (place). But while generating programs by combining parts of other programs is an old-fashioned way to do Program Synthesis, in the standard sense, the end result is always desired to be a program. The LLMs used in the article to do what F. Chollet calls "Porgram Synthesis" generate no code.

Re: OpenAI O3 breakthrough high score on ARC-AGI-PUB

#918

It sucks that I would love to be excited about this... but I mostly feel anxiety and sadness.

We’re enabling a huge swath of humanity being put out of work so a handful of billionaires can become trillionaires.

It would happen in China regardless what is done here. Removing billionaires does not fix this. The ship has sailed.

Re: OpenAI O3 breakthrough high score on ARC-AGI-PUB

#919

Incredibly impressive. Still can't really shake the feeling that this is o3 gaming the system more than it is actually being able to reason. If the reasoning capabilities are there, there should be no reason why it achieves 90% on one version and 30% on the next. If a human maintains the same performance across the two versions, an AI with reason should too.

But does it matter if it "really, really" reasons in the human sense, if it's able to prove some famous math theorem or come up with a novel result in theoretical physics? While beyond current motels, that would be the final test of AGI capability.

If it's gaming the system, then it's much less likely to reliably come up with novel proofs or useful new theoretical ideas.
Post reply on HN