Live data from Hacker News

Poetiq shatters ARC-AGI 2 benchmark at half the cost

poetiq.ai

1–4 of 4 posts

Re: Poetiq shatters ARC-AGI 2 benchmark at half the cost

#3

Considering the release of GPT-5.2, this article is worth discussing together, as it managed to achieve the same high score as GPT-5.2 using Gemini 3 Pro

Am I crazy to think these models have actually surpassed human performance on ARC 2? https://www.lesswrong.com/posts/DX3EmhmwZjTYp9PBf/ai-perform...

Re: Poetiq shatters ARC-AGI 2 benchmark at half the cost

#4
post #3

Considering the release of GPT-5.2, this article is worth discussing together, as it managed to achieve the same high score as GPT-5.2 using Gemini 3 Pro

Am I crazy to think these models have actually surpassed human performance on ARC 2? https://www.lesswrong.com/posts/DX3EmhmwZjTYp9PBf/ai-perform...

This is not surprising, rather, it's the 100% figure that makes me skeptical. In fact, the intelligence level of ordinary people isn't that high, and AI can indeed surpass it. Otherwise, why would we use it?