Earlier quoted context omitted.
Based on the past history with frontier-math & AIME 2025 [1],[2] I would not trust announcements which cant be independently verified. I am excited to try it out though. Also, the performance of LLMs on imo 2025 was not even bronze [3]. Finally, this article shows that LLMs were just mostly bluffing [4] on usamo 2025. [1] https://www.reddit.com/r/slatestarcodex/comments/1i53ih7/fro ... [2] https://x.com/DimitrisPapai…
The solutions were publicly posted to GitHub: https://github.com/aw31/openai-imo-2025-proofs/tree/main
My skepticism stems from the past frontier math announcement which turned out to be a bluff.