Earlier quoted context omitted.
Impressive prediction, especially pre-ChatGPT. Compare to Gary Marcus 3 months ago: https://garymarcus.substack.com/p/reports-of-llms-mastering-... We may certainly hope Eliezer's other predictions don't prove so well-calibrated.
Gary Marcus is so systematically and overconfidently wrong that I wonder why we keep talking about this clown.
OpenAI claims gold-medal performance at IMO 2025
341–350 of 737 posts
Re: OpenAI claims gold-medal performance at IMO 2025
#342Not even bronze. https://news.ycombinator.com/item?id=44615695
Re: OpenAI claims gold-medal performance at IMO 2025
#343Earlier quoted context omitted.
Gary Marcus is so systematically and overconfidently wrong that I wonder why we keep talking about this clown.
People like him and Zitron do serve a useful purpose in balancing the hype from the other side, which, while justified to a great extent, is often a bit too overwhelming.
Re: OpenAI claims gold-medal performance at IMO 2025
#344Re: OpenAI claims gold-medal performance at IMO 2025
#345Noam Brown: > this isn’t an IMO-specific model. It’s a reasoning LLM that incorporates new experimental general-purpose techniques. > it’s also more efficient [than o1 or o3] with its thinking. And there’s a lot of room to push the test-time compute and efficiency further. > As fast as recent AI progress has been, I fully expect the trend to continue. Importantly, I think we’re close to AI substantially contributing…
The new "Full Self-Driving next year"?
Re: OpenAI claims gold-medal performance at IMO 2025
#346Earlier quoted context omitted.
>> Why is that less exciting? A machine competing in an unconstrained natural language difficult math contest and coming out on top by any means is breath taking science fiction a few years ago - now it’s not exciting? Half the internet is convinced that LLMs are a big data cheating machine and if they're right then, yes, boldly cheating where nobody has cheated before is not that exciting.
I don't get it, how do you "big data cheat" an AI into solving previously unencountered problems? Wouldn't that just be engineering?
Re: OpenAI claims gold-medal performance at IMO 2025
#347Re: OpenAI claims gold-medal performance at IMO 2025
#348Has anyone independently reviewed these solutions? My proving skills are extremely rusty so I can’t look at these and validate them. They certainly are not traditional proofs though.
It reads like someone who found the correct answer but seemingly had no understanding of what they did and just handed in the draft paper.
Which seems odd, shouldn't an LLM be better at prose?
Re: OpenAI claims gold-medal performance at IMO 2025
#349The cynicism/denial on HN about AI is exhausting. Half the comments are some weird form of explaining away the ever increasing performance of these models I've been reading this website for probably 15 years, its never been this bad. many threads are completely unreadable, all the actual educated takes are on X, its almost like there was a talent drain
Re: OpenAI claims gold-medal performance at IMO 2025
#350Earlier quoted context omitted.
[flagged]
I feel like I've noticed you you making the same comment 12 places in this thread -- incorrectly misrepresenting the difficulty of this tournament and ultimately it comes across as a bitter ex. Here's an example problem 5: Let a1,a2,…,an be distinct positive integers and let M=max1≤i Find the maximum number of pairs (i,j) with 1≤i<j≤n for which (ai +aj )(aj −ai )=M.