AI models and in particular LLMs are not capable of logic reasoning. See for example this paper: https://arxiv.org/abs/2506.06941 Ergo, they can't prove any theorem whatsoever. How do people at OpenAI expect that we believe in claims like that? This is yet another before-the-IPO stunt in my opinion.. Personally, I won't believe any of these claims until the community of mathematicians says otherwise.
How An AI math breakthrough ignited a controversy
161–170 of 243 posts
Re: How An AI math breakthrough ignited a controversy
#162Regardless of what you think of the priority dispute issue discussed on sibling threads, I’m highly skeptical of the closing quote that this Navier Stokes result means that the same approach of casually spending a few million on agentic computation is going to solve end to end materials design or drug development. Those problems can’t be formally verified with an automated theorem prover. We have a lot of physics bas…
Re: How An AI math breakthrough ignited a controversy
#163Regardless of what you think of the priority dispute issue discussed on sibling threads, I’m highly skeptical of the closing quote that this Navier Stokes result means that the same approach of casually spending a few million on agentic computation is going to solve end to end materials design or drug development. Those problems can’t be formally verified with an automated theorem prover. We have a lot of physics bas…
Yeah, I think you can't just throw money randomly at problems and expect results unless you know a line of attack that can get you all the way. OpenAI chose the line of attack only after it became known to them via rumors. They "front-ran" the researchers.
Re: How An AI math breakthrough ignited a controversy
#164Earlier quoted context omitted.
> but only if Alpöge’s name was removed This is blatant scientific misconduct.
IIUC, the accusation was not to try to remove Alpöge from the paper he wrote with Buckmaster solving the "easier" conjeture, but to exclude Alpöge in the followup paper where Buckmaster review the OpenAI solution of the "full" conjeture. For comparison, if you offer me to collaborate in a paper about Algebra I may agree to go alone, but if the paper is about Quantum Chemistry I have to piggyback a few coworkers becau…
Re: How An AI math breakthrough ignited a controversy
#165Earlier quoted context omitted.
Yeah, I think you can't just throw money randomly at problems and expect results unless you know a line of attack that can get you all the way. OpenAI chose the line of attack only after it became known to them via rumors. They "front-ran" the researchers.
but the researchers were also largely relying on AI
Re: How An AI math breakthrough ignited a controversy
#166Re: How An AI math breakthrough ignited a controversy
#167Earlier quoted context omitted.
but the researchers were also largely relying on AI
In the same way you rely on a keyboard or touchscreen to type this comment. It doesn't mean the tool is the brain behind the work.
Re: How An AI math breakthrough ignited a controversy
#168Re: How An AI math breakthrough ignited a controversy
#169Earlier quoted context omitted.
Not really. There's a non-pedantic, charitable interpretation widely available. It's below for reference. > Navier-Stokes is one of six [open] “Millennium Problems” on a list compiled by the Clay Mathematics Institute in 2000.
If solved problems don’t count anymore, the sentence also doesn’t make sense because now navier stokes apparently isn’t open anymore too, right?
Re: How An AI math breakthrough ignited a controversy
#170Earlier quoted context omitted.
Yeah, I think you can't just throw money randomly at problems and expect results unless you know a line of attack that can get you all the way. OpenAI chose the line of attack only after it became known to them via rumors. They "front-ran" the researchers.
but the researchers were also largely relying on AI
If I write a book and pass it through a spelling and polish checker, I still wrote the book and its core IP. I didn’t “rely on” the tool to create the IP.