Earlier quoted context omitted.
It is not. You're operating under the assumption that all open math problems are difficult and novel. This particular problem was about improving the lower bound for a function tracking a property of hypergraphs (undirected graphs where edges can contain more than two vertices). Both constructing hypergraphs (sets) and lower bounds are very regular, chore type tasks that are common in maths. In other words, there's p…
> nice that the LLMs solved something for once. That sentence alone needs unpacking IMHO, namely that no LLM suddenly decided that today was the day it would solve a math problem. Instead a couple of people who love mathematics, doing it either for fun or professionally, directly ask a model to solve a very specific task that they estimated was solvable. The LLM itself was fed countless related proofs. They then guid…
Epoch confirms GPT5.4 Pro solved a frontier math open problem
321–330 of 744 posts
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#322Earlier quoted context omitted.
The difference is whether an entity that can "feel" is in the loop and how much they have contributed to it even if it is a remix.
I think there's demonstrably very little difference at all between human and AI outputs, and that's exactly what freaks people out about it. Else they wouldn't be so obsessed with trying to find and define what makes it different. The Thesis of Everything is a Remix is that there is no difference in how any culture is produced. Different models will have a different flavor to their output in the same way as different…
Is there "demonstrably" a lot of difference between Shakespeare and an HN comment?
The point is exactly that there is no such difference. And that it enables slop to be sold as art. And that exactly is the danger. But another point is we had the even before LLMs. And LLMs just make it more explicit and makes it possible at scale.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#323I am kind of amazed at how many commenters respond to this result by confidently asserting that LLMs will never generate 'truly novel' ideas or problem solutions. > AI is a remixer; it remixes all known ideas together. It won't come up with new ideas > it's not because the model is figuring out something new > LLMs will NEVER be able to do that, because it doesn't exist It's not enough to say 'it will never be able t…
I guess when it can't be tripped up by simple things like multiplying numbers, counting to 100 sequentially or counting letters in a string without writing a python program, then I might believe it. Also no matter how many math problems it solves it still gets lost in a codebase
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#324Earlier quoted context omitted.
Care to link some?
I think they would be hard to find due to how many posts exists along with how things aren't as funny the second time around.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#325I am kind of amazed at how many commenters respond to this result by confidently asserting that LLMs will never generate 'truly novel' ideas or problem solutions. > AI is a remixer; it remixes all known ideas together. It won't come up with new ideas > it's not because the model is figuring out something new > LLMs will NEVER be able to do that, because it doesn't exist It's not enough to say 'it will never be able t…
Beliefs are not rooted in facts. Beliefs are a part of you, and people aren't all that happy to say "this LLM is better than me"
Yet, whenever I ask it to do something novel or creative, it falls very short. But humans are ingenious beasts and I'm sure or later they will design an architecture able to be creative - I just doubt it will be Transformer-based, given the results so far.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#326Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#327Earlier quoted context omitted.
I actually still don't see the source for them trying several times, but we can take that for granted. Regardless, as I said: 1. It's labeled as "moderately interesting" 2. They said that they expect an expert could solve it in 1-3 months 3. They had already come up with the solution that the AI had but weren't convinced it would have worked So how big was the gap here, do you think?
Also, the full write up does not say the researchers solved it.
They didn't solve it, that's fair. They did consider the approach already.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#328Earlier quoted context omitted.
I actually still don't see the source for them trying several times, but we can take that for granted. Regardless, as I said: 1. It's labeled as "moderately interesting" 2. They said that they expect an expert could solve it in 1-3 months 3. They had already come up with the solution that the AI had but weren't convinced it would have worked So how big was the gap here, do you think?
Yes, a "moderately interesting" Open problem. I can't think of any chores that would take an expert months to complete. I can't think of any chores that I've completed but was then 'unconvinced could work'. Please sit down and think about what you are saying here. Are we still talking about chores ? One of the more strange phenomena with machines getting better and the incessant need (seemingly driven by human except…
Again, I'm not trying to downplay this, but to frame this accurately. I think an AI being able to build a parser/ compiler is cool too.
> One of the more strange phenomena with machines getting better and the incessant need (seemingly driven by human exceptionalism) to downplay each result, is that you just end up belittling humans in the process.
I don't believe in human exceptionalism at all, don't attribute positions to me.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#329I have long said I am an AI doubter until AI could print out the answers to hard problems or ones requiring tons of innovation. Assuming this is verified to be correct (not by AI) then I just became a believer. I would like to see a few more AI inventions to know for sure, but wow, it really is a new and exciting world. I really hope we use this intelligence resource to make the world better.
We already have a few years of experience with this.
> I really hope we use this intelligence resource to make the world better.
We already have a few years of experience with this.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#330I am kind of amazed at how many commenters respond to this result by confidently asserting that LLMs will never generate 'truly novel' ideas or problem solutions. > AI is a remixer; it remixes all known ideas together. It won't come up with new ideas > it's not because the model is figuring out something new > LLMs will NEVER be able to do that, because it doesn't exist It's not enough to say 'it will never be able t…
I always found this argument very weak. There isn't that much truly new anyway. Creativity is often about mixing old ideas. Computers can do that faster than humans if they have a good framework. Especially with something as simple as math - limited set of formal rules and easy to verify results - I find a belief computers won't beat humans at it to be very naive.