Live data from Hacker News

Epoch confirms GPT5.4 Pro solved a frontier math open problem

epoch.ai

301–310 of 744 posts

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#301
post #247

Earlier quoted context omitted.

Next-token-prediction cannot do calculations. That is fundamental. It can produce outputs that resemble calculations. It can prompt an agent to input some numbers into a separate program that will do calculations for it and then return them as a prompt. Neither of these are calculations.

So you don't think 50T parameter neural networks can encode the logic for adding two n-bit integers for reasonably sized integers? That would be pretty sad.

They do not. The fundamental technology behind LLMs does not allow that to be the case. You are hoping that an LLM can do something that it cannot do.

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#302
post #272

Earlier quoted context omitted.

AI can both explore new things and exploit existing things. Nothing forces it to only rehash old stuff. >without actually making you like really happy or good. What are you basing this off of. I've shared several AI songs with people in real life due to how much I've enjoyed them. I doing see why an AI playlist couldn't be good or make people happy. It just needs to find what you like in music. Again coming back to e…

>What are you basing this off of. Jokes. LLMs are not able to make me laugh all day by generating infinite stream of hilarious original jokes.. Does it work for you?

I've found several posts on moltbook funny. I don't really like regular jokes in general and I don't find human ones particularly funny either. I don't think we are at the point of being able to be reliable funny, but it definitely seems possible from my perspective.

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#303
post #294

Earlier quoted context omitted.

Because economy. Look at marvel movies, do you think the latest one is really new? Or just a rehash of what they found working commercially? Look at all the AI generated blog posts that is flooding the internet.. LLMs might produce something new once in a long while due to blind luck, but if it can generate something that pushes the right buttons (aka not really creative) to majority of population, then that is what…

That's literally all culture: https://www.youtube.com/watch?v=nJPERZDfyWc

The difference is whether an entity that can "feel" is in the loop and how much they have contributed to it even if it is a remix.

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#304
post #272

Earlier quoted context omitted.

>What are you basing this off of. Jokes. LLMs are not able to make me laugh all day by generating infinite stream of hilarious original jokes.. Does it work for you?

I've found several posts on moltbook funny. I don't really like regular jokes in general and I don't find human ones particularly funny either. I don't think we are at the point of being able to be reliable funny, but it definitely seems possible from my perspective.

Care to link some?

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#305

Earlier quoted context omitted.

https://epochai.substack.com/p/first-ai-solution-on-frontier... >The newly-solved problem came from Will Brian, who had placed it in the Moderately Interesting category. It is a conjecture from a paper he wrote with Paul Larson in 2019. They were unable to solve it at the time, or in several attempts since. Brian had this to say.

I actually still don't see the source for them trying several times, but we can take that for granted. Regardless, as I said: 1. It's labeled as "moderately interesting" 2. They said that they expect an expert could solve it in 1-3 months 3. They had already come up with the solution that the AI had but weren't convinced it would have worked So how big was the gap here, do you think?

Yes, a "moderately interesting" Open problem.

I can't think of any chores that would take an expert months to complete. I can't think of any chores that I've completed but was then 'unconvinced could work'. Please sit down and think about what you are saying here. Are we still talking about chores ?

One of the more strange phenomena with machines getting better and the incessant need (seemingly driven by human exceptionalism) to downplay each result, is that you just end up belittling humans in the process.

This is significant. Your analogy is wrong. It's fine to admit it.

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#306

I am kind of amazed at how many commenters respond to this result by confidently asserting that LLMs will never generate 'truly novel' ideas or problem solutions. > AI is a remixer; it remixes all known ideas together. It won't come up with new ideas > it's not because the model is figuring out something new > LLMs will NEVER be able to do that, because it doesn't exist It's not enough to say 'it will never be able t…

> e.g. 167,383 * 426,397 = 71,371,609,051

They may be wrong, but so are you.

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#307

I am thinking there’s a large category of problems that can be solved by resampling existing proofs. It’s the kind of brute force expedition machine can attempt relentlessly where humans would go mad trying. It probably doesn’t really advance the field, but it can turn conjectures into theorems.

I'm of the opinion that everything we've discovered is via combinatorial synthesis. Standing on the shoulders of giants and all that. I'm not sure I've seen any convincing argument that we've discovered anything ex nihilo.

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#308

I am kind of amazed at how many commenters respond to this result by confidently asserting that LLMs will never generate 'truly novel' ideas or problem solutions. > AI is a remixer; it remixes all known ideas together. It won't come up with new ideas > it's not because the model is figuring out something new > LLMs will NEVER be able to do that, because it doesn't exist It's not enough to say 'it will never be able t…

Beliefs are not rooted in facts. Beliefs are a part of you, and people aren't all that happy to say "this LLM is better than me"

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#309
post #304

Earlier quoted context omitted.

I've found several posts on moltbook funny. I don't really like regular jokes in general and I don't find human ones particularly funny either. I don't think we are at the point of being able to be reliable funny, but it definitely seems possible from my perspective.

Care to link some?

I think they would be hard to find due to how many posts exists along with how things aren't as funny the second time around.

Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem

#310
"In this scaffold, several other models were able to solve the problem as well: Opus 4.6 (max), Gemini 3.1 Pro, and GPT-5.4 (xhigh)."

I find that very surprising. This problem seems out of reach 3 months ago but now the 3 frontier models are able to solve it.

Is everybody distilling each others models? Companies sell the same data and RL environment to all big labs? Anybody more involved can share some rumors? :P

I do believe that AI can solve hard problems, but that progress is so distributed in a narrow domain makes me a bit suspicious somehow that there is a hidden factor. Like did some "data worker" solve a problem like that and it's now in the training data?

Post reply on HN