Earlier quoted context omitted.
Next-token-prediction cannot do calculations. That is fundamental. It can produce outputs that resemble calculations. It can prompt an agent to input some numbers into a separate program that will do calculations for it and then return them as a prompt. Neither of these are calculations.
So you don't think 50T parameter neural networks can encode the logic for adding two n-bit integers for reasonably sized integers? That would be pretty sad.
Epoch confirms GPT5.4 Pro solved a frontier math open problem
301–310 of 744 posts
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#302Earlier quoted context omitted.
AI can both explore new things and exploit existing things. Nothing forces it to only rehash old stuff. >without actually making you like really happy or good. What are you basing this off of. I've shared several AI songs with people in real life due to how much I've enjoyed them. I doing see why an AI playlist couldn't be good or make people happy. It just needs to find what you like in music. Again coming back to e…
>What are you basing this off of. Jokes. LLMs are not able to make me laugh all day by generating infinite stream of hilarious original jokes.. Does it work for you?
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#303Earlier quoted context omitted.
Because economy. Look at marvel movies, do you think the latest one is really new? Or just a rehash of what they found working commercially? Look at all the AI generated blog posts that is flooding the internet.. LLMs might produce something new once in a long while due to blind luck, but if it can generate something that pushes the right buttons (aka not really creative) to majority of population, then that is what…
That's literally all culture: https://www.youtube.com/watch?v=nJPERZDfyWc
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#304Earlier quoted context omitted.
>What are you basing this off of. Jokes. LLMs are not able to make me laugh all day by generating infinite stream of hilarious original jokes.. Does it work for you?
I've found several posts on moltbook funny. I don't really like regular jokes in general and I don't find human ones particularly funny either. I don't think we are at the point of being able to be reliable funny, but it definitely seems possible from my perspective.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#305Earlier quoted context omitted.
https://epochai.substack.com/p/first-ai-solution-on-frontier... >The newly-solved problem came from Will Brian, who had placed it in the Moderately Interesting category. It is a conjecture from a paper he wrote with Paul Larson in 2019. They were unable to solve it at the time, or in several attempts since. Brian had this to say.
I actually still don't see the source for them trying several times, but we can take that for granted. Regardless, as I said: 1. It's labeled as "moderately interesting" 2. They said that they expect an expert could solve it in 1-3 months 3. They had already come up with the solution that the AI had but weren't convinced it would have worked So how big was the gap here, do you think?
I can't think of any chores that would take an expert months to complete. I can't think of any chores that I've completed but was then 'unconvinced could work'. Please sit down and think about what you are saying here. Are we still talking about chores ?
One of the more strange phenomena with machines getting better and the incessant need (seemingly driven by human exceptionalism) to downplay each result, is that you just end up belittling humans in the process.
This is significant. Your analogy is wrong. It's fine to admit it.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#306I am kind of amazed at how many commenters respond to this result by confidently asserting that LLMs will never generate 'truly novel' ideas or problem solutions. > AI is a remixer; it remixes all known ideas together. It won't come up with new ideas > it's not because the model is figuring out something new > LLMs will NEVER be able to do that, because it doesn't exist It's not enough to say 'it will never be able t…
They may be wrong, but so are you.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#307I am thinking there’s a large category of problems that can be solved by resampling existing proofs. It’s the kind of brute force expedition machine can attempt relentlessly where humans would go mad trying. It probably doesn’t really advance the field, but it can turn conjectures into theorems.
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#308I am kind of amazed at how many commenters respond to this result by confidently asserting that LLMs will never generate 'truly novel' ideas or problem solutions. > AI is a remixer; it remixes all known ideas together. It won't come up with new ideas > it's not because the model is figuring out something new > LLMs will NEVER be able to do that, because it doesn't exist It's not enough to say 'it will never be able t…
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#309Earlier quoted context omitted.
I've found several posts on moltbook funny. I don't really like regular jokes in general and I don't find human ones particularly funny either. I don't think we are at the point of being able to be reliable funny, but it definitely seems possible from my perspective.
Care to link some?
Re: Epoch confirms GPT5.4 Pro solved a frontier math open problem
#310I find that very surprising. This problem seems out of reach 3 months ago but now the 3 frontier models are able to solve it.
Is everybody distilling each others models? Companies sell the same data and RL environment to all big labs? Anybody more involved can share some rumors? :P
I do believe that AI can solve hard problems, but that progress is so distributed in a narrow domain makes me a bit suspicious somehow that there is a hidden factor. Like did some "data worker" solve a problem like that and it's now in the training data?