Live data from Hacker News

Ten advances in mathematics and theoretical computer science

openai.com

111–120 of 1000 posts

Re: Ten advances in mathematics and theoretical computer science

#111

In a way the most remarkable thing about this is that it isn't even at the top of the HN homepage. Even if this is a step up from what we've seen before, we're no longer astonished by the idea that AI can make significant advances in mathematics and computer science.

What about AI research itself? Is OpenAI close to automating its human staff out of a job?

Re: Ten advances in mathematics and theoretical computer science

#112
post #6

My main gripe here is the lack of transparency around the total experiment and construction. I doubt that they simply pointed their model at these ten specific problems alone and gave the model one shot; therefore the $2000 number could be completely misleading, similar to P-value hacking by not disclosing the total experimental setup. I want to know: 1. How many total problems were given to the model, and what perce…

[flagged]

Re: Ten advances in mathematics and theoretical computer science

#113
post #111

In a way the most remarkable thing about this is that it isn't even at the top of the HN homepage. Even if this is a step up from what we've seen before, we're no longer astonished by the idea that AI can make significant advances in mathematics and computer science.

What about AI research itself? Is OpenAI close to automating its human staff out of a job?

Yes, but they wouldn't publish that bit lest other companies steal the ideas.

Re: Ten advances in mathematics and theoretical computer science

#114
post #5

I wonder what the total cost of this research was, including the salary for their mathematicians and engineers.

Why would you factor in salary unless they had to baby it through. You would only count the hours for setting up the harness and prompt and checking the result.

Training the model is going to be amortized over other uses.

Re: Ten advances in mathematics and theoretical computer science

#115
Now that we've seen AI produce a fair number of proofs (and disproofs), I'm curious when we'll start seeing it build genuinely novel theory. Does anyone have predictions on when and how we'll get there and will it take new architectures/ training paradigms, or is the current approach enough?

Re: Ten advances in mathematics and theoretical computer science

#117

Earlier quoted context omitted.

I don't think you want to bring cost into this argument. Even if the cost was $1 mil for these 10 problems, that's maybe 10-20 math researchers for a year. Do you really think that if you paid that to humans, they will deliver the same results?

Grad students on zero pay solve problems like this everyday. What exactly is your point here?

Even if they can solve problems like this every day, you still have a very limited number of grad students who can solve them. With model capabilities like this, you can have the equivalent of millions of grad students who can solve problems like this.

Re: Ten advances in mathematics and theoretical computer science

#118
one of the early premises of how ai takeoff would go was that a system that could solve open problems in advanced mathematics would also discover novel advances in math and computer science that directly unlock drastically better software performance. we are seeing frontier level math breakthroughs (ie performance that would put it in the top 100 or 1000 mathematicians in the world if it were a human, meaning top .00001% or 800/8B). we are also seeing incredible advances in software performance. open ai announced like 15% improvement by fixing gpu kernel issues. these are clearly linked in the sense of scaling laws and generalization of intelligence: a huge model gets capabilities in both math and software engineering that isn't possible at smaller scales.

but it seems less likely to me than before that the types of math/science discoveries will explicitly unlock better software performance. in some sense this fits our intuitions. when top tech companies use math PhD type employees, they have them stop doing pure math research and instead focus on software engineering. these people are often very good at software engineering but not due to recent discoveries in academic mathematics, it's due to their general intelligence. to me, this is evidence that the models are getting better but does not make me think we are on the cusp of a foom style fast takeoff enabled by revolutions in frontier math (i also posted this on twitter @mlipman13)

Re: Ten advances in mathematics and theoretical computer science

#119
post #6

My main gripe here is the lack of transparency around the total experiment and construction. I doubt that they simply pointed their model at these ten specific problems alone and gave the model one shot; therefore the $2000 number could be completely misleading, similar to P-value hacking by not disclosing the total experimental setup. I want to know: 1. How many total problems were given to the model, and what perce…

I don't think you want to bring cost into this argument. Even if the cost was $1 mil for these 10 problems, that's maybe 10-20 math researchers for a year. Do you really think that if you paid that to humans, they will deliver the same results?

It is comical at this point. Some people just can not stand the thought of AI actually delivering and are trying to find whatever ways to discredit it.

Re: Ten advances in mathematics and theoretical computer science

#120
post #81

Earlier quoted context omitted.

Dunno about the parent commenter, but I personally interpret the concept as having a hidden representation of self that is continually tended to, and influences future choices. This implies statefulness, which models are intentionally not at inference time (*). (*) Even if we hack around this and just do the usual trick of simply laundering statefulness to a higher level, in this case the context window being fed in,…

Before reading, know that I am uncertain in either direction. > a hidden representation of self that is continually tended to This sounds like a personality? They act like they have one of those. It may be an illusion, and even if it isn't an illusion it is unlikely to be anything like the source (us), but they act like it. > I further fail to identify how it could be hidden or maintained, considering I control like…

> This sounds like a personality?

Not quite what I meant, but it's also not entirely unrelated I guess? Personality to me is like a natural bias. It does also shift over time, and is also an internal bit of state. I guess in some respects it can also be self-referential, like personal convictions.

> Perhaps they were losing their self-awareness at the time?

I do think it is entirely possible for people's self-awareness to shift, yes. Or more precisely, I do model things that way.

> Do you mean like these, or something else?

They're adjacent, but I more meant something like these:

https://arxiv.org/abs/2410.03768

https://arxiv.org/abs/2310.18512

https://arxiv.org/abs/2605.26537

So basically, steganography. The difference is that these papers investigate from the perspective of separate LLM instances covertly exchanging information between each other. This is in contrast with the scenario I'm laying out, where an LLM's past state is exchanging information with its future state, continuously representing and modulating a concealed internal state of some sort. And then that state just so happening to be some sort of self-referential meta state.

And the best inkling I have towards this is basically: https://www.youtube.com/shorts/WP5_XJY_P0Q

But then I don't think there's enough covert channel bandwidth in the agent replies for anything interesting like this.

Post reply on HN