Earlier quoted context omitted.
It's in the link above, but you can look at #1051 or #851 on the erdosproblems website.
The erdosproblems website shows 851 was proved in 1934. https://www.erdosproblems.com/851 I guess 1051 qualifies - from the paper: "Semi-autonomous mathematical discovery with gemini" https://arxiv.org/pdf/2601.22401 "We tentatively believe Aletheia’s solution to Erdős-1051 represents an early example of an AI system autonomously resolving a slightly non-trivial open Erdős problem of somewhat broader (mild) mathemati…
GPT-5.2 derives a new result in theoretical physics
221–230 of 430 posts
Re: GPT-5.2 derives a new result in theoretical physics
#222Re: GPT-5.2 derives a new result in theoretical physics
#223I'll read the article in a second, but let me guess ahead of time: Induction. Okay read it: Yep Induction. It already had the answer. Don't get me wrong, I love Induction... but we aren't having any revolutions in understanding with Induction.
Re: GPT-5.2 derives a new result in theoretical physics
#224AI can be an amazing productivity multiplier for people who know what they're doing. This result reminded me of the C compiler case that Anthropic posted recently. Sure, agents wrote the code for hours but there was a human there giving them directions, scoping the problem, finding the test suites needed for the agentic loops to actually work etc etc. In general making sure the output actually works and that it's a s…
>AI can be an amazing productivity multiplier for people who know what they're doing. >[...] >The "AI replaces humans in X" narrative is primarily a tool for driving attention and funding. You're sort of acting like it's all or nothing. What about the the humans that used to be that "force multiplier" on a team with the person guiding the research? If a piece of software required a team of ten to people, and instead…
That, of course, assumes that there are 9 other projects that are both known (or knowable) and worth doing. And in the case of Uber/Lyft drivers, there's a skillset mismatch between the "deprecated" jobs and their replacements.
Re: GPT-5.2 derives a new result in theoretical physics
#225Earlier quoted context omitted.
[flagged]
[flagged]
Re: GPT-5.2 derives a new result in theoretical physics
#226Earlier quoted context omitted.
It's a stupid point then. Are you able to work with a world leading physicist to any significant degree? No
It's like saying: calculator drives new result in theoretical physics (In the hands of leading experts.)
The humans put in significant effort and couldn’t do it. They didn’t then crank it out with some search/match algorithm.
They tried a new technology, modeled (literally) on us as reasoners, that is only just being able to reason at their level and it did what they couldn’t.
The fact that the experts were a critical context for the model, doesn’t make the models performance any less significant. Collaborators always provide important context for each other.
Re: GPT-5.2 derives a new result in theoretical physics
#227The headline may make it seem like AI just discovered some new result in physics all on its own, but reading the post, humans started off trying to solve some problem, it got complex, GPT simplified it and found a solution with the simpler representation. It took 12 hours for GPT pro to do this. In my experience LLM’s can make new things when they are some linear combination of existing things but I haven’t been to g…
When chess engines were first developed, they were strictly worse than the best humans. After many years of development, they became helpful to even the best humans even though they were still beatable (1985–1997). Eventually they caught up and surpassed humans but the combination of human and computer was better than either alone (~1997–2007). Since then, humans have been more or less obsoleted in the game of chess.…
Re: GPT-5.2 derives a new result in theoretical physics
#228Earlier quoted context omitted.
It's easy to fall into a negative mindset when there are legions of pointy haired bosses and bandwagoning CEOs who (wrongly) point at breakthroughs like this as justification for AI mandates or layoffs.
Yes, all of these stories, and frequent model releases are just intended to psyop "decision makers" into validating their longstanding belief that the labour shouldn't be as big of a line item in a companies expenses, and perhaps can be removed altogether.. They can finally go back to the good old days of having slaves (in the form of "agentic" bots), they yearn to own slaves again. CEOs/decision makers would rather…
> Competition will be dynamic because people have agency. The country that is ahead at any given moment will commit mistakes driven by overconfidence, while the country that is behind will feel the crack of the whip to reform. … That drive will mean that competition will go on for years and decades.
https://danwang.co/ (2025 Annual letter)
The future is not predetermined by trends today. So it’s entirely possible that the dinosaur companies of today can’t figure out how to automate effectively, but get outcompeted by a nimble team of engineers using these tools tomorrow. As a concrete example, a lot of SaaS companies like Salesforce are at risk of this.
Re: GPT-5.2 derives a new result in theoretical physics
#229Re: GPT-5.2 derives a new result in theoretical physics
#230It's interesting to me that whenever a new breakthrough in AI use comes up, there's always a flood of people who come in to handwave away why this isn't actually a win for LLMs. Like with the novel solutions GPT 5.2 has been able to find for erdos problems - many users here (even in this very thread!) think they know more about this than Fields medalist Terence Tao, who maintains this list showing that, yes, LLMs hav…
They never surrender.