Live data from Hacker News

Erdos 281 solved with ChatGPT 5.2 Pro

twitter.com

61–70 of 310 posts

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#62
post #24
post #19

Earlier quoted context omitted.

It's pattern matching. Which is actually what we measure in IQ tests, just saying.

I call it matching. Pattern matching had a different meaning.

what are you referring to? LLMs are neural networks at their core and the most simple versions of neural networks are all about reproducing patterns observed during training

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#63

Earlier quoted context omitted.

Why not plan for a future where a lot of non-trivial tasks are automated instead of living on the edge with all this anxiety?

[flagged]

come out of the irony layer for a second -- what do you believe about LLMs?

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#64
post #47

Earlier quoted context omitted.

[flagged]

I suspect this is AI generated, but it’s quite high quality, and doesn’t have any of the telltale signs that most AI generated content does. How did you generate this? It’s great.

Your intuition on AI is out of date by about 6 months. Those telltale signs no longer exist.

It wasn't AI generated. But if it was, there is currently no way for anyone to tell the difference.

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#65
post #10

FWIW, I just gave Deepseek the same prompt and it solved it too (much faster than the 41m of ChatGPT). I then gave both proofs to Opus and it confirmed their equivalence. The answer is yes. Assume, for the sake of contradiction, that there exists an \(\epsilon > 0\) such that for every \(k\), there exists a choice of congruence classes \(a_1^{(k)}, \dots, a_k^{(k)}\) for which the set of integers not covered by the f…

I am not familiar with the field, but any chance that the deepseek is just memorizing the existing solution? Or different. https://news.ycombinator.com/item?id=46664976

Sure but if so wouldn't ChatGPT 5.2 Pro also "just memorizing the existing solution?"?

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#66
post #16

I guess the first question I have is if these problems solved by LLMs are just low-hanging fruit that human researchers either didn't get around to or show much interest in - or if there's some actual beef here to the idea that LLMs can independently conduct original research and solve hard problems.

That's the first warning from the wiki : > https://github.com/teorth/erdosproblems/wiki/AI-contribution...

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#67

Earlier quoted context omitted.

It's bizarre. The same account was previously arguing in favor of emergent reasoning abilities in another thread ( https://news.ycombinator.com/item?id=46453084 ) -- I voted it up, in fact! Turing test failed, I guess. (edit: fixed link)

I thought the mockery and sarcasm in my piece was rather obvious.

Poe's Law is the real Bitter Lesson.

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#68
post #16

I guess the first question I have is if these problems solved by LLMs are just low-hanging fruit that human researchers either didn't get around to or show much interest in - or if there's some actual beef here to the idea that LLMs can independently conduct original research and solve hard problems.

There is still value on letting these LLMs loose on the periphery and knocking out all the low hanging fruit humanity hasn’t had the time to get around to. Also, I don’t know this, but if it is a problem on Erdos I presume people have tried to solve it atleast a little bit before it makes it to the list.

Is there though? If they are "solved" (as in the tickbox mark them as such, through a validation process, e.g. another model confirming, formal proof passing, etc) but there is no human actually learning from them, what's the benefit? Completing a list?

I believe the ones that are NOT studied are precisely because they are seen as uninteresting. Even if they were to be solved in an interesting way, if nobody sees the proof because they are just too many and they are again not considered valuable then I don't see what is gained.

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#69
post #65

Earlier quoted context omitted.

I am not familiar with the field, but any chance that the deepseek is just memorizing the existing solution? Or different. https://news.ycombinator.com/item?id=46664976

Sure but if so wouldn't ChatGPT 5.2 Pro also "just memorizing the existing solution?"?

No it's not, you can refer to my link and subsequent discussion.

Re: Erdos 281 solved with ChatGPT 5.2 Pro

#70
post #7

I have 15 years of software engineering experience across some top companies. I truly believe that ai will far surpass human beings at coding, and more broadly logic work. We are very close

HN will be the last place to admit it; people here seem to be holding out with the vague 'I tried it and it came up with crap'. While many of us are shipping software without touching (much) code anymore. I have written code for over 40 years and this is nothing like no-code or whatever 'replacing programmers' before, this is clearly different judging from the people who cannot code with a gun to their heads but stil…

Both can be correct : you might be making a lot of money using the latest tools while others who work on very different problems have tried the same tools and it's just not good enough for them.

The ability to make money proves you found a good market, it doesn't prove that the new tools are useful to others.

Post reply on HN