Earlier quoted context omitted.
Sorry to be nihilist, but you never had any objective value if you're thinking in these terms. As far as we know, the universe "just is". There is no universal objective value of human beings, at all, any one of us. You have to make or find your own value in the universe. I try not to think too hard about the nihilist side and try to appreciate that for some unfathomable reason, I seem to have what I call consciousne…
> As far as we know, the universe "just is". I don't know this. In fact, billions of people around the world don't know this. In fact, all evidence points to the contrary. You have objective value being made in the image of a personal God. Denying that leads to a lot of pain, namely nihilistic suffering because it's on you to "pull yourself up by the bootstraps" in any endeavor involving your own self-worth.
GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
411–420 of 467 posts
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#412If they tried this on 1000 problems and this is the one that succeeded, it still means that there are 999 open problems that an LLM cannot one-shot. It seems likely that this would remain the situation until the next model.
If this is the first one they tried, maybe we’re totally hosed.
The conclusions are so different in these cases that it is impossible to know what to think. Though it is reasonable, I think, to assume that a company is willing to push the maximally misleading narrative —- especially a company known for questionable ethical direction at the top, and one that is still circling an IPO, and one that is in the tech industry, where conjuring an illusion of growth and progress is sufficient for success.
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#413Earlier quoted context omitted.
I think a lot of this has to do with the post-training these models normally get. They are designed to answer basic questions with straightforward and short summary answers. They have the capacity to reason deeply, but they are not biased towards that unless prompted. I think it's because LLMs as they are in 2026 are both highly capable but also parlor tricks. They are not sentient, you just set them up with the cont…
Even Fable hallucinates. I had it tracking down some very obscure Ancient Greek inscriptions and the response just made up a translation/context for one inscription after "looking it up." Now, it was still a very particular thing and I really had to get into the weeds to push it to that point, but who knows how many other gaps, near or far, it will happily skip over just for the sake of coherence. I think this is an…
Everything that an LLM outputs is just a statistical language-based (no real grounding) prediction. Luckily with a model based on a large training set most common questions may elicit coherent responses from the training data, but you don't need to veer too far off into "questions less asked" territory to get responses based on training data mashups that amount to best guesses that are wrong, aka hallucinations. The unfortunate part of this is that as a user you may only catch this when asking a question about something you are already fairly knowledgeable about, then you give some pushback to the model and it cheerfully acknowledges "you're right - I made that up".
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#414Earlier quoted context omitted.
No one here actually cares about folding laundry. I can demonstrate it by pointing out at the absence of posts on that subject. ...but when an affordable robot that folds laundry becomes available, people here pay attention.
Oh, for laundry to be solved!
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#415Both impressive and terrifying. But as always, the methodology is buried: how many open problems were tried until they found a success? If they tried this on 1000 problems and this is the one that succeeded, it still means that there are 999 open problems that an LLM cannot one-shot. It seems likely that this would remain the situation until the next model. If this is the first one they tried, maybe we’re totally hos…
Not only that, but they have like 500 world-leading experts in mathematics and IMO alumni, so how do we know one of the agents wasn't hardcoded to return a proof that the mathematicians had found?
I'm a mathematician/graph theorist, and I've tried ChatGPT 5.3, 5.4, 5.5, and now 5.6 on a bunch of simple-ish open problems, and I've never gotten a solution.
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#416are the references real? how do you think it got access to those papers? were they somehow already in the training data, or a result of web searches, Google scholar, etc? None of them include a web URL but in text some are super specific ("[3, Sections 2.1 and 3.1]" and "[8, p. 367]"). The references go back to 1954 (Chronologically sorted: 1954, 1973, 1975, 1976, 1978, 1979, 1981, 1985, 1987 and 1994.) Since referen…
Yes, reference 10 jumped out at me as well. I thought personal correspondence references typically include one of the authors of the paper.
https://scholar.google.com/scholar?q=W.T.%20Tutte%2C%20Perso....
Sloppy scholarship. On the other hand, it's simply a credit attribution of posing the problem, so it's not material in evaluating the results. I observe that the majority of references I can find that attribute this to Tutte are very indirect - i.e., citing sources that themselves claim Tutte was one of the people who formulated it - so it would take someone with a little more time on their hands (or perhaps an LLM) to track down the original...
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#417Earlier quoted context omitted.
It is very concise, and reads precisely as you suggest: to exploit properties already discovered and therefore combined in a novel way. I'm just delighted by the prose. It reads like an old paper. The ones that were just straightforward theorems with proofs that do exactly what they say.
In my (very) limited use of GPT-5.6, I have noticed it is quite concise in general, and significantly better at abstract thinking. Doing a PR review of a large change it was interesting to see Fable and 5.6 mention a few similar points with Fable much more long-winded and less readable, while 5.6 caught more "second-level" concerns and Fable more "in the code" concerns, so they both are quite useful in concert. In ge…
Fortunately, OpenAI APIs expose the verbosity parameter, which is separate to effort. If you want longer responses, you can. Or just prompt it.
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#418Earlier quoted context omitted.
This is one reason why I can't stand hn. Edgelord nihilistic comments that say nothing matters and take everything for granted get up-voted while credible and rational statements about God get down-voted.
That is an incredibly insulting comment. I am married, two children, have lead a fantastic fulfilling life. Just because I don't believe the Flying Spaghetti Monster created the universe doesn't mean I am an "Edgelord". Remember the phrase, you also don't believe in God. There are hundreds of gods you don't believe in.
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#419Unrelated to the accomplishment or proof itself, but it's interesting how much of the prompt, even in this latest-and-greatest model, is spent essentially telling the model to actually solve the problem. Things like "Reject status reports, vague optimism, and claims that an unproved global compatibility statement is 'routine'." Also a lot prompt spent feeding it strategies, which feel like they should/will eventually…
Re: GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]
#420Unrelated to the accomplishment or proof itself, but it's interesting how much of the prompt, even in this latest-and-greatest model, is spent essentially telling the model to actually solve the problem. Things like "Reject status reports, vague optimism, and claims that an unproved global compatibility statement is 'routine'." Also a lot prompt spent feeding it strategies, which feel like they should/will eventually…
“Kick logic out and do the impossible! Remember that, that’s the way Team Gurren rolls!”
The sweet irony is all the jailbreak style fixes could hamper this approach.