Earlier quoted context omitted.
I assume you're using the "regular" Pro version of Gemini 3.1 for the above, rather than the Deep Think mode, which is more comparable to GPT-5.5 Pro. To my knowledge, regular 3.1 Pro is a tier below and often makes mistakes. Moreover, there's no reason to believe the progress of LLMs, which couldn't reliably solve high-school math problems just 3–4 years ago, will stop anytime soon. You might want to track the progr…
> there's no reason to believe the progress of LLMs [...] will stop anytime soon Wrong. Every advancement has followed a s curve. Where we are on that curve is anyones guess. Or maybe "this time its different".
A recent experience with ChatGPT 5.5 Pro
261–270 of 558 posts
Re: A recent experience with ChatGPT 5.5 Pro
#262Earlier quoted context omitted.
I assume you're using the "regular" Pro version of Gemini 3.1 for the above, rather than the Deep Think mode, which is more comparable to GPT-5.5 Pro. To my knowledge, regular 3.1 Pro is a tier below and often makes mistakes. Moreover, there's no reason to believe the progress of LLMs, which couldn't reliably solve high-school math problems just 3–4 years ago, will stop anytime soon. You might want to track the progr…
> there's no reason to believe the progress of LLMs [...] will stop anytime soon Wrong. Every advancement has followed a s curve. Where we are on that curve is anyones guess. Or maybe "this time its different".
Re: A recent experience with ChatGPT 5.5 Pro
#263Earlier quoted context omitted.
Great. You see a shape in graphs. And that shape tells you that _at some unknown point in the future_ progress will slow (but likely not stop). Now back to the point, what reason do you have to believe progress will stop soon ? If you have no reason, then it sounds like you agree with OP. Which makes the patronizing sarcasm all that much more nauseating.
Nausea aside, what evidence does anyone have that “super intelligence” of the sort your argument alludes to is even possible? Because that’s what we’re really talking about; greater than human intelligence on this sort of academic task. For example; When llms start contributing meaningfully to their own development, that would be a convincing indicator imo.
This has been the case for awhile now already…
https://kersai.com/the-48-hours-that-changed-ai-forever-clau...
Re: A recent experience with ChatGPT 5.5 Pro
#264Earlier quoted context omitted.
> there's no reason to believe the progress of LLMs [...] will stop anytime soon Wrong. Every advancement has followed a s curve. Where we are on that curve is anyones guess. Or maybe "this time its different".
There are advancements that do not follow s curves - consider for instance total data transmitted over all networks, or financial derivatives volumes. I think a better question for AI is “is it more like a network effect, liquidity effect, or a biological/physical effect”?
Re: A recent experience with ChatGPT 5.5 Pro
#265Earlier quoted context omitted.
I actually consider compilers more impressive, and a compiler was an important part of making this possible.
To each their own. I mean compilers didn’t produce trillions of dollars of investment, and produce serious and profound philosophical questions about the nature of consciousness but you’re right, thank god we have C
Re: A recent experience with ChatGPT 5.5 Pro
#266Earlier quoted context omitted.
Great. You see a shape in graphs. And that shape tells you that _at some unknown point in the future_ progress will slow (but likely not stop). Now back to the point, what reason do you have to believe progress will stop soon ? If you have no reason, then it sounds like you agree with OP. Which makes the patronizing sarcasm all that much more nauseating.
Nausea aside, what evidence does anyone have that “super intelligence” of the sort your argument alludes to is even possible? Because that’s what we’re really talking about; greater than human intelligence on this sort of academic task. For example; When llms start contributing meaningfully to their own development, that would be a convincing indicator imo.
As the blog points out - this is one particular subfield where LLMs have much easier prospects - lots of low hanging fruit that “just” requires a couple weeks of PHD candidate research.
Mathematics itself is one of a small handful of endeavors where automated reinforcement training is extremely straightforward and can be done at massive scale without humans.
Neither of these factors place a structural bound on the kind of thing LLMs can be good at, but we are far from certain we can achieve performance at this level in other fields economically and in the near future.
Re: A recent experience with ChatGPT 5.5 Pro
#267Earlier quoted context omitted.
> there's no reason to believe the progress of LLMs [...] will stop anytime soon Wrong. Every advancement has followed a s curve. Where we are on that curve is anyones guess. Or maybe "this time its different".
This could be right for the current architecture of LLMs, but you can come up with specialized large language models that can more efficiently use tokens for a specific subset of problems by encoding the information differently ( https://www.nature.com/articles/d41586-024-03214-7 ). So if instead of text we come up with a different representation for mathematical or physical problems, that could both improve the qual…
That's precisely what happens on the bad side of a S curve.
Re: A recent experience with ChatGPT 5.5 Pro
#268Earlier quoted context omitted.
> there's no reason to believe the progress of LLMs [...] will stop anytime soon Wrong. Every advancement has followed a s curve. Where we are on that curve is anyones guess. Or maybe "this time its different".
There are advancements that do not follow s curves - consider for instance total data transmitted over all networks, or financial derivatives volumes. I think a better question for AI is “is it more like a network effect, liquidity effect, or a biological/physical effect”?
Or Roman trade volume before the Fall of Rome.
Not to mention what you describe is not technological improvement but increase in data or money flows, not the same.
Re: A recent experience with ChatGPT 5.5 Pro
#269It's a very long post with a mix of technical (math) and philosophical sections. Here are the most striking points to reflect upon IMHO. > It seems to me that training beginning PhD students to do research [...] has just got harder, since one obvious way to help somebody get started is to give them a problem that looks as though it might be a relatively gentle one. If LLMs are at the point where they can solve “gentl…
> Would we regard that as a major achievement of the mathematician? I don’t think we would. 1. Does it matter, really? 2. Is it very different from previous computer-aided proofs, philosophically?