Live data from Hacker News

A recent experience with ChatGPT 5.5 Pro

gowers.wordpress.com

291–300 of 558 posts

Re: A recent experience with ChatGPT 5.5 Pro

#291
post #35

Earlier quoted context omitted.

But perhaps we should regard it as a major achievement.

I mean in the same way getting Wolfram Alpha to solve a really hard/ugly differential equation I suppose

Mario Andretti could never have won a motor race without a car, yet we say he won the Indy500.

Re: A recent experience with ChatGPT 5.5 Pro

#292
post #136

Earlier quoted context omitted.

I assume you're using the "regular" Pro version of Gemini 3.1 for the above, rather than the Deep Think mode, which is more comparable to GPT-5.5 Pro. To my knowledge, regular 3.1 Pro is a tier below and often makes mistakes. Moreover, there's no reason to believe the progress of LLMs, which couldn't reliably solve high-school math problems just 3–4 years ago, will stop anytime soon. You might want to track the progr…

There are many indications that model progress is slowing down, so that is not entirely accurate.

Model progress at spitting out unhallucinated facts is slowing down hard. Model progress at solving hard math challenges/programming tasks doesn't seem to be slowing down that I can tell.

Re: A recent experience with ChatGPT 5.5 Pro

#293

Earlier quoted context omitted.

Compilers just made it all possible, but they are not new and shiny. LLMs did not produce the philosophical questions, but they do raise them. It's worth noting that computers have been changing the way we think about consciousness long before LLMs, largely thanks to compilers.

Yea there’s no logical stopping point when you use that logic. Why not say electricity or the element silicon?

I don't think the level of investment in an idea is equivalent to how impressive it may be. Most of the investment in AI is based on the idea that it will make professions and human labour obsolete, which means whoever has the reins at the moment it "solves" the "problem" of human labour will effectively reign over everyone else. The level of investment is then somewhat orthogonal to how technically impressive it is.

Not to mention that the less easily-explainable a technical achievement is, the less investment it will attract simply because fewer people will grasp the ramifications. You can describe AI in two words ("machine human") while it would take a few more to describe compilers in an instantly understandable way.

Re: A recent experience with ChatGPT 5.5 Pro

#295
The vast, vast majority of students going into higher education this fall will not contribute much to science until 4-5 years down the road (should they do research). Realistically 6-7 when they're in full swing with their Ph.D.

If we look where these models were 5-7 years ago...the existential threat of the Ph.D. was not even on the radar back then. The people finishing up their doctorate now are the first that can truly leverage these tools.

Now, if these to-be researcher students feel defeated (enough to quit), or completely lean on AI models the work for them, we're going to have a problem. Same with the funding of those Ph.D. positions. If we move away from "funding to produce researchers" to "funding to achieve results", will money that was usually spent to fund Ph.D. students start to flow towards compute?

If we look at it a bit cynically: Some researcher will be able to pump out a lot more papers by spending money on compute, than a couple of years of training students.

Interesting times. But also so much uncertainty. I feel terrible for the students that will have to decide now what they want to do, with all this knowledge.

Re: A recent experience with ChatGPT 5.5 Pro

#296
post #141

Earlier quoted context omitted.

> there's no reason to believe the progress of LLMs [...] will stop anytime soon Wrong. Every advancement has followed a s curve. Where we are on that curve is anyones guess. Or maybe "this time its different".

This could be right for the current architecture of LLMs, but you can come up with specialized large language models that can more efficiently use tokens for a specific subset of problems by encoding the information differently ( https://www.nature.com/articles/d41586-024-03214-7 ). So if instead of text we come up with a different representation for mathematical or physical problems, that could both improve the qual…

> So if instead of text we come up with a different representation for mathematical or physical problems, that could both improve

But then, wouldn't we first have to translate all of our current math and physics knowledge into that new representation in order to be able to train a model on it? Looks like a tremendous amount of work to me.

Re: A recent experience with ChatGPT 5.5 Pro

#297
post #110

Earlier quoted context omitted.

While this sounds generous (and in some ways it is), it does not address the general point that GP is making. That is, the systematic disadvantage which large parts of humanity have w.r.t. to access to the tools. You could say they can't drive a Lambhorgini either, but that also doesn't solve the problem.

Its a problem of the individual institutions and countries. The budget required for AI tools currently is negligible compared to other university expenses. We don't need to call everything a systemic disadvantage when the disadvantaged (at the institution level) have agency here.

> The budget required for AI tools currently is negligible compared to other university expenses.

Is it? Do you have any idea what the salary of a mid-tier university researcher in an Eastern European country is? Or in Africa or south-east Asia? With sota LLM pricing you easily get into the same order of magnitude, so essentially labour cost would double for researchers at such universies. Not "negligible" at all.

Re: A recent experience with ChatGPT 5.5 Pro

#298
post #166

> "Even though I can motivate it in retrospect, ChatGPT’s idea to use h^2-dissociated sets to control relations of order at most h feels quite ingenious. As far as I can tell, this idea is completely original." The question that keep bothering me is can an LLM generate an idea that is truly novel? How would/could that actually happen? But then that leads to the question - what are we actually doing when we think? Per…

Yes, they can. Some people like to parrot "next token prediction", "LLMs can only interpolate", and other nonsense, but it is obviously not true for many reasons, in particular since we introduced RL. Humans do not have the monopoly on generating novel ideas, modern AI models using post training, RL etc can come to them in the same way we do, exploration. See also verifier's law [0]: "The ease of training AI to solve…

RL or no RL, AI cannot escape the distribution it's trained on. It's just that the labs will put so much into the distribution that we won't be able to tell the difference that easily, nor will it matter for most tasks. The reason AI does well on ARC-AGI-2 is because the labs created synthetic training data using similar puzzles.

Re: A recent experience with ChatGPT 5.5 Pro

#300
post #110

Earlier quoted context omitted.

Its a problem of the individual institutions and countries. The budget required for AI tools currently is negligible compared to other university expenses. We don't need to call everything a systemic disadvantage when the disadvantaged (at the institution level) have agency here.

Can you tell me what is the budget necessary to supply AI tools capable of substantial research assistance to all academic staff at a university? You seem to have a good estimate in your head; I definitely do not. From personal experience, ChatGPT 5.5 (the Plus tier) is excellent for programming tasks and also for various teaching related tasks but I have not observed the research benefits that Tim Gowers has when I…

Which is good, since public money is tax money, so it better be spent wisely and not just thrown at the latest hype without thinking properly about it. It's a feature that public spending moves slowly, we should all be thankful for it.
Post reply on HN