Live data from Hacker News

GPT-5.2

openai.com

71–80 of 1001 posts

Re: GPT-5.2

#71

The only table where they showed comparisons against Opus 4.5 and Gemini 3: https://x.com/OpenAI/status/1999182104362668275 https://i.imgur.com/e0iB8KC.png

100% on the AIME (assuming its not in the training data) is pretty impressive. I got like 4/15 when I was in HS...

Re: GPT-5.2

#73
post #31

For me the last remaining killer feature of ChatGPT is the quality of the voice chat. Do any of the competitors have something like that?

Try elevenlabs

Does elevenlabs have a real-time conversational voice model? It seems like like their focus is largely on text to speech and speech to text. Which can approximate that type of thing but it's not at all the same as the native voice to voice that 4o does.

Re: GPT-5.2

#74
post #9
post #3

Everything is still based on 4 4o still right? is a new model training just too expensive? They can consult deepseek team maybe for cost constrained new models.

Apparently they have not had a successful pre training run in 1.5 years

What kind of issues could prevent a company with such resources from that?

Re: GPT-5.2

#75

Earlier quoted context omitted.

We're also in benchmark saturation territory. I heard it speculated that Anthropic emphasizes benchmarks less in their publications because internally they don't care about them nearly as much as making a model that works well on the day-to-day

How do you measure whether it works better day to day without benchmarks?

Subscriptions.

Re: GPT-5.2

#76
post #9

Earlier quoted context omitted.

Apparently they have not had a successful pre training run in 1.5 years

I want to read a short scify story set in 2150 about how, mysteriously, no one has been able to train a better LLM for 125 years. The binary weights are studied with unbelievably advanced quantum computers but no one can really train a new AI from scratch. This starts cults, wars and legends and ultimately (by the third book) leads to the main protagonist learning to code by hand, something that no human left alive s…

You can ask 2025 Ai to write such a book, it's happy to comply and may or may not actually write the book

https://www.pcgamer.com/software/ai/i-have-been-fooled-reddi...

Re: GPT-5.2

#77
Wish they would include or leak more info about what this is, exactly. 5.1 was just released, yet they are claiming big improvements (on benchmarks, obviously). Did they purposely not release the best they had to keep some cards to play in case of Gemini 3 success or is this a tweak to use more time/tokens to get better output, or what?

Re: GPT-5.2

#78

Pricing is the same?

ChatGPT pricing is the same. API pricing is +40% per token, though greater token efficiency means that cost per task is not always that much higher. On some agentic evals we actually saw costs per task go down with GPT-5.2. It really depends on the task though; your mileage may vary.

Re: GPT-5.2

#79
post #51

Are there any specifics about how this was trained? Especially when 5.1 is only a month old. I'm a little skeptical of benchmarks these days and wish they put this up on llmarena edit: noticed 5.2 is ranked in the webdev arena (#2 tied with gemini-3.0-pro), but not yet in text arena (last update 22hrs ago)

Unfortunately there are never any real specifics about how any of their models were trained. It's OpenAI we're talking about after all.

Re: GPT-5.2

#80
Man this was rushed, typo in the first section:

> Unlike the previous GPT-5.1 model, GPT-5.2 has new features for managing what the model "knows" and "remembers to improve accuracy.

Post reply on HN