Remember when everyone was predicting that GPT-5 would take over the planet?
GPT-5.4
311–320 of 868 posts
Re: GPT-5.4
#312Earlier quoted context omitted.
The model was released less than an hour ago, and somehow you've been able to form such a strong opinion about it. Impressive!
I am actually super impressed with Codex-5.3 extra high reasoning. Its a drop in replacement (infact better than Claude Opus 4.6. lately claude being super verbose going in circles in getting things resolved). I stopped using claude mostly and having a blast with Codex 5.3. looking forward to 5.4 in codex.
I've found that 5.3-Codex is mostly Opus quality but cheaper for daily use.
Curious to see if 5.4 will be worth somewhat higher costs, or if I'll stick to 5.3-Codex for the same reasons.
Re: GPT-5.4
#313Re: GPT-5.4
#314[flagged]
Re: GPT-5.4
#315Earlier quoted context omitted.
That last benchmark seemed like an impressive leg up against Opus until I saw the sneaky footnote that it was actually a Sonnet result. Why even include it then, other than hoping people don't notice?
It's only that one number that is for sonnet.
Re: GPT-5.4
#316Re: GPT-5.4
#317>Today, we’re releasing GPT‑5.4 in ChatGPT (as GPT‑5.4 Thinking),
>Note that there is not a model named GPT‑5.3 Thinking
They held out for eight months without a confusing numbering scheme :)
Re: GPT-5.4
#318It's interesting that they charge more for the > 200k token window, but the benchmark score seems to go down significantly past that. That's judging from the Long Context benchmark score they posted, but perhaps I'm misunderstanding what that implies.
[flagged]
Re: GPT-5.4
#319Earlier quoted context omitted.
Ironically this would actually be a good thing. As we can see from Iran Claude doesn’t quite have these bugs ironed out yet…
This is the exact attitude that lead to a chat bot being used to identify a school for girls as a valid target. The chatbot cannot be held responsible. Whoever is using chatbots for selecting targets is incompetent and should likely face war crime charges.