Live data from Hacker News

GPT-5.4

openai.com

451–460 of 868 posts

Re: GPT-5.4

#453

Earlier quoted context omitted.

This is the low quality reddit-style garbage that gets upvoted on HN these days?

As programmers become intelligently irrelevant in the whole picture, you would see more posts like this

"This account belongs to a lazy person" true

Re: GPT-5.4

#454

Earlier quoted context omitted.

in my testing codex actually planned worse than claude but coded better once the plan is set, and faster. it is also excellent to cross check claude's work, always finding great weakness each time.

That’s why I think the sweet spot is to write up plans with Claude and then execute them with Codex

Weird. It used to be the opposite. My own experience is that Claude’s behind-the-scenes support is a differentiator for supporting office work. It handles documents, spreadsheets and such much better than anyone else (presumably with server side scripts). Codex feels a bit smarter, but it inserts a lot of checkpoints to keep from running too long. Claude will run a plan to the end, but the token limits have become so small in the last couple months that the $20 pla basically only buys one significant task per day. The iOS app is what makes me keep the subscription.

Re: GPT-5.4

#456

I've only used 5.4 for 1 prompt (edit: 3@high now) so far (reasoning: extra high, took really long), and it was to analyse my codebase and write an evaluation on a topic. But I found its writing and analysis thoughtful, precise, and surprisingly clearly written, unlike 5.3-Codex. It feels very lucid and uses human phrasing. It might be my AGENTS.md requiring clearer, simpler language, but at least 5.4's doing a good…

> It might be my AGENTS.md requiring clearer, simpler language If you gave the exact same markdown file to me and I posted ed the exact same prompts as you, would I get the same results?

I'm not sure if the model (under its temperature/other settings) produces deterministic responses. But I do think models' style and phrasing are fairly changeable via AGENTS.md-style guidelines.

5.4's choice of terms and phrasing is very precise and unambiguous to me, whereas 5.3-Codex often uses jargon and less precise phrases that I have to ask further about or demand fuller explanations for via AGENTS.md.

Re: GPT-5.4

#457

Earlier quoted context omitted.

Plasma physicist here, I haven't tried 5.4 yet, but in general I am very impressed with the recent upgrades that started arriving in the fall of 2025: for tasks like manipulating analytic systems of equations, quickly developing new features for simulation codes, and interpreting and designing experiments (with pictures) they have become much stronger. I've been asking questions and probing them for several years now…

Youre just chatting yourself out of a job.

If we don't need plasma physicists anymore then we probably have fusion reactors or something, which seems like a fine trade. (In reality we're going to want humans in the loop for for the forseeable future)

Re: GPT-5.4

#459
post #195

[flagged]

What makes you think that they see bombing civilians as a bug, not a feature?

first real comment, I thought that at first but this could lower the possible users that could be using chatGPT and that would be against us (shareholders)

Re: GPT-5.4

#460

Earlier quoted context omitted.

But, why include the non-functional chat box in the article?

Welcome to a big company

Welcome to a big company where pretty much everyone has been working full steam for years, in order to take advantage of having a job at a company during a once-in-a-lifetime moment.
Post reply on HN