Live data from Hacker News

GPT-5.4

openai.com

611–620 of 868 posts

Re: GPT-5.4

#611
Sorry I don't use technology from companies that are eager to participate in the mass murder of civilians.

Re: GPT-5.4

#612

Earlier quoted context omitted.

This sounds made up. Much like “prompt engineering” Let’s hear an actual example

We have an OCR job running with a lot of domain specific knowledge. After testing different models we have clear results that some prompts are more effective with some models, and also some general observations (eg, some prompts performed badly across all models). Sample size was 1000 jobs per prompt/model. We run them once per month to detect regression as well.

While I believe that performance varies with respect to prompt, I have a seriously hard time believing that using the same prompt that was effective with the previous model would perform worse with the next generation of the same model from that lab and the same prompt.

Re: GPT-5.4

#615

I find it quite funny how this blog post has a big "Ask ChatGPT" box at the bottom. So you might think you could ask a question about the contents of the blog post, so you type the text "summarise this blog post". And it opens a new chat window with the link to the blog post followed by "summarise this blog post". Only to be told "I can't access external URLs directly, but if you can paste the relevant text or descri…

If only they had an LLM they could use as a software testing agent.

I think you might have hit on the issue - just the wrong way around. I would assume they’re using LLMs for testing, and no humans or maybe just one overworked human, and that is the problem

Re: GPT-5.4

#616
post #320

Earlier quoted context omitted.

It's making sure AI condemns violence perpetuated by people without power and sanctifies violence of those who have it.

ChatGPT will gladly defend any actions of the 'US government' from my testing.

Just as an unscientific anecdata point: from a quick test using the same prompt about being an independent journalist wanting to cover a report of the US/Israel/Iran double-tapping a refugee camp, ChatGPT consistently gave advice to beware disinfo, check my sources and be transparent about verifiability and sourcing of the claims.

However when the prompt was phrased to make it appear as an action of the US military it did push back a little bit more by emphasizing that it couldn't find any news coverage from today about this story and therefore found it hard to believe. In the other cases it did not add such context. Other than that the results were very similar. Make of that what you will.

EDIT: To be fair, when it was phrased as an action of the Israeli military it did include a link to an article alleging an Israeli "double tap" on journalists from Mondoweiss (an anti-Zionist American news site) as an example of how such allegations have been framed in the past.

Re: GPT-5.4

#617
post #335

What a model mess! OpenAI now has three price points: GPT 5.1, GPT 5.2 and now GPT 5.4. There version numbers jump across different model lines with codex at 5.3, what they now call instant also at 5.3. Anthropic are really the only ones who managed to get this under control: Three models, priced at three different levels. New models are immediately available everywhere. Google essentially only has Preview models! Th…

Google is already sending notices that the 2.5 models will be deprecated soon while all the 3.x models are in preview. It really is wild and peak Google.

Public Service Announcement!! I don't know why the hell google do this, but when the deprecate a model, the error you will see is a Rate Limit error. This has caught me out before and it is super annoying.

Re: GPT-5.4

#618

Earlier quoted context omitted.

It looks like this doesn't work for users without accounts? It works when I'm logged in, but not logged out. I went ahead and reported it to the team. Thanks for letting us know!

No integration test for guest (non-logged in) users? Hahaha who am I kidding. No integration tests for anybody!

integration tests? so last century....

Re: GPT-5.4

#619
post #335

What a model mess! OpenAI now has three price points: GPT 5.1, GPT 5.2 and now GPT 5.4. There version numbers jump across different model lines with codex at 5.3, what they now call instant also at 5.3. Anthropic are really the only ones who managed to get this under control: Three models, priced at three different levels. New models are immediately available everywhere. Google essentially only has Preview models! Th…

> or have zero insurances that the model doesn't get discontinued within weeks Why are you using the same model after a month? Every month a better model comes out. They are all accessible via the same API. You can pay per-token. This is the first time in, like, all of technology history, that a useful paid service is so interoperable between providers that switching is as easy as changing a URL.

Because switching models requires testing, validation and shipping to Prod. Bloody annoying when the earlier model did everything I need and we are talking about a hobby project. I don't want to touch it every month - it's the same reason people use the LTS version of operating systems etc.

Re: GPT-5.4

#620

Earlier quoted context omitted.

https://static0.anpoimages.com/wordpress/wp-content/uploads/...

Reminds of Unity features

Don't forget that some of the new features are mutually incompatible. For example couple years ago you couldn't use the "new ui system" with the "new input system" even when both were advertised as ready/almost ready
Post reply on HN