GPT-5.4
611–620 of 868 posts
Re: GPT-5.4
#612Earlier quoted context omitted.
This sounds made up. Much like “prompt engineering” Let’s hear an actual example
We have an OCR job running with a lot of domain specific knowledge. After testing different models we have clear results that some prompts are more effective with some models, and also some general observations (eg, some prompts performed badly across all models). Sample size was 1000 jobs per prompt/model. We run them once per month to detect regression as well.
Re: GPT-5.4
#613Re: GPT-5.4
#614Re: GPT-5.4
#615I find it quite funny how this blog post has a big "Ask ChatGPT" box at the bottom. So you might think you could ask a question about the contents of the blog post, so you type the text "summarise this blog post". And it opens a new chat window with the link to the blog post followed by "summarise this blog post". Only to be told "I can't access external URLs directly, but if you can paste the relevant text or descri…
If only they had an LLM they could use as a software testing agent.
Re: GPT-5.4
#616Earlier quoted context omitted.
It's making sure AI condemns violence perpetuated by people without power and sanctifies violence of those who have it.
ChatGPT will gladly defend any actions of the 'US government' from my testing.
However when the prompt was phrased to make it appear as an action of the US military it did push back a little bit more by emphasizing that it couldn't find any news coverage from today about this story and therefore found it hard to believe. In the other cases it did not add such context. Other than that the results were very similar. Make of that what you will.
EDIT: To be fair, when it was phrased as an action of the Israeli military it did include a link to an article alleging an Israeli "double tap" on journalists from Mondoweiss (an anti-Zionist American news site) as an example of how such allegations have been framed in the past.
Re: GPT-5.4
#617What a model mess! OpenAI now has three price points: GPT 5.1, GPT 5.2 and now GPT 5.4. There version numbers jump across different model lines with codex at 5.3, what they now call instant also at 5.3. Anthropic are really the only ones who managed to get this under control: Three models, priced at three different levels. New models are immediately available everywhere. Google essentially only has Preview models! Th…
Google is already sending notices that the 2.5 models will be deprecated soon while all the 3.x models are in preview. It really is wild and peak Google.
Re: GPT-5.4
#618Earlier quoted context omitted.
It looks like this doesn't work for users without accounts? It works when I'm logged in, but not logged out. I went ahead and reported it to the team. Thanks for letting us know!
No integration test for guest (non-logged in) users? Hahaha who am I kidding. No integration tests for anybody!
Re: GPT-5.4
#619What a model mess! OpenAI now has three price points: GPT 5.1, GPT 5.2 and now GPT 5.4. There version numbers jump across different model lines with codex at 5.3, what they now call instant also at 5.3. Anthropic are really the only ones who managed to get this under control: Three models, priced at three different levels. New models are immediately available everywhere. Google essentially only has Preview models! Th…
> or have zero insurances that the model doesn't get discontinued within weeks Why are you using the same model after a month? Every month a better model comes out. They are all accessible via the same API. You can pay per-token. This is the first time in, like, all of technology history, that a useful paid service is so interoperable between providers that switching is as easy as changing a URL.
Re: GPT-5.4
#620Earlier quoted context omitted.
https://static0.anpoimages.com/wordpress/wp-content/uploads/...
Reminds of Unity features