Live data from Hacker News

GPT-4.1 in the API

openai.com

11–20 of 513 posts

Re: GPT-4.1 in the API

#11
post #4

> We will also begin deprecating GPT‑4.5 Preview in the API, as GPT‑4.1 offers improved or similar performance on many key capabilities at much lower cost and latency. GPT‑4.5 Preview will be turned off in three months, on July 14, 2025, to allow time for developers to transition. Well, that didn't last long.

so we're going back... .4 of a gpt? make it make sense openai..

Re: GPT-4.1 in the API

#16
post #4

> We will also begin deprecating GPT‑4.5 Preview in the API, as GPT‑4.1 offers improved or similar performance on many key capabilities at much lower cost and latency. GPT‑4.5 Preview will be turned off in three months, on July 14, 2025, to allow time for developers to transition. Well, that didn't last long.

[deleted]

Re: GPT-4.1 in the API

#17

No benchmark comparisons to other models, especially Gemini 2.5 Pro, is telling.

Gemini 2.5 Pro gets 64% on SWE-bench verified. Sonnet 3.7 gets 70%

They are reporting that GPT-4.1 gets 55%.

Re: GPT-4.1 in the API

#19

It seems that OpenAI is really differentiating itself in the AI market by developing the most incomprehensible product names in the history of software.

They learned from the best: Microsoft

Re: GPT-4.1 in the API

#20
> Note that GPT‑4.1 will only be available via the API. In ChatGPT, many of the improvements in instruction following, coding, and intelligence have been gradually incorporated into the latest version (opens in a new window) of GPT‑4o, and we will continue to incorporate more with future releases.

The lack of availability in ChatGPT is disappointing, and they're playing on ambiguity here. They are framing this as if it were unnecessary to release 4.1 on ChatGPT, since 4o is apparently great, while simultaneously showing how much better 4.1 is relative to GPT-4o.

One wager is that the inference cost is significantly higher for 4.1 than for 4o, and that they expect most ChatGPT users not to notice a marginal difference in output quality. API users, however, will notice. Alternatively, 4o might have been aggressively tuned to be conversational while 4.1 is more "neutral"? I wonder.

Post reply on HN