Live data from Hacker News

OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

developers.openai.com

111–120 of 174 posts

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#112
post #78

Earlier quoted context omitted.

Yes, the prompt is slim by design. I might be wrong, but the point was to see what the model can do "on it's own". The eval prompt is quite extensive: https://github.com/guilamu/llms-wordpress-plugin-benchmark/b...

That’s the thing, not everyone wants and values the model based on that. But I guess it works for you, and that benchmark achieves it. I personally develop with very detailed spec, and I don’t want nothing more and nothing less compared to the spec. I found 5.4/5.5 much better at following spec while Opus makes some things up, which aligns with your benchmark but that makes 5.4/5.5 better for me while worse for you.

Yeah as I said this a benchmark for my usecase only, a single use case, which is obvisouly not representative of everybody's needs.

What strike me as very strange though is that 0 model were able to just use the search input already present in GravitYForms forms list page and all created a second input.

Also, I know it's not in the prompt, but adding a ctrl+f shortcut to a search input? Is that that crazy? I don't know.

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#113
post #29
post #14

API page lists the knowledge cutoff as Dec 01, 2025 but when prompting the model it says June 2024. Knowledge cutoff: 2024-06 Current date: 2026-04-24 You are an AI assistant accessed via an API.

I don't know why this keeps coming up. This has always been the least reliable way to know the cutoff date (and indeed, it may well have been trained on sites with comments like these!) Just ask it about an event that happened shortly before Dec 1, 2025. Sporting event, preferably.

[dead]

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#116
post #69

Earlier quoted context omitted.

That’s actually crazy, what kind of task is that? And is that a recurring kind of task like some analysis, or coding related?

Coding (along with docs, tests obviously), rewriting a huge chunk of the KVM hypervisor (in Kernel 7, started in the -rc2) and KSM and other modules, can't say too much about it yet (might do an announcement in coming weeks) . The coding is automated but the plan took days of manual arguing (with all models possible) prior (while doing other things during waiting times as I currently manage 70 repos for an upcoming r…

Please do a post about this (though I realize that takes time). This sounds amazing. I have always dreamed of doing this too but just don't have the budget.

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#117

Earlier quoted context omitted.

The doctor would be responsible for the accuracy of their translation tool, something they can't verify but you expect them to use?

"what you see is all there is." it's generally much easier to verify something you've been made aware of than it is to know of it in the first place (and still verify it.)

The irony is that licensed interpreters / translators usually perform worse than AI.

Only the liability shifts from OpenAI to them.

Furthermore, where the alternative to a licensed professional was nothing, or a random untrained person or a weak professional, then it's harming the user on the pretext of protecting him.

(like in the other mentioned contexts).

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#118
post #67

Earlier quoted context omitted.

> a lot of doctors are using ChatGPT both to search diagnosis and communicate with non-English speaking patients I think that's the problem. Who's going to claim responsibility when ChatGPT hallucinates or mistranslates a patient's diagnosis and they die? For OpenAI, this would at best be a PR nightmare, so that's why they have safeguards.

Adults bear responsibility for choices about their own lives. In fact, the more educated they are, the better choices they can make. A doctor who gets refused by ChatGPT doesn't stop needing to communicate with the patient; they fall back to a worse option (Google Translate, a family member interpreting, guessing). Refusal isn't safety, it's liability-shifting dressed up as safety. If there's no doctor, no interprete…

[deleted]

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#119
post #69

Earlier quoted context omitted.

That’s actually crazy, what kind of task is that? And is that a recurring kind of task like some analysis, or coding related?

Coding (along with docs, tests obviously), rewriting a huge chunk of the KVM hypervisor (in Kernel 7, started in the -rc2) and KSM and other modules, can't say too much about it yet (might do an announcement in coming weeks) . The coding is automated but the plan took days of manual arguing (with all models possible) prior (while doing other things during waiting times as I currently manage 70 repos for an upcoming r…

I’m vague on a specific reason for this feeling because there are a few to choose from and no one overpowers the other, but the emotion that comes to mind when I read this is disgust. As a society I feel we will look back on the subsidized opulence of this moment with total and utter contempt.

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#120
post #119

Earlier quoted context omitted.

Coding (along with docs, tests obviously), rewriting a huge chunk of the KVM hypervisor (in Kernel 7, started in the -rc2) and KSM and other modules, can't say too much about it yet (might do an announcement in coming weeks) . The coding is automated but the plan took days of manual arguing (with all models possible) prior (while doing other things during waiting times as I currently manage 70 repos for an upcoming r…

I’m vague on a specific reason for this feeling because there are a few to choose from and no one overpowers the other, but the emotion that comes to mind when I read this is disgust. As a society I feel we will look back on the subsidized opulence of this moment with total and utter contempt.

Or nostalgia for simpler times
Post reply on HN