Input: $5/M tokens at Output: $30/M tokens at Cache read: $0.50/M tokens at Significantly more expensive than Opus 4.7 beyond 272K and at least in my tasks, I haven't seen the model that much more token efficient, certainly not to such a degree that it'd compensate this difference. GPT-5.4 had a solid context window at 400k with reliable compaction, both appear somewhat regressed, though still to early to truly say whether compaction is less reliable. Also, I have found frontend output to still skew towards that one very distinct, easily noticeable, card laden, bluesy hue overindulged template that made me skeptical of Horizon Alpha/Beta pre GPT-5s release. Ended up doing amazing at the time for task adherence, which made it very useful for me outside that one major deficit. The fact that GPT-5.5 is still so restricted in that area is weird considering it's supposed to be an entirely new foundation.
OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
101–110 of 174 posts
Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#102Earlier quoted context omitted.
This is snark. Since when has a junior level dev managed to debug and deploy say a cloudformation stack and follow up with notes under 3 minutes?
Heard this analogy elsewhere, but worth repeating: AI is like having the greatest developer who ever lived, but she is always on 4 beers.
Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#103Earlier quoted context omitted.
Heard this analogy elsewhere, but worth repeating: AI is like having the greatest developer who ever lived, but she is always on 4 beers.
personifying ai is incredibly cringe no matter how weird your comparison is
Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#104Earlier quoted context omitted.
> a lot of doctors are using ChatGPT both to search diagnosis and communicate with non-English speaking patients I think that's the problem. Who's going to claim responsibility when ChatGPT hallucinates or mistranslates a patient's diagnosis and they die? For OpenAI, this would at best be a PR nightmare, so that's why they have safeguards.
Adults bear responsibility for choices about their own lives. In fact, the more educated they are, the better choices they can make. A doctor who gets refused by ChatGPT doesn't stop needing to communicate with the patient; they fall back to a worse option (Google Translate, a family member interpreting, guessing). Refusal isn't safety, it's liability-shifting dressed up as safety. If there's no doctor, no interprete…
I think AI proves the contrary. There are plenty of examples of things that are getting worse because of technological advancement, particularly AI. Software quality, writing, online discourse, misinformation have all suffered over the last few years. I truly believe the internet is a worse place than it was 5 years ago, and I can't imagine bringing that to medicine would work out differently.
The medical system shouldn't rely on falling back to crappy workarounds, it should aspire to build the best system it reasonably can.
Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#105Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#106Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#107Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#108Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#109what's the real world comparison to opus 4.7 fellow coders?
4.6 did very well. 90% perfect on first try, got to 100% with just a few followups. 4.7 failed horribly. First produced garbage output and claimed it was done, admitted it did that when called out, proceeded to work at it a lot longer and then IT GAVE UP. GPT 5.5 codex was shockingly good. Achieved 90% perfect on first try in about a fourth of the time. Got to 100% faster and with fewer follow-ups.
I’m impressed.
Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API
#110Earlier quoted context omitted.
The doctor would be responsible. I had a choice better a doctor that used AI or not, I would much prefer one that did...
The doctor would be responsible for the accuracy of their translation tool, something they can't verify but you expect them to use?