Live data from Hacker News

OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

developers.openai.com

151–160 of 174 posts

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#151
post #79

Earlier quoted context omitted.

gemma4-e4b is 50% better than gemma4-26b in your benchmark, something's wrong

Yes those two models were tested on my own PC (local inference using my own CPU/GPU). So something my be bugged on my setup. gemma4-26b should be far better than gemma4-e4b.

The early quants for Gemma4 26b had issues and needed to be updated, might be worth checking

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#154
post #56

Just tested it on my homemade Wordpress+GravityForms benchmark and it's one of the worst model of the leaderboard performance wise and the worst value wise: https://github.com/guilamu/llms-wordpress-plugin-benchmark I know it's only on a single benchmark, but I dont understand how it can be so bad...

A junior tinkering in their garage in domains they have little experience executed a flawed test and decided to call it a benchmark. It's extremely common nowadays because words dont mean anything anymore. The forums that used to be filled with technical people doing real work are now filled with the masses of vibe researchers doing this kind of stuff. This is what happens when anything goes over some popularity threshold.

HN is the last bastion of serious inquiry these days. But its not immune as OPs comment proves.

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#155

Earlier quoted context omitted.

personifying ai is incredibly cringe no matter how weird your comparison is

Imagine a drunk developer. Sparks of brilliance while missing obvious trees.

I know of a publicly traded company which in its early years was built on beer. Literally. 3 guys in a co-working space in Cambridge, MA. Beer fueled their progress. 15 years later the software is still the backbone of the org.

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#156

Earlier quoted context omitted.

I gave 4.6, 4.7 and GPT 5.5 the same prompt and task to reverse engineer a collection of sample vector files from an obscure Amiga CAD program and create a detailed txt specification and a python converter that converts to SVG and produce a report so I can visually verify. 4.6 did very well. 90% perfect on first try, got to 100% with just a few followups. 4.7 failed horribly. First produced garbage output and claimed…

Interesting that 4.7 failed like that. Seems 5.5 is impressive but is oh so expensive. Would be interesting if you ran your same test with Deepseek v4 and some of the other Chinese models.

Just tried with DeepSeek V4 Pro with OpenCode. It didn't do great. First attempt produced somewhat correct drawings for some of the original samples, but most were just a spaghetti messs of lines. Some prodding got it to do a little better, but still not right. A third prod and it went down a wild rabbit hole and was much worse. I gave up.

I also tried GLM 5.1, it's first attempt was such a disaster I didn't bother working with it any further. It also took by far the longest and wasted a bunch of time/tokens trying to find other converters online (and failing) instead of just reverse engineering the format from the sample files given.

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#157

Earlier quoted context omitted.

Please do an update when you're ready, this sounds like madness to me so I'd love to see what the output is. Whatever it is I have to know.

Typical AI psychosis. They might notice it soon or stay in this condition for months.

I don't think you really grasp the direction the world is taking or even really understand AI capabilities when it's put together to reach high automation, you might not agree or embrace it yet, but you will be joining the loop wagon, soon enough.

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#158
post #123

Earlier quoted context omitted.

Has that task accomplished anything yet?

It made Sam richer.

I don't know their margin so I can't really say, but do we have 8 OpenAI accounts, I doubt they are making that much with us seeing that there isn't a single hour where we don't saturate the accounts.

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#159
post #69

Earlier quoted context omitted.

That’s actually crazy, what kind of task is that? And is that a recurring kind of task like some analysis, or coding related?

Coding (along with docs, tests obviously), rewriting a huge chunk of the KVM hypervisor (in Kernel 7, started in the -rc2) and KSM and other modules, can't say too much about it yet (might do an announcement in coming weeks) . The coding is automated but the plan took days of manual arguing (with all models possible) prior (while doing other things during waiting times as I currently manage 70 repos for an upcoming r…

I'm also in that boat of not understanding how people fail to get a huge productivity boost from GenAI. And it's not just novices but sometimes seriously accomplished coders. It can't be they're just typing 'Make me an ERP' and then go 'these thing are dumb slop machines' right?

Re: OpenAI releases GPT-5.5 and GPT-5.5 Pro in the API

#160
post #122

Earlier quoted context omitted.

Or nostalgia for simpler times

That as well. But everyone reading GP’s posts knows in their bones that it’s unsustainable. It’s economically unsustainable and environmentally unsustainable, and in that context it strikes me as pure hoarding behaviour. Taking as much as they can for themselves before the house of cards crashes down. I have no sympathy for OpenAI or Anthropic as corporations, but if these are the new tools of the trade, then platfor…

I mostly hate it because the token crunch is now coming for us regular users because of people like this. A few people always ruin it for the rest of us.
Post reply on HN