Live data from Hacker News

GPT-5.3-Codex

openai.com

121–130 of 634 posts

Re: GPT-5.3-Codex

#121

The behind the scenes on deciding when to release these models has got to be pretty insanely stressful if they're coming out within 30 minutes-ish of each other.

I wonder if their "5.3" was continuously being updated, with regenerated benchmarks with each improvement, and they just stayed ready to release it when claude released

Re: GPT-5.3-Codex

#122
post #74

Earlier quoted context omitted.

What's funny is that most of this "progress" is new datasets + post-training shaping the model's behavior (instruction + preference tuning). There is no moat besides that.

>There is no moat besides that. Compute. Google didn't announce $185 billion in capex to do cataloguing and flash cards.

Google didn't buy 30% of Anthropic to starve them of compute

Re: GPT-5.3-Codex

#123

,,GPT‑5.3-Codex is the first model we classify as High capability for cybersecurity-related tasks under our Preparedness Framework , and the first we’ve directly trained to identify software vulnerabilities. While we don’t have definitive evidence it can automate cyber attacks end-to-end, we’re taking a precautionary approach and deploying our most comprehensive cybersecurity safety stack to date. Our mitigations inc…

That’s just classical OpenAI trying to make us believe they’re closing on AGI… Like all « so called » research from them and Anthropic about safety alignment and that their tech is so incredibly powerful that guardrails should be put on them.

Re: GPT-5.3-Codex

#124

May AI not write the code for me. May I at least understand what it has "written". AI help is good but don't replace real programmers completely. I'm enough copy pasting code i don't understand. What if one day AI will fall down and there will be no real programmers to write the software. AI for help is good but I don't want AI to write whole files into my project. Then something may broke and I won't know what's bro…

Honest question: have you tried evolving your code architecture when adding features instead of just "promting more and more things"?

Re: GPT-5.3-Codex

#125
post #93

I remember when AI labs coordinated so they didn't push major announcements on the same day to avoid cannibalizing each other. Now we have AI labs pushing major announcements within 30 minutes .

I wish they’d just stop pretending to care about safety, other than a few researchers at the top they care about safety only as long as they aren’t losing ground to the competition. Game theory guarantees the AI labs will do what it takes to ensure survival. Only regulation can enforce the limits, self policing won’t work when money is involved.

As long as China continues to blitz forward, regulation is a direct path to losing.

Re: GPT-5.3-Codex

#126
post #91

Earlier quoted context omitted.

The labs have fully embraced the cutthroat competition, the arms race has fully shed the civilized facade of beneficient mutual cooperation. Dirty tricks and underhanded tactics will happen - I think Demis isn't savvy in this domain, but might end up stomping out the competition on pure performance. Elon, Sam, and Dario know how to fight ugly and do the nasty political boardroom crap. 26 is gonna be a very dramatic y…

>civilized facade of mutual cooperation >Dirty tricks and underhanded tactics As long the tactics are legal ( i.e. not corporate espionage, bribes etc), the no holds barred full free market competition is the best thing for the market and the consumers.

The consumers are getting huge wins.

Model costs continue to collapse while capability improves.

Competition is fantastic.

Re: GPT-5.3-Codex

#127

Earlier quoted context omitted.

yea but i feel like we are over the hill on benchmaxxing, many times a model has beaten anthropic on a specific bench, but the 'feel' is that it is still not as good at coding

'feel' is no more accurate not saying there's a better way but both suck

Speak for yourself. I've been insanely productive with Codex 5.2.

With the right scaffolding these models are able to perform serious work at high quality levels.

Re: GPT-5.3-Codex

#128
post #57

Earlier quoted context omitted.

Yeah they are both fighting for survival. No surprise really. Need to keep the hype going if they are both IPO'ing later this year.

How many IPOs can a company really do?

As many as they want. They can "spin off" and then "merge" again.

Re: GPT-5.3-Codex

#129
I am on a max subscription for Claude, and hate the fact that OpenAI have not figured out that $20 => $200 is a big jump. Good luck to them. In terms of model, just last night, Codex 5.2 solved a problem for me which other models were going round and round. Almost same instructions. That said, I still plan to be on $100 Claude (overall value across many tasks, ability to create docs, co-work), and may bump up OpenAI subscription to the next tier should they decide to introduce one. Not going to $200 even with 5.3, unless my company pays for it.

Re: GPT-5.3-Codex

#130

Earlier quoted context omitted.

yea but i feel like we are over the hill on benchmaxxing, many times a model has beaten anthropic on a specific bench, but the 'feel' is that it is still not as good at coding

'feel' is no more accurate not saying there's a better way but both suck

The variety of tasks they can do and will be asked to do is too wide and dissimilar, it will be very hard to have a transversal measurement, at most we will have area specific consensus that model X or Y is better, it is like saying one person is the best coder at everything, that does not exist.
Post reply on HN