GPT-5.3-Codex
41–50 of 634 posts
Re: GPT-5.3-Codex
#42While I love Codex and believe it's amazing tool, I believe their preparedness framework is out of date. As it is more and more capable of vibe coding complex apps, it's getting clear that the main security issues will come up by having more and more security critical software vibe coded.
It's great to look at systems written by humans and how well Codex can be used against software written by humans, but it's getting more important to measure the opposite: how well humans (or their own software) are able to infiltrate complex systems written mostly by Codex, and get better on that scale.
In simpler terms: Codex should write secure software by default.
Re: GPT-5.3-Codex
#43Re: GPT-5.3-Codex
#44Is this me or Sam is being absolute sore loser he is and trying to steal Opus thunder?
BTW, loser is spelled with a single o.
Re: GPT-5.3-Codex
#45Is this me or Sam is being absolute sore loser he is and trying to steal Opus thunder?
For downvoters, you must be naive to think these companies are not surveilling each other through various means.
Re: GPT-5.3-Codex
#46Re: GPT-5.3-Codex
#47Earlier quoted context omitted.
I do not trust the AI benchmarks much, they often do not line up with my experience. That said ... I do think Codex 5.2 was the best coding model for more complex tasks, albeit quite slow. So very much looking forward to trying out 5.3.
Another day, another hn thread of "this model changes everything" followed immediately by a reply stating "actually I have the literal opposite experience and find competitor's model is the best" repeated until it's time to start the next day's thread.
Re: GPT-5.3-Codex
#48It is absurd to release 5.3-Codex before first releasing 5.3. Also, there is no reason for OpenAI and Anthropic to be trying to one-up each other's releases on the same day. It is hell for the reader.
Re: GPT-5.3-Codex
#49Is this me or Sam is being absolute sore loser he is and trying to steal Opus thunder?
Re: GPT-5.3-Codex
#50Earlier quoted context omitted.
I do not trust the AI benchmarks much, they often do not line up with my experience. That said ... I do think Codex 5.2 was the best coding model for more complex tasks, albeit quite slow. So very much looking forward to trying out 5.3.
Another day, another hn thread of "this model changes everything" followed immediately by a reply stating "actually I have the literal opposite experience and find competitor's model is the best" repeated until it's time to start the next day's thread.