GPT-5.3-Codex
101–110 of 634 posts
Re: GPT-5.3-Codex
#102> GPT‑5.3‑Codex is our first model that was instrumental in creating itself. The Codex team used early versions to debug its own training
I'm happy to see the Codex team moving to this kind of dogfooding. I think this was critical for Claude Code to achieve its momentum.
Re: GPT-5.3-Codex
#103It's so difficult to compare these models because they're not running the same set of evals. I think literally the only eval variant that was reported for both Opus 4.6 and GPT-5.3-Codex is Terminal-Bench 2.0, with Opus 4.6 at 65.4% and GPT-5.3-Codex at 77.3%. None of the other evals were identical, so the numbers for them are not comparable.
I encourage people to try. You can even timebox it and come up with some simple things that might look initially insufficient but that discomfort is actually a sign that there's something there. Very similar to moving from not having unit/integration tests for design or regression and starting to have them.
Re: GPT-5.3-Codex
#104So can I use this from Opencode? Because Anthropic started to enforce their TOS to kill the Opencode integration
Re: GPT-5.3-Codex
#105Re: GPT-5.3-Codex
#106Re: GPT-5.3-Codex
#107Re: GPT-5.3-Codex
#108So can I use this from Opencode? Because Anthropic started to enforce their TOS to kill the Opencode integration
Re: GPT-5.3-Codex
#109When 2 multi billion giants advertise same day, it is not competition but rather a sign of struggle and survival. With all the power of the "best artificial intelligence" at your disposition, and a lot of capital also all the brilliant minds, THIS IS WHAT YOU COULD COME UP WITH? Interesting
Yeah they are both fighting for survival. No surprise really. Need to keep the hype going if they are both IPO'ing later this year.