Live data from Hacker News

Building more with GPT-5.1-Codex-Max

openai.com

141–150 of 332 posts

Re: Building more with GPT-5.1-Codex-Max

#141

Earlier quoted context omitted.

Couldn't agree more about the google product offerings. Vertex AI? AI Studio? Maker studio? Gemini? The documentation is fragmented with redundant offerings making it confusing to determine what is what. GCS billing is complicated to figure out vs OpenAI billing or anthropic. Sad part is Google does offer a ChatML/OpenAI compliant endpoint to do LLM calls and I believe they in an experiment also reduced friction in g…

I've just found myself using OpenRouter if we need Google models for a project, it's worth the extra 5% just not to have to deal with the utter disaster that is their product offering.

FWIW I had to bail on the same thing because my results were drastically different. There was something happening with images through open router. Although outside of that I’d absolutely do the same thing, their apis are awful and billing worse. Maybe it makes sense for huge orgs but it’s a nightmare on the smaller scale.

Re: Building more with GPT-5.1-Codex-Max

#142

not sure if I am actually using 5.1-codex-max or just normal 5.1-codex (is there even 5.1-codex?) trying to continue work where gemini 3 left off and couple prompts in I had to switch back since it was reimplementing and changing things that didn't need changing and attempted to solve typos by making the code implementing those things work with the typo, weird behavior - probably is not compatible with the style gemi…

Just run the /model command in codex and select the model which you want.

Re: Building more with GPT-5.1-Codex-Max

#143
post #91

I rarely used Codex compared to Claude because it was extremely slow in GitHub copilot . Like maybe 2-5X slower than Claude Sonnet. I really wish they just made their models faster than “better”

OpenAI doesn't want you to use their models outside of their own products, which is why the API and integrations like Github Copilot are super slow.

That does not make business sense though. If people want to use Open AI models in Copilot and other tools and they dont perform they will just switch to another model and not come back they are not going to use Codex.

Re: Building more with GPT-5.1-Codex-Max

#144

Somewhat related, after seeing the praise for codex in the Sonnet 4.5 release thread I gave it a go, and I must say, that CLI is much worse than Claude Code (even if the model is great, I’m not sure where the issue really lies between the two). It was extremely slow (like, multiple times slower than Sonnet with Claude Code, though that’s partially on me for using thinking-high I guess) to finish the task, with the ba…

sure claude code has better ux but honestly its hard to get any good amount of usage out of the subscriptions vs what codex offers at the same price with claude im constantly hitting rate limits with codex getting substantially more and "slow" isn't really a problem for me as long as it keep working the only complaint i have is that codex itself has usage limited now (Either due to outstanding git issues around tools…

> if claude manages to release a smaller model

have you tried Haiku?

Re: Building more with GPT-5.1-Codex-Max

#145
post #49
post #40

I would love to see all the big players put 1% of the effort they put into model training into making the basic process of paying and signing in suck less. Claude: they barely have a signin system at all. Multiple account support doesn’t exist. The minimum seat count for business is nonsense. The data retention policies are weak. OpenAI: Make ZDR a thing you can use or buy without talking to sales, already. And for t…

Agree 1,000%. I just won’t even waste my time with the google stuff cuz I can’t figure out how to pay with it. And that’s a problem everywhere at google. Our google play account is suspended cuz I can’t verify the company. It won’t let me cuz it says I’m not the owner. I’ve always been the owner of my company. For 18 years. There is no one else. Once some error said make sure the owner email matches your profile in g…

[deleted]

Re: Building more with GPT-5.1-Codex-Max

#146
post #56

I've been using a lot of Claude and Codex recently. One huge difference I notice between Codex and Claude code is that, while Claude basically disregards your instructions (CLAUDE.md) entirely, Codex is extremely, painfully, doggedly persistent in following every last character of them - to the point that i've seen it work for 30 minutes to convolute some solution that was only convoluted because of some sentence I t…

GPT-5 is like that

Re: Building more with GPT-5.1-Codex-Max

#147
post #40

I would love to see all the big players put 1% of the effort they put into model training into making the basic process of paying and signing in suck less. Claude: they barely have a signin system at all. Multiple account support doesn’t exist. The minimum seat count for business is nonsense. The data retention policies are weak. OpenAI: Make ZDR a thing you can use or buy without talking to sales, already. And for t…

And stop asking for phone numbers for "fraud prevention" when I've already given you my name, address and credit card.

Can't people spoof the first two and use a stolen credit card number?

Re: Building more with GPT-5.1-Codex-Max

#149
post #15

Rest assured that we are better at training models than naming them ;D - New benchmark SOTAs with 77.9% on SWE-Bench-Verified, 79.9% on SWE-Lancer, and 58.1% on TerminalBench 2.0 - Natively trained to work across many hours across multiple context windows via compaction - 30% more token-efficient at the same reasoning level across many tasks Let us know what you think!

[deleted]

Re: Building more with GPT-5.1-Codex-Max

#150
post #15

Rest assured that we are better at training models than naming them ;D - New benchmark SOTAs with 77.9% on SWE-Bench-Verified, 79.9% on SWE-Lancer, and 58.1% on TerminalBench 2.0 - Natively trained to work across many hours across multiple context windows via compaction - 30% more token-efficient at the same reasoning level across many tasks Let us know what you think!

[flagged]
Post reply on HN