Live data from Hacker News

GPT-5.6 Sol Ultra will be in Codex

twitter.com

241–250 of 433 posts

Re: GPT-5.6 Sol Ultra will be in Codex

#241

Earlier quoted context omitted.

The craziest for me is companies that sticking stochastic agents into automated business processes and expecting stable/reliable outcomes. Businesses want deterministic processes in the vast majority of cases.

People are stochastic. You build reliable processes out of unreliable parts with feedback and self-correcting mechanisms. AI is not actually magically special in this regard. It has higher variance and we're still figuring out how to get all the tradeoffs right.

The big problem is that a person making a mistake can be taught to not make that mistake again. That's also not foolproof but at least it works a lot of the times. AI are unteachable, if you have given them a good prompt and they do something wrong 90% of the time you are shit out of luck.

That is to say I do agree that building reliable processes out of unreliable parts with feedback is the modus operandi. However AI cannot meaningfully handle feedback and learn. And that is a key unsolved problem.

Re: GPT-5.6 Sol Ultra will be in Codex

#242
post #175

Earlier quoted context omitted.

The craziest to me was someone saying “we are using AI in daily processes, now we need to automate”. But of course to some asshole non-technical people it meant asking for their vibe coded bullshit to be merged into production without review and fighting about it.

The craziest for me is companies that sticking stochastic agents into automated business processes and expecting stable/reliable outcomes. Businesses want deterministic processes in the vast majority of cases.

Yeah because "works many times in a row" = "deterministic" to many people.

Re: GPT-5.6 Sol Ultra will be in Codex

#243

Earlier quoted context omitted.

OpenAI's agreement with the Pentagon was "No use of OpenAI technology to direct autonomous weapons systems". And about the mass surveillance, I don't see why the military should not use AI to do surveillance abroad.

Maybe because it will make people abroad like you less and that has flow-on effects, mostly economic.

I think that applies to military involvement abroad generally.

If you are dropping bombs on someone I'm unconvinced the use of AI will make them like you more or less.

Re: GPT-5.6 Sol Ultra will be in Codex

#244

Earlier quoted context omitted.

Still avialable through the API. According to people that have tried both Fable nad 5.6, Fable is clearly better at coding. So i expect a lot of people to pay extra for it.

Who is going to pay API prices for using Claude? I don't know one single company. It's not gonna happen.

Every company above 150 employees has to pay API prices as Anthropic won't give you subscriptions.

Re: GPT-5.6 Sol Ultra will be in Codex

#245

Has anyone already tried 5.6 Sol in their day to day coding/development activities? How does it compare to GPT-5.5?

Has it been released yet? I'm not seeing GPT 5.6 anywhere in the selectors, nor any announcement about it, but you're talking about it as it's already been released?

There's something like 20 companies with access.

Re: GPT-5.6 Sol Ultra will be in Codex

#246

I wonder if it's related that that OpenAI has found a way to cut inference costs by half, according to The Information. https://www.theinformation.com/newsletters/ai-agenda/openai-...

https://archive.ph/NEwVz "However, these inference optimizations, which rival Anthropic refers to as “compute multipliers,” are a big focus for all the labs. Anthropic CEO Dario Amodei has been publicly talking about the concept since at least mid-2023, when he said on a podcast that the company limits “the number of people who are aware of a given compute multiplier” because it could give other AI labs a leg up if t…

And Anthropic sure reads and applies all the open research.

This 2023 thread about this issue is prescient: https://old.reddit.com/r/MachineLearning/comments/11sboh1/d_... (just add Anthropic to OpenAI)

Re: GPT-5.6 Sol Ultra will be in Codex

#248
post #48

Earlier quoted context omitted.

I recently have been testing ChatGPT business at work and the quota seems to disappear almost instantly even using weaker models. Unless they dramatically increase their quotas it’ll be unusable.

My solution for the ones stuck with that: use 5.5 for planning and 5.3-mini for the grunt work. 5.3 is remarkably useful still but you need to hold its hand.

I actually meant 5.4-mini..

Re: GPT-5.6 Sol Ultra will be in Codex

#249

Earlier quoted context omitted.

People are stochastic. You build reliable processes out of unreliable parts with feedback and self-correcting mechanisms. AI is not actually magically special in this regard. It has higher variance and we're still figuring out how to get all the tradeoffs right.

The big problem is that a person making a mistake can be taught to not make that mistake again. That's also not foolproof but at least it works a lot of the times. AI are unteachable, if you have given them a good prompt and they do something wrong 90% of the time you are shit out of luck. That is to say I do agree that building reliable processes out of unreliable parts with feedback is the modus operandi. However A…

"AI are unteachable, if you have given them a good prompt and they do something wrong 90% of the time you are shit out of luck."

please take a look at the error(s) made in the prior run. what could've been done better? create or modify an existing skill to emphasize this, or suggest additional language in AGENTS.md.

Re: GPT-5.6 Sol Ultra will be in Codex

#250

Earlier quoted context omitted.

Has it been released yet? I'm not seeing GPT 5.6 anywhere in the selectors, nor any announcement about it, but you're talking about it as it's already been released?

There's something like 20 companies with access.

Isn't it like 99% sure all those companies got that access to it together with NDAs and other goodies?
Post reply on HN