Live data from Hacker News

GPT‑5.3 Instant

openai.com

161–170 of 330 posts

Re: GPT‑5.3 Instant

#161
"Instant" is really going to age poorly as far as a brand name goes, especially with Taalas ( https://chatjimmy.ai ) proving out that baked silicon models can be truly instant.

I was literally posting about this earlier this morning[1], but all data indicates that we'll have models equivalent to Opus 4.6 / GPT 5.3 with a truly instant (ie > 10k t/s) response time by 2028. Small models are getting better faster, and their ability to be baked into silicon in a power and speed efficient way is likely going to completely disrupt things.

[1] https://x.com/pwnies/status/2028831699736637912

Re: GPT‑5.3 Instant

#162

I kind of chuckled when I read the headline "GPT‑5.3 Instant: Smoother, more ..." LLM companies starting to sound like cigarette advertisements.

GPT-5.3 Instant: It's toasted...

Sounds more like the tagline for consumer GPUs these days.

Re: GPT‑5.3 Instant

#163
post #9

I'm a bit confused by this branding (never even noticed that there was a 5.2-Instant), it's not a super fast 1000tok/s Cerebras based model which they have for codex-spark, it's just 5.2 w/out the router / "non-thinking" mode? I feel like openai is going to get right back to where they were pre GPT-5 with a ton of different options and no one knows which model to use for what.

Yeah, for a while ChatGPT Plus has been powered by two series of models under the hood. One series is the Instant series, which is faster and more tuned to ChatGPT, but less accurate. The second series is the Thinking series, which is more accurate and more tuned to professional knowledge work, but slower (because it uses more reasoning tokens). We'd also prefer to have simple experience with just one option, but pic…

Is there a way to get sticky model selection back, or the reason is that it is just too expensive to serve alternative models?

For coding I love codex-5.3-xhigh, but for non-coding prompts I still far prefer o3 even if it's considered a legacy model.

I can imagine that its higher tool use is too expensive to serve, but as a pro user I would love it to come back.

Re: GPT‑5.3 Instant

#164
post #9

I'm a bit confused by this branding (never even noticed that there was a 5.2-Instant), it's not a super fast 1000tok/s Cerebras based model which they have for codex-spark, it's just 5.2 w/out the router / "non-thinking" mode? I feel like openai is going to get right back to where they were pre GPT-5 with a ton of different options and no one knows which model to use for what.

Yeah, for a while ChatGPT Plus has been powered by two series of models under the hood. One series is the Instant series, which is faster and more tuned to ChatGPT, but less accurate. The second series is the Thinking series, which is more accurate and more tuned to professional knowledge work, but slower (because it uses more reasoning tokens). We'd also prefer to have simple experience with just one option, but pic…

Auto will never work, because for the exact same prompt sometimes you want a quick answer because it's not something very important to you, and sometimes you want the answer to be as accurate as possible, even if you have to wait 10 minutes.

In my case it would be more useful to have a slider of how much I'm willing to wait. For example instant, or think up to 1 minute, or think up to 15 minutes.

Re: GPT‑5.3 Instant

#168

I unsubbed because ChatGPT was no longer SOTA. They def got cheap. Reminds me of that graph where late customers are abused. OpenAI is already abusing the late customers. Claude is pretty great.

It's odd because I no longer really like ChatGPT. For chat-type requests, I prefer Claude, or if it's knowledge-intensive then Gemini 3 Pro (which is better for history, old novels, etc). But GPT 5.3 Codex is great. Significantly better than Opus, in the TUI coding agent.

May be trying Codex on your suggestion. I was recently let down by its regular thinking.

Re: GPT‑5.3 Instant

#169

Earlier quoted context omitted.

This applies to any US company. Have we forgotten everything we learned in 2012? If your data is shared with Google, Anthropic, Meta, Amazon, or any of their US competitors, it is within reach of the NSA. Whether or not a company provides support to the DoW is orthogonal to that fact.

Some companies are more evil than others. OpenAI is more evil than Anthropic. Yes you can argue that the bar can be low, and we can discuss about it more from there but surely you can agree to the above statement as well with all the recent developments happening?

I think the distinction is pointless. If OpenAI and Anthropic both subscribe to a high baseline of evil, the difference in their principles or conduct isn't worth applauding. Anthropic's moral high-road means nothing contextualized with their Palantir partnership and preexisting DoD contract.

It all feels reminiscent of Google and Apple fighting over who had the more secure ecosystem, when they had both already assented to hidden surveillance measures. Neither Anthropic nor OpenAI can be trusted for anyone that has even the slightest fear of the US government lashing out against them.

Post reply on HN