I was literally posting about this earlier this morning[1], but all data indicates that we'll have models equivalent to Opus 4.6 / GPT 5.3 with a truly instant (ie > 10k t/s) response time by 2028. Small models are getting better faster, and their ability to be baked into silicon in a power and speed efficient way is likely going to completely disrupt things.
GPT‑5.3 Instant
161–170 of 330 posts
Re: GPT‑5.3 Instant
#162Re: GPT‑5.3 Instant
#163I'm a bit confused by this branding (never even noticed that there was a 5.2-Instant), it's not a super fast 1000tok/s Cerebras based model which they have for codex-spark, it's just 5.2 w/out the router / "non-thinking" mode? I feel like openai is going to get right back to where they were pre GPT-5 with a ton of different options and no one knows which model to use for what.
Yeah, for a while ChatGPT Plus has been powered by two series of models under the hood. One series is the Instant series, which is faster and more tuned to ChatGPT, but less accurate. The second series is the Thinking series, which is more accurate and more tuned to professional knowledge work, but slower (because it uses more reasoning tokens). We'd also prefer to have simple experience with just one option, but pic…
For coding I love codex-5.3-xhigh, but for non-coding prompts I still far prefer o3 even if it's considered a legacy model.
I can imagine that its higher tool use is too expensive to serve, but as a pro user I would love it to come back.
Re: GPT‑5.3 Instant
#164I'm a bit confused by this branding (never even noticed that there was a 5.2-Instant), it's not a super fast 1000tok/s Cerebras based model which they have for codex-spark, it's just 5.2 w/out the router / "non-thinking" mode? I feel like openai is going to get right back to where they were pre GPT-5 with a ton of different options and no one knows which model to use for what.
Yeah, for a while ChatGPT Plus has been powered by two series of models under the hood. One series is the Instant series, which is faster and more tuned to ChatGPT, but less accurate. The second series is the Thinking series, which is more accurate and more tuned to professional knowledge work, but slower (because it uses more reasoning tokens). We'd also prefer to have simple experience with just one option, but pic…
In my case it would be more useful to have a slider of how much I'm willing to wait. For example instant, or think up to 1 minute, or think up to 15 minutes.
Re: GPT‑5.3 Instant
#165Re: GPT‑5.3 Instant
#166Re: GPT‑5.3 Instant
#167I kind of chuckled when I read the headline "GPT‑5.3 Instant: Smoother, more ..." LLM companies starting to sound like cigarette advertisements.
Re: GPT‑5.3 Instant
#168I unsubbed because ChatGPT was no longer SOTA. They def got cheap. Reminds me of that graph where late customers are abused. OpenAI is already abusing the late customers. Claude is pretty great.
It's odd because I no longer really like ChatGPT. For chat-type requests, I prefer Claude, or if it's knowledge-intensive then Gemini 3 Pro (which is better for history, old novels, etc). But GPT 5.3 Codex is great. Significantly better than Opus, in the TUI coding agent.
Re: GPT‑5.3 Instant
#169Earlier quoted context omitted.
This applies to any US company. Have we forgotten everything we learned in 2012? If your data is shared with Google, Anthropic, Meta, Amazon, or any of their US competitors, it is within reach of the NSA. Whether or not a company provides support to the DoW is orthogonal to that fact.
Some companies are more evil than others. OpenAI is more evil than Anthropic. Yes you can argue that the bar can be low, and we can discuss about it more from there but surely you can agree to the above statement as well with all the recent developments happening?
It all feels reminiscent of Google and Apple fighting over who had the more secure ecosystem, when they had both already assented to hidden surveillance measures. Neither Anthropic nor OpenAI can be trusted for anyone that has even the slightest fear of the US government lashing out against them.
Re: GPT‑5.3 Instant
#170I kind of chuckled when I read the headline "GPT‑5.3 Instant: Smoother, more ..." LLM companies starting to sound like cigarette advertisements.