Live data from Hacker News

GPT‑5.3‑Codex‑Spark

openai.com

171–180 of 415 posts

Re: GPT‑5.3‑Codex‑Spark

#171

> Our latest frontier models have shown particular strengths in their ability to do long-running tasks, working autonomously for hours, days or weeks without intervention. I have yet to see this (produce anything actually useful).

Can I just say how funny this metric is?

"Our model is so slow and our tokens/second is so low that these tasks can take hours!" is not the advertising they think it is.

Re: GPT‑5.3‑Codex‑Spark

#172
post #129

Earlier quoted context omitted.

> What am I missing? Largest production capacity maybe? Also, market demand will be so high that every player's chips will be sold out.

> Largest production capacity maybe? Anyone can buy TSMC's output...

Can anyone buy TSMC though?

Re: GPT‑5.3‑Codex‑Spark

#173
post #170

My stupid pelican benchmark proves to be genuinely quite useful here, you get a visual representation of the quality difference between GPT-5.3-Codex-Spark and full GPT-5.3-Codex: https://simonwillison.net/2026/Feb/12/codex-spark/

These are the ones I look for every time a new model is released. Incorporates so many things into one single benchmark.

Also your blog is tops. Keep it up, love the work.

Re: GPT‑5.3‑Codex‑Spark

#174
post #65

Earlier quoted context omitted.

I love the probabilistic nature of this. Presentations could be anywhere from extremely impressive to hilariously embarrassing.

It would be so cool if it generated live in the presentation and adjusted live as you spoke, so you’d have to react to whatever popped on screen!

[deleted]

Re: GPT‑5.3‑Codex‑Spark

#177
post #66

Continue to believe that Cerebras is one of the most underrated companies of our time. It's a dinner-plate sized chip. It actually works. It's actually much faster than anything else for real workloads. Amazing

Yet investors keep backing NVIDIA.

At this point Tech investment and analysis is so divorced from any kind of reality that it's more akin to lemmings on the cliff than careful analysis of fundamentals

Re: GPT‑5.3‑Codex‑Spark

#179
post #65

Earlier quoted context omitted.

I love the probabilistic nature of this. Presentations could be anywhere from extremely impressive to hilariously embarrassing.

It would be so cool if it generated live in the presentation and adjusted live as you spoke, so you’d have to react to whatever popped on screen!

There was a pre-LLM version of this called "battledecks" or "PowerPoint Karaoke"[0] where a presenter is given a deck of slides they've never seen and have to present on it. With a group of good public speakers it can be loads of fun (and really impressive the degree that some people can pull it off!)

0. https://en.wikipedia.org/wiki/PowerPoint_karaoke

Re: GPT‑5.3‑Codex‑Spark

#180
post #163

Earlier quoted context omitted.

Just pick some reasonable values. Also, keep in mind that this hardware must still be useful 3 years from now. What’s going to happen to cerebras in 3 years? What about nvidia? Which one is a safer bet? On the other hand, competition is good - nvidia can’t have the whole pie forever.

> Just pick some reasonable values. And that's the point - what's "reasonable" depends on the hardware and is far from fixed. Some users here are saying that this model is "blazing fast" but a bit weaker than expected, and one might've guessed as much. > On the other hand, competition is good - nvidia can’t have the whole pie forever. Sure, but arguably the closest thing to competition for nVidia is TPUs and future c…

AMD
Post reply on HN