Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

691–700 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#691

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

This is something Xioami already did with MiMo-2.5-Pro a month ago, and at a higher speed (1,000 t/s).

750 tps at GPT-5.5-Pro prices would be ruinous!

Re: Previewing GPT‑5.6 Sol: a next-generation model

#692
post #578

Earlier quoted context omitted.

Hopefully like this (but smarter): https://chatjimmy.ai/

Why is the insane speed of 13KTPS of this site is not more on the the top of the AI conversations?

Because I just tested it and it took 3-4 clarifications before it actually gave a correct response vs gemini/google search. It's not great, but good.

I'd rather wait 3x as long.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#694

Earlier quoted context omitted.

I get that, but not at 15k tokens/s.

But it’s irrelevant. 750 tokens/s on a full frontier model is useful. 15000 poor quality tokens is much less useful no matter how much scaffolding you put around it.

I’ve been using 1,000 t/s on a near frontier model for a month now. It’s very useful for agentic coding.

It does require new approaches for me personally since I get a lot less time to think or read its output.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#695

Earlier quoted context omitted.

Try gpt-5.3-codex-spark - it's 1000 TPS and from my experience more capable than 5.4 mini. If you have a subscription it's a different pool of usage.

Used it, very fast but tiny context window and doesn't have good reasoning. (good for quick simple code changes)

MIMO 2.5 Pro ultraspeed has a 1M window. 1,000 tok/sec is great for planning since you can have a rapid conversation with a lot of turns.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#696

Easily the most interesting part of this announcement is buried in the second to last paragraph: "We're also launching GPT‑5.6 Sol on Cerebras at up to 750 tokens per second in July, bringing frontier intelligence to customers at unprecedented speed. Access will initially be limited to select customers as we expand capacity." 750 tokens/s on a frontier model is going to be extremely interesting. I doubt this new vers…

I saw videos of coding with Mimo-V2.5-Pro UltraSpeed, which is advertised at 1,000 tokens/s, which is very impressive.: https://www.bilibili.com/video/BV1fME16uEW7 If the time-to-first-token latency also greatly improved, this could be very useful for end-to-end in controls, like autonomous driving for example.

It’s awesome, particularly since it’s at DeepSeek tier prices (3X of DS-V4-Pro). At 1,000 tok/sec though you can really rip through tokens. (About $9 an hour if you manage to run the output nonstop.)

It tends to cost more than DS since it doesn’t seem to have as many input cache hits.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#698

Here is a trend I'm noticing: - GPT-5 mini costs $0.25/$2 and will be discontinued in December. - GPT-5.4 mini costs $0.75/$4.5 and is supposed to be the replacement. - GPT-5.4 nano costs $0.2/$1.25 and, while it ranks better in benchmarks than GPT-5 mini, it's not even close when you test it in real scenarios. So you're left being forced to go to GPT 5.4 mini if you use 5 mini today. The same thing is happening here…

Why not self host or go to openrouter if you don't need SOTA frontier?

Re: Previewing GPT‑5.6 Sol: a next-generation model

#699

Earlier quoted context omitted.

Other people's computers famously can't be taken away.

You're right, there is an all powerful wizard who can take away all the world's computers. You got me. There's a reason Yudkowsky only thinks AI can be stopped by literal missile strikes against data centers.

[deleted]

Re: Previewing GPT‑5.6 Sol: a next-generation model

#700
post #497

Earlier quoted context omitted.

This is genuinely confusing to my senses. The future is going to be so strange/neat/me unemployed.

The future is totally illegible to me. I love these AI models, but I feel like I'm going to be jobless within 10 years. Anomie is at an all time high right now.

10 years? An optimist, I see.
Post reply on HN