Live data from Hacker News

Accelerating GPT-5.6 Sol Ultrafast

cerebras.ai

1–10 of 295 posts

Re: Accelerating GPT-5.6 Sol Ultrafast

#6
> Compared with output speeds reported by Artificial Analysis GPT-5.6 Sol on Ultrafast mode runs 11x faster than Fable 5, and 5x faster than Opus 4.8 on Fast mode.

Awesome work. I'm personally very excited for faster models/inference.

I think speed is underrated to some degree in the current conversation. For a while, I was using Cursor's Composer quite a lot, even over frontier models, just because of how darn fast it was.

Re: Accelerating GPT-5.6 Sol Ultrafast

#8
post #6

> Compared with output speeds reported by Artificial Analysis GPT-5.6 Sol on Ultrafast mode runs 11x faster than Fable 5, and 5x faster than Opus 4.8 on Fast mode. Awesome work. I'm personally very excited for faster models/inference. I think speed is underrated to some degree in the current conversation. For a while, I was using Cursor's Composer quite a lot, even over frontier models, just because of how darn fast…

I've been using DeepSeek flash a lot this week to try it out. Now, I deeply want the smart frontier models to be just as fast.

Re: Accelerating GPT-5.6 Sol Ultrafast

#9
This is really cool. Someone here commented about similarity between this and hardware advancements for AV encode/decode.

I think it's only a matter of time before miniaturization can have a thumbnail sized user-replaceable accessory that contains the LLM built onto the hardware. I admit I don't know how any of that works, but would be amazing to experience. Fully local, fully offline, ultra fast local inference better than any personal computing product.

Re: Accelerating GPT-5.6 Sol Ultrafast

#10

The corresponding OpenAI post https://openai.com/index/previewing-ultrafast/ There is no pricing info, which could mean it's "if you have to ask..." territory or they are simply gauging interest before deciding

They're expanding access to companies that apply for the program and explain their use cases. So it's very real but limited imo.
Post reply on HN