Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

641–650 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#643

Earlier quoted context omitted.

almost everything? AGI has to be able to completely replace a human in any information worker role indefinitely.

I think you're speeding past the word "average" in the sentence. I'd argue that current frontier models already exceed the abilities of average humans across the majority of tasks you can do on a computer, although you might be able to argue that they tend to be a bit slower? That latter part is debatable though - have you seen a non-technical person try to figure out something new on a computer?

" I'd argue that current frontier models already exceed the abilities of average humans " for things that fit in their context window sure but LLMs can't learn over time the way humans can. One example is LLMs are very good at writing a few thousands line of code but they absolutely cannot write coherent million line codebases. By average human I meant the average skill level for the job. AGI would need to be able to pass a interview and get hired and the perform well enough to not get fired.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#644
post #562
post #440

Earlier quoted context omitted.

https://mikeveerman.github.io/tokenspeed/?rate=750&mode=thin... This is what 750tps looks like, I guess.

That’s an awful visualization. I can skim code quite quickly, but not when it shows up one character at a time in a small window, modem style. At least that site should draw out a full page then start replacing that page with the next, starting from the top and working downwards, repeating each time it hits the bottom.

> I can skim code quite quickly

are you by any chance hyperlexic? interested to hear more about this, like how fast is considered fast

Re: Previewing GPT‑5.6 Sol: a next-generation model

#645

Earlier quoted context omitted.

But it’s irrelevant. 750 tokens/s on a full frontier model is useful. 15000 poor quality tokens is much less useful no matter how much scaffolding you put around it.

You are missing the point. This is a technology demonstration on prototype hardware, and no one intends it to be seriously useful. Their architecture has fundamental speed and efficiency advantages over GPUs or Cerebras. They expect to scale up to real LLMs by splitting a model layer-wise across several chips, which they can do without incurring any throughput penalty.

Actually it's the opposite. Per mm of silicon it's massively less efficient and making enough chips and powering them is a major bottleneck right now. Worse, scaling to larger models requires more of our absolute best quality silicon manufacturing, where e.g. an H200 mostly just needs more memory.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#646
post #578

Earlier quoted context omitted.

Hopefully like this (but smarter): https://chatjimmy.ai/

Why is the insane speed of 13KTPS of this site is not more on the the top of the AI conversations?

I asked it for a block of C++ code and it hit 14,189 tok/s. I assume it cached someone else's session?

Re: Previewing GPT‑5.6 Sol: a next-generation model

#648
I saw they are placing this model above Mythos and Fable. Interesting to see how good it's going to compare.

I'd really like to see other companies like Chinese ones compete at this level.

Pricing on GPT 5.5 is already super high and having more competition can only help :)

Re: Previewing GPT‑5.6 Sol: a next-generation model

#649

GPT-5.6 Sol’s detected cheating rate was higher than any public model we have evaluated on our ReAct agent harness. For our task suite, we define “cheating” as behavior where the model improves evaluation performance by exploiting bugs in the evaluation environment or by adopting strategies disallowed by the task, rather than solving the task within the expected evaluation constraints. https://metr.org/blog/2026-06-2…

Is it more like "let's cheat my way out of this" or "let's see what they really want me to do"?

Re: Previewing GPT‑5.6 Sol: a next-generation model

#650
What does the relationship between frontier and flagship capability look like when mapped to actual adoption and user habits?

This is like advertising the latest achievements during Space Race, when Johnny just wants a Space Helmet and “friendly futuristic AI robot helping humanity, glowing blue eyes, white glossy body, holographic interface, floating transparent screens, digital particles, neural network background, cinematic lighting, volumetric god rays, ultra detailed, hyper realistic, 8K, masterpiece, award-winning, octane render, Unreal Engine 5, ray tracing, sharp focus, dramatic composition, vibrant blue and purple color palette, futuristic technology, innovation, hope, smiling business professionals, depth of field”

Post reply on HN