Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

451–460 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#451

Earlier quoted context omitted.

I think you missed the point and don't understand / aren't considerate of SLM utility.

But I’m not missing the point. If you can run one frontier model at 750t/s, then you can probably run many many instances of an SLM in parallel at a rate that exceeds 15k/s. That’s kinda the point of the flash or ultrafast variants. And they’re on something much more modern than llama3.1.

Yes, you are missing the point. 1) It's a demo. [0] 2) It hasn't been updated for 4+ months.

You don't need LLMs for everything. That is 100% the point. You can burn down the world with all of your frontier LLMs that are being used for simple queries OR we can do something faster and more efficient like this. Just because you can run a SotA model at "fast" speeds, again, severely misses the point.

And no, you can't run anything from Anthropic or OAI on-prem, so until you can there's really no comparison. If people want to continue down the path of gate-kept models with no other options then we'll all follow you off the cliff.

[0] https://taalas.com/products/

Re: Previewing GPT‑5.6 Sol: a next-generation model

#452

Earlier quoted context omitted.

Yes. The difference is obviously that full, fat Linux runs on a superset of anything a layperson would call a computer, and can be built from source on roughly the same set of hardware. Running the full, fat Deepseek (as in the 1.6T model, unquantized) is too big to run on anything a layperson would call a computer, and being able to actually build it is even harder.

It's famously difficult to find people willing to rent you time on big computers over the internet.

Other people's computers famously can't be taken away.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#453
post #391

Earlier quoted context omitted.

When attacking archetypes of people, there is some responsibility to make clear who you’re attacking and why, even to someone who’s not being hyper-open-minded. At least if you want them to learn from you: which may or may not be your goal. When you attack/signal you’re on the offensive, it is foolish to believe that they won’t knee-jerk attack back and become closed minded at least a little. Regardless, the “misinte…

You used a lot of words to defend a strawman argument

As you ironically strawman me. Your hypocrisy knows no bounds!

Re: Previewing GPT‑5.6 Sol: a next-generation model

#454

Earlier quoted context omitted.

When attacking archetypes of people, there is some responsibility to make clear who you’re attacking and why, even to someone who’s not being hyper-open-minded. At least if you want them to learn from you: which may or may not be your goal. When you attack/signal you’re on the offensive, it is foolish to believe that they won’t knee-jerk attack back and become closed minded at least a little. Regardless, the “misinte…

When you read someone's comment there is some responsibility to read the words they wrote and not attempt to attack them for an argument no reasonable person would extract from those words.

[deleted]

Re: Previewing GPT‑5.6 Sol: a next-generation model

#455
Another year, and OpenAI comes up with yet another naming scheme for their models. First it was integers (GPT2, GPT3). Then they added friendly names (remember Ada, Babbage, Curie, Davinci?), but decided against it. Instead we got dot integers (GPT3.5), then then letter-number modifiers (o1), plus word modifiers like o1-pro, o3-mini, or -mini-high, or codex, codex-max, Pro, etc.

Now they've got friendly cosmic names. And this time they want us to believe that this time they're gonna stick to a naming convention? I'll believe it when they do 3 releases in a row without inventing a new naming scheme.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#456

Earlier quoted context omitted.

I get that, but not at 15k tokens/s.

But it’s irrelevant. 750 tokens/s on a full frontier model is useful. 15000 poor quality tokens is much less useful no matter how much scaffolding you put around it.

You are missing the point. This is a technology demonstration on prototype hardware, and no one intends it to be seriously useful.

Their architecture has fundamental speed and efficiency advantages over GPUs or Cerebras. They expect to scale up to real LLMs by splitting a model layer-wise across several chips, which they can do without incurring any throughput penalty.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#457
post #406

A question I always have is, how to the AI labs safeguard the leak of their model? Training a cutting edge model basically cost a minimum of hundreds of millions of dollars. And its all contained within a file. Okay, that file might be 500GB large, but its still just one blob that is worth almost a billion dollars. And they need to train new models every few weeks, have lots of people with access to it to debug it, r…

Employees naturally jump from one company to another, and they know the secret sauce.

The difference is in the dataset mostly and to extract this dataset, competitors use a process called distillation (= extract data through actual queries) from the other models.

This yield to "funny" cases as well, like Gemini who claims "I am ChatGPT" occasionally, or ChatGPT calling itself Claude, etc.

https://note.com/maudi/n/n821a6308437b?hl=en

They all copy on each other.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#458

Earlier quoted context omitted.

When attacking archetypes of people, there is some responsibility to make clear who you’re attacking and why, even to someone who’s not being hyper-open-minded. At least if you want them to learn from you: which may or may not be your goal. When you attack/signal you’re on the offensive, it is foolish to believe that they won’t knee-jerk attack back and become closed minded at least a little. Regardless, the “misinte…

When you read someone's comment there is some responsibility to read the words they wrote and not attempt to attack them for an argument no reasonable person would extract from those words.

Reasonable people could interpret the original comment in many other ways than was probably intended.

I like when people are open minded to people who are closed minded/attacking them. It’s an admirable and difficult trait to attain. But to expect that from others is foolish. Most people can’t stay objective/curious after being punched in the face.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#459

Earlier quoted context omitted.

When attacking archetypes of people, there is some responsibility to make clear who you’re attacking and why, even to someone who’s not being hyper-open-minded. At least if you want them to learn from you: which may or may not be your goal. When you attack/signal you’re on the offensive, it is foolish to believe that they won’t knee-jerk attack back and become closed minded at least a little. Regardless, the “misinte…

Angry girlfriend SMS essay

You have a lot of angry girls texting you?
Post reply on HN