Live data from Hacker News

Previewing GPT‑5.6 Sol: a next-generation model

openai.com

791–797 of 797 posts

Re: Previewing GPT‑5.6 Sol: a next-generation model

#791

Earlier quoted context omitted.

It seems to rely on a definition of "understand" which is more about spirituality than actual observable evidence "Understanding" has enough philosophical leeway in its use to allow at least the possibility of sentience as a prerequisite. This is where the discussion about LLM capabilities becomes genuinely difficult, and dismissing that difficulty as "word games" or "spirituality vs evidence" is not helpful.

Considering that "sentience" has enough "philosophical leeway" that it's just as reasonable to assert that LLMs are sentient (and at extremes, that they have been sentient for years) -- especially if we are, as you suggest, supposed to include any philosophically possible definition -- I don't think that's a meaningful rebuttal. If no one can agree on whether it's sentient, it's bad faith to choose a fringe definitio…

it's bad faith to choose a fringe definition that hands off its definition to such a nebulous term.

I do not think it is at all unreasonable or "fringe" to regard understanding as involving intentionality: ie a directedness of thought toward the object-relations being "grasped". That may not be the only possible conception of understanding but it is a mainstream philosophical idea.

In fact, I'd argue that statements about what "is" and "is not" sentient relies on even more spirituality and word games for anything that isn't a terran tetrapod.

Then you seem to be confusing "hard to understand" with "meaningless".

you're just expressing dogma rather than having a discussion.

Anything else is bad faith, or assuming bad faith on the part of the participants

Have a think about that (repeated) tone before responding.

Fwiw I am a long-time believer in consciousness being fully realisable in machines; I think the jury is still out on LLMs.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#793

Earlier quoted context omitted.

AGI should be able to do every job a human can do using a computer at least as well as the average human.

That's already been true for a while, you're overestimating the average human. They just have different failure modes.

It isn't even close to true. The biggest problem is that humans performance improves over time.

https://www.linkedin.com/pulse/announcing-aa-briefcase-bench...

AA-Briefcase is a new benchmark for testing models on realistic knowledge work tasks in complex projects built by industry experts. Models are evaluated on multi-week knowledge work projects, each with many linked tasks and thousands of input source files. AA-Briefcase combines rubric and pairwise grading to evaluate verifiable task success, analytical quality, and presentation quality, giving a holistic view of overall agentic capability in knowledge work.

Tasks with many messy input files, conflicting information, and complex deliverables remain difficult for all models. Under a strict all-or-nothing grading scheme per task, Claude Fable 5 leads overall, but achieves a perfect task score on only 3% of tasks. On 31 of 91 tasks, no model scores above 50%.

Re: Previewing GPT‑5.6 Sol: a next-generation model

#794

Earlier quoted context omitted.

The "it's not X it's Y" where Y qnd X are the same indicates a lack of understanding.

Consider the number of humans I've seen make statements that fit that description (about AI, no less!), I don't think that's a strong argument against it.

Would you say those humans understand what they’re talking about?

Re: Previewing GPT‑5.6 Sol: a next-generation model

#795

Earlier quoted context omitted.

Consider the number of humans I've seen make statements that fit that description (about AI, no less!), I don't think that's a strong argument against it.

Would you say those humans understand what they’re talking about?

Touché

Re: Previewing GPT‑5.6 Sol: a next-generation model

#796

Earlier quoted context omitted.

What How? Which model is behind it?

It’s pure silicon. Llama3.

What does that even mean? There must be a software stack somewhere feeding the inputs to the model. Do you mean the weights are baked in sillicon?

Re: Previewing GPT‑5.6 Sol: a next-generation model

#797

>> We are taking this short-term step because we believe it is the strongest path... >>During this preview, we will continue testing and coordinating closely with partners as we work toward broader availability. Instead of generating negative publicity, can't they just wait for the preview period to get over?. What does openAI announce when they know others can't access it?. Curious question - what do they gain from…

probably so people don't think they lost the race
Post reply on HN