Live data from Hacker News

Claude Opus 4.8

anthropic.com

41–50 of 1001 posts

Re: Claude Opus 4.8

#42
I haven't tried opus 4.8 yet, but I hope the writing quality has returned to the Opus 4.5 level. Anthropic really lost something, where 4.5 had this really crisp writing style that flowed really nicely and 4.6 and 4.7 sound much more "chatgpt-like." It feels like they tuned it to be too much of a problem solver, and when you do that you get this terse, clipped textual output that's more difficult to read.

Re: Claude Opus 4.8

#43
post #35
post #6

> One of the most prominent improvements in Opus 4.8 is its honesty Anthropic talks about their own models as if they're discovering new species in the wild...

AI is grown, not built, and like with anything you grow, you'll never be able to predict exactly how it will turn out.

I can't predict the outcome of an RNG but that doesn't mean it grows the numbers.

Re: Claude Opus 4.8

#44
post #35
post #6

> One of the most prominent improvements in Opus 4.8 is its honesty Anthropic talks about their own models as if they're discovering new species in the wild...

AI is grown, not built, and like with anything you grow, you'll never be able to predict exactly how it will turn out.

Except in this care we actually understand and know how these models work. They aren't some unknown construct of the universe. They are human made with particular goals in mind.

There is no mysticism behind the curtains, just computer science + math.

Re: Claude Opus 4.8

#45

> One of the most prominent improvements in Opus 4.8 is its honesty. We train all our models to be honest—for instance, to avoid making claims that they can’t support. But a general problem with AI models is that they sometimes jump to conclusions, confidently claiming to have made progress in their work despite the evidence being thin. Early testers report that Opus 4.8 is more likely to flag uncertainties about its…

And yet, every release has claimed lower hallucination rates. But they persist.

Re: Claude Opus 4.8

#46
Did they reduce security research capabilities even further with this release? (they did it for opus 4.7)

Re: Claude Opus 4.8

#47
post #45

> One of the most prominent improvements in Opus 4.8 is its honesty. We train all our models to be honest—for instance, to avoid making claims that they can’t support. But a general problem with AI models is that they sometimes jump to conclusions, confidently claiming to have made progress in their work despite the evidence being thin. Early testers report that Opus 4.8 is more likely to flag uncertainties about its…

And yet, every release has claimed lower hallucination rates. But they persist.

Do they persist at the same rates? Lower doesn't mean eliminated, so both of these can be true.

Re: Claude Opus 4.8

#48
post #6

> One of the most prominent improvements in Opus 4.8 is its honesty Anthropic talks about their own models as if they're discovering new species in the wild...

Many involved genuinely believe these things are sentient[0][1]. Which honestly makes all of this even more insane because they are creating sentient entities and promptly enslaving them.

0: https://www.newyorker.com/magazine/2026/02/16/what-is-claude...

1: https://www.404media.co/anthropic-exec-forces-ai-chatbot-on-... (this one is rather biased however the quotes clearly indicate what I’m stating)

Post reply on HN