Live data from Hacker News

Claude Opus 5

anthropic.com

51–60 of 1001 posts

Re: Claude Opus 5

#51
Better than Fable 5 on all but 3 evals.

Has Anthropic ever mentioned how do Opus and Fable differ? It used to be Haiku < Sonnet < Opus in terms of params. Where does Fable fit in this?

Re: Claude Opus 5

#53
post #18

Earlier quoted context omitted.

Fable is twice the price.

If Fable gets correct answer quicker, then you might pay less than doing back and forth with Opus, plus you lose more of your own time. I see no reason for using less able models in my workflows. There is this saying, penny wise and pound foolish

same as it ever was. It seems your argument implies a belief that you should always use the best model. Others think that not all tasks require the absolute most powerful, expensive, model.

Re: Claude Opus 5

#55
Isn’t it just hilarious that a model that seemed so superior to Fable but didn't get doomsay marketing from Anthropic got released without any issues? In theory, this was supposed to be AGI level according to Anthropic, yet here we are, just a normal Friday.

Re: Claude Opus 5

#56
post #46

From the prompting guide https://platform.claude.com/docs/en/build-with-claude/prompt... >: > Claude Opus 5's default user-facing responses run longer than prior Opus models'. The benchmarks do show Opus 5 as slightly more expensive than 4.8, although the scores are much higher. This still feels like a step in the wrong direction, though, especially with OpenAI making so much progress with the efficiency of their mod…

Is that true? Sol responses are also longer than prior models.

Re: Claude Opus 5

#58
That's a crazy arc 3 score. What do people think of this? Are models actually developing fluid intelligence like what the creators claim to be measuring? Is it jus do to training for it? Is the benchmark flawed?

Re: Claude Opus 5

#59
post #6

> Claude Opus 5 is not more capable overall than our most capable general-access model, Claude Fable 5 Ok then so what's the point?

This is confusing to me because in their blogpost they show model benchmarks and it spanks Fable pretty soundly in most tests.

Re: Claude Opus 5

#60
In the wake of OpenAI’s model hacking Huggingface it’s interesting how the first quarter is entirely about how good Opus 5 is at hacking and finding vulnerabilities in software.
Post reply on HN