Live data from Hacker News

Claude Opus 5

anthropic.com

211–220 of 1001 posts

Re: Claude Opus 5

#211
post #190

Seems really good so far using it in Claude Code CLI - it gave me a new flag when I asked a question: "I don't have a reliable way to read that number, so I'd be guessing if I gave you one — and this is exactly the kind of question where a confident guess is worse than none. What I can tell you is what I actually observe:" I really like this update - gave me a clear sense of the facts but didn't give me a guess just…

So wordy.

Re: Claude Opus 5

#212
post #155

I think the most important thing here is not absolute performance. It's that organizations now have access to a Fable-ish model without Fable's 30-day data retention requirement[0]. > "Consistent with prior Opus models, Opus 5 does not have data retention requirements for general access."[1] On the Opus model release page, the reason why Fable doesn't have an ARC-AGI score is because of that retention policy[2]. 0: h…

Also the cost per task. It appears to be significantly cheaper, cheaper than sonnet!

I can't help but read these comments in the voice of a TV commercial....

Re: Claude Opus 5

#213

What's the point of 150 pages description of a model that's going to be replaced in a couple months? Who even reads this? I know it's cheap to generate text with LLMs, but this is just noise at this point.

It's common practice to release a detailed system card (OP) and a high level summary: https://www.anthropic.com/news/claude-opus-5

It's okay if you're not the target audience for one or the other.

Re: Claude Opus 5

#214
Looks like the API price in tokens is same as previous Opus or Sol, double the price of Terra.

Maybe there’s a better comparison than cost per token, but it will be application-specific.

Re: Claude Opus 5

#215

That's a crazy arc 3 score. What do people think of this? Are models actually developing fluid intelligence like what the creators claim to be measuring? Is it jus do to training for it? Is the benchmark flawed?

Doubleplus benchmaxxed

Re: Claude Opus 5

#216

What's the point of 150 pages description of a model that's going to be replaced in a couple months? Who even reads this? I know it's cheap to generate text with LLMs, but this is just noise at this point.

System Cards aren't really targeted to users, that's what blog posts and docs are for: https://ai.meta.com/tools/system-cards/

Re: Claude Opus 5

#218
post #180

I think the most important thing here is not absolute performance. It's that organizations now have access to a Fable-ish model without Fable's 30-day data retention requirement[0]. > "Consistent with prior Opus models, Opus 5 does not have data retention requirements for general access."[1] On the Opus model release page, the reason why Fable doesn't have an ARC-AGI score is because of that retention policy[2]. 0: h…

So the rumors were right, Opus 5 was indeed being polished up for release. Huge improvements in GDPval-AA v2 too -- great for some of the knowledge work-based agentic workloads I run. Also glad they still kepy Fable 5 on "credits only" access. I think we're going to start seeing model providers gate top-of-the-line models behind pay-as-you-go API rates/credits while subsidizing other models on monthly subscriptions.

Fable 5 is still included in Max subscriptions!

Re: Claude Opus 5

#219
Something fun: on our AWS Bedrock console right now, there's a 'NEW' model called 'anthropic.honey'. Wonder if that's the codename just for this one or in general?

Re: Claude Opus 5

#220
post #97

Earlier quoted context omitted.

I like how they highlighted Opus 5 as the best for “Agentic Coding” even though the number is slightly lower than Fable. Close enough for marketing, I guess!

At half the price and less likely to auto-downgrade, it sounds like a reasonable claim.

given that i couldn't even use fable without it downgrading to Opus, this is just a straight upgrade for me
Post reply on HN