Live data from Hacker News

Claude Fable 5

anthropic.com

991–1000 of 1001 posts

Re: Claude Fable 5

#991
post #37

An 11% jump over opus 4.8 and a 22% jump over gpt 5.5 on Agentic Coding Benchmarks is certainly impressive. Obviously still need to verify it for myself to see if it's truely a leap. But am I the only one wondering, "What can I do today that I couldnt do yesterday?" Previously I would think "Oh I wonder if I can finally get it to do X now?" However now I feel like yesterdays models were more that capable to handle ne…

[flagged]

Re: Claude Fable 5

#996
In Indian arranged marriages, families sometimes meet for an hour, everyone is on their best engineered behavior, and suddenly people are ready to make lifelong commitments based on smiles, tea, and a few photos. My mom would come home after one afternoon saying, “What wonderful people!”

That is where we are with every new model release.

The people yelling “Fable will take your job” are still at the first meeting. I have used it for 16 hours, and spent one of those hours fighting it over git. It rebased wrong, stashed changes, forgot it had stashed them, merged a stash it claimed did not exist, then reset to HEAD. By the end, I had lost the code we had just worked on.

Maybe wait until first kid before minting trillionaires :)

Re: Claude Fable 5

#997

What a (genuinely) surprising choice: >"We’ve therefore launched the model with safeguards that mean queries on some topics will instead receive a response from our next-most-capable model, Claude Opus 4.8" That's a very surprising solution. Imagine being asked to do something you feel you shouldn't do, and rather than refusing, you say, "Yeah I could do that but given that I don't want you to succeed at this task, I…

It makes more sense if Anthropic is assuming that most flagged conversations are false positives (but it wants to keep Mythos away from the true positives).
Post reply on HN