Live data from Hacker News

Claude Fable 5.1 and Claude Mythos 5.1

anthropic.com

521–530 of 1001 posts

Re: Claude Fable 5.1 and Claude Mythos 5.1

#523
post #467

Earlier quoted context omitted.

While I can't speak for everyone in academia, I personally don't feel comfortable in putting my research questions and outputs to a private website, before the idea is at least arxived. Especially as all the Fable/Mythos prompts are said to be human reviewed. So I believe that, at least in the short run, we might be seeing breakthroughs in hard open problems or in low hanging problems which are not that interesting t…

Isn't it showing a problem with an academia? "I don't want to live in a world where someone else makes the world a better place than we do."

This feels like an unwarranted strawman. There are plenty of reasons for researchers to share openly at times and plenty of times it makes sense to wait until the meal is ready to serve before publishing.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#524

(I work at Anthropic) Beyond all the benchmarks, I think Fable 5.1 is a big improvement in writing style. It sounds a lot less stereotypically like other Claude models, has (imho) a much more natural style, and responds to my style instructions more reliably. More work to be done (and we will!) but reading better prose makes me so much happier. Another point I expect not to get much attention until it all happens at…

Could you share what you use internally to make Fable not sound like a word salad generator?

Re: Claude Fable 5.1 and Claude Mythos 5.1

#525
> These patterns invalidate every later thinking block:

• [...] Rebuilding the top-level system prompt or tools array between requests in the same conversation.

Many people unknowingly do this (at a high cost to them because of the cache busts), this change will finally force them to stop.

Especially if you're generating your system prompt via a template that can change mid conversation, it's so easy to fall into this trap.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#526

Somewhat ironically, Fable 5.1 was flagged by the biology safeguards after I asked it to have a dig around the Fable 5.1 system card :)

Same for me. Every single time I tried it got flagged. I think this will be my litnus test for if the safeguards are good enough for benign requests.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#527
post #435

Anyone ever seen the SouthPark episode making fun of Game of Thrones: A Song of Ass and Fire? Anthropic's announcements reminds me of "The Dragons Are Coming" running joke. What they have done: * Nerfed Fable, as many of noted it's useless * Leverage Mythos as a marketing strategy, claiming its too good to release * Removed thought traces, one of the only useful things to make sure your prompts are working correctly…

and yet we still have people saying the rate of change is increasing my view is we had a leap over the last fe years and it's tapering off. this is fine, but for the IPOs

We had a leap because of the introduction and refinement of agents - the rest has been minor

Re: Claude Fable 5.1 and Claude Mythos 5.1

#528
post #479

My main gripe with LLMs is the cringe AI phrasings that they use in UI elements. Pompous things like "Your keys, supercharged" or weird yoda-speak stuff like "searches the app remembers" instead of just naming the thing "Learned searches" .. you know, proper GUI copy like it was done for the past decades. I jumped when I saw a mention about "writing style improvements" so I gave it a try on a recent feature in rcmd […

> Pompous things like "Your keys, supercharged" or weird yoda-speak stuff like "searches the app remembers"...

It's copywriting. They fed these models the internet, which is loaded with it.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#529
post #246

Earlier quoted context omitted.

I've developed a habit of adding into my prompts "please keep your response concise and succinct" or "I'm trying to cram, please only provide the minimum level of technical detail necessary to understand this topic" I find it helps immensely but it'd be nice if I didn't have to do that.

why so many people add 'please' when asking machine to do something? Was there actually research that when you SCREAM or curse it follows your instructions better? P.S. Although my wife insists that I should stay polite in case AI overlords remember how I treat them ...

I think about removing please/thanks, but then I accidentally add them back in during some edit/rewrite of the prompt... It's just how I'm used to asking for things

Re: Claude Fable 5.1 and Claude Mythos 5.1

#530
I think what’s the industry is interested to see now isn’t “the best and latest super intelligent frontier model ever!!”, but rather the ability to run good enough models locally or better, on consumer or laptop grade specs. So I am not that impressed, plus haven’t used Claude for a while nor planning to, their models are useless with their “safe guards”.
Post reply on HN