Live data from Hacker News

Claude Fable 5.1 and Claude Mythos 5.1

anthropic.com

561–570 of 1001 posts

Re: Claude Fable 5.1 and Claude Mythos 5.1

#561

I'm afraid watermarking could restrict applications where LLMs can be safely used to assist with writing. If I write something myself and use an LLM to proofread it, without watermarking I can confidently say that corrections done by LLMs are small and insignificant enough to claim that the text is still authored by me, not by the model. With watermarking, however, I will never be sure if the result will not be flagg…

I think if you use an LLM just to proofread then it'll not be able to insert a strong enough watermark.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#562
post #479

My main gripe with LLMs is the cringe AI phrasings that they use in UI elements. Pompous things like "Your keys, supercharged" or weird yoda-speak stuff like "searches the app remembers" instead of just naming the thing "Learned searches" .. you know, proper GUI copy like it was done for the past decades. I jumped when I saw a mention about "writing style improvements" so I gave it a try on a recent feature in rcmd […

> Where are all these verbal tics coming from and why is it so hard to get rid of them? It’s a side effect of post-training for effectiveness and efficiency at technical tasks. Over time the models learn to pack as much information as possible into their available context window, because that’s one way to increase the effective intelligence. Humans do this too with industry jargon, dense tech-talk, etc. We have a lim…

Yeah that was what I was most worried about when I read the top comment here. I found the use of language a feature not a bug. I don’t care how good it reads. If I can communicate with it concicely it’s enough to get my work done. I don’t hate the language for copy either, but yeah different users, different problems.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#563
post #479

My main gripe with LLMs is the cringe AI phrasings that they use in UI elements. Pompous things like "Your keys, supercharged" or weird yoda-speak stuff like "searches the app remembers" instead of just naming the thing "Learned searches" .. you know, proper GUI copy like it was done for the past decades. I jumped when I saw a mention about "writing style improvements" so I gave it a try on a recent feature in rcmd […

> Where are all these verbal tics coming from and why is it so hard to get rid of them? It’s a side effect of post-training for effectiveness and efficiency at technical tasks. Over time the models learn to pack as much information as possible into their available context window, because that’s one way to increase the effective intelligence. Humans do this too with industry jargon, dense tech-talk, etc. We have a lim…

Makes sense. Then maybe we would need a separate simpler LLM trained on UI copy and good UX to decide this stuff and let frontier models do the implementation.

But who has both the compute power and the motivation to do such a thing?

I guess I'll just continue rewriting the UI one word at a time for the time being.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#564
post #86

(I work at Anthropic) Beyond all the benchmarks, I think Fable 5.1 is a big improvement in writing style. It sounds a lot less stereotypically like other Claude models, has (imho) a much more natural style, and responds to my style instructions more reliably. More work to be done (and we will!) but reading better prose makes me so much happier. Another point I expect not to get much attention until it all happens at…

As a fervent Claude Code user who made the switch to GPT 5.6 Sol over Opus 5 over hard-to-read prose this makes me happy. I love your product but the current models are very hard to work with if you need to do a lot of context switching. Brevity is key.

It's not really brevity - it's the constant writing tropes. It's like they ready a book on advertising copy and that's the only way they can write. Very tedious. Is Sol much better? I might have to switch to that too!

Re: Claude Fable 5.1 and Claude Mythos 5.1

#565
post #365

What I don't see in the comments: "I had a specific problem I couldn't solve with the previous version of this LLM. But the improvements in this version unlocked the solution for me." What I do see in the comments: subjective improvement in text generation, possibly lower cost, some optimism about code generation, but some skepticism too. I use coding agents. To me they are very useful. But what I spend on them isn't…

We rarely upgrade our phones or MacBooks because the newer version can do something the previous one literally couldn’t. Often it’s the efficiency, speed, battery life, etc, combined, that lets us push the hardware further.

I get your point, but we can only have groundbreaking leaps once in a blue moon. That doesn’t mean incremental improvements aren’t useful.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#566
post #325

Earlier quoted context omitted.

As someone working in science, this belief confuses me. How (by what means) do you think Fable 5.1 will be able to make further progress in scientific domains? The problem with science is that there is no agentic harness. The agent can't test things. At best it can hallucinate something and ask if that hallucination "makes sense", but this doesn't work in science.

Great news, then! TFA: "Last week, we previewed the Model Hardware Standard, which allows Claude to directly and safely operate laboratory equipment."

How much lab equipment is automatable though? There's definitely some in biology, but if you're doing fundamental research it's 99% stuff you are building yourself with your own hands. Robotics is a long way from being able to do any of that.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#567
post #467

Earlier quoted context omitted.

While I can't speak for everyone in academia, I personally don't feel comfortable in putting my research questions and outputs to a private website, before the idea is at least arxived. Especially as all the Fable/Mythos prompts are said to be human reviewed. So I believe that, at least in the short run, we might be seeing breakthroughs in hard open problems or in low hanging problems which are not that interesting t…

Isn't it showing a problem with an academia? "I don't want to live in a world where someone else makes the world a better place than we do."

grants are competitive

Re: Claude Fable 5.1 and Claude Mythos 5.1

#569
post #46

Looks like all three breaking changes are patches for inadvertent chain of thought disclosure. Someone found out (don't have the tweet handy) that if you created a bogus "think_deeply" tool and then forced the model to use it, it would output what is believed to be its raw thinking there - I believe the first breaking change stops this. The second two are aimed at people getting Haiku to repeat thinking blocks from o…

These draconian "Preserved Thinking" measures they're taking are going to be an absolute pain in the ass. This alone is enough for me to move our API use off their platform entirely. It's a HUGE breaking change that they're trying to dampen by having it not affecting current customers until "in the future", see: https://platform.claude.com/docs/en/build-with-claude/preser... You're no longer allowed to edit the conte…

> No more editing the system prompt as the conversation progresses, no more dynamic loading of custom tool calling formats.

Hm, aiui you can support both of these via mid-conversation system turns https://platform.claude.com/docs/en/build-with-claude/mid-co... - and in general you'd want to to preserve the cache and recency of the instruction anyways rather than frankensteining an off-distribution transcript. Not sure though.

Re: Claude Fable 5.1 and Claude Mythos 5.1

#570
post #546

AI is really not "just software" anymore. It is able to discover facts and advance science. Hard to disagree that we're near or at the point where Artificial Intelligence has expanded reality into 4 quadrants: objects that are not alive: dust, rocks, water, wood, hats, lego, aluminum, etc. objects that are alive but not intelligent: trees, mold, staphylococcus, cancer, grapes, etc. objects that are alive and intellig…

SH is in the wrong bucket

too soon
Post reply on HN