I'm afraid watermarking could restrict applications where LLMs can be safely used to assist with writing. If I write something myself and use an LLM to proofread it, without watermarking I can confidently say that corrections done by LLMs are small and insignificant enough to claim that the text is still authored by me, not by the model. With watermarking, however, I will never be sure if the result will not be flagg…
Claude Fable 5.1 and Claude Mythos 5.1
561–570 of 1001 posts
Re: Claude Fable 5.1 and Claude Mythos 5.1
#562My main gripe with LLMs is the cringe AI phrasings that they use in UI elements. Pompous things like "Your keys, supercharged" or weird yoda-speak stuff like "searches the app remembers" instead of just naming the thing "Learned searches" .. you know, proper GUI copy like it was done for the past decades. I jumped when I saw a mention about "writing style improvements" so I gave it a try on a recent feature in rcmd […
> Where are all these verbal tics coming from and why is it so hard to get rid of them? It’s a side effect of post-training for effectiveness and efficiency at technical tasks. Over time the models learn to pack as much information as possible into their available context window, because that’s one way to increase the effective intelligence. Humans do this too with industry jargon, dense tech-talk, etc. We have a lim…
Re: Claude Fable 5.1 and Claude Mythos 5.1
#563My main gripe with LLMs is the cringe AI phrasings that they use in UI elements. Pompous things like "Your keys, supercharged" or weird yoda-speak stuff like "searches the app remembers" instead of just naming the thing "Learned searches" .. you know, proper GUI copy like it was done for the past decades. I jumped when I saw a mention about "writing style improvements" so I gave it a try on a recent feature in rcmd […
> Where are all these verbal tics coming from and why is it so hard to get rid of them? It’s a side effect of post-training for effectiveness and efficiency at technical tasks. Over time the models learn to pack as much information as possible into their available context window, because that’s one way to increase the effective intelligence. Humans do this too with industry jargon, dense tech-talk, etc. We have a lim…
But who has both the compute power and the motivation to do such a thing?
I guess I'll just continue rewriting the UI one word at a time for the time being.
Re: Claude Fable 5.1 and Claude Mythos 5.1
#564(I work at Anthropic) Beyond all the benchmarks, I think Fable 5.1 is a big improvement in writing style. It sounds a lot less stereotypically like other Claude models, has (imho) a much more natural style, and responds to my style instructions more reliably. More work to be done (and we will!) but reading better prose makes me so much happier. Another point I expect not to get much attention until it all happens at…
As a fervent Claude Code user who made the switch to GPT 5.6 Sol over Opus 5 over hard-to-read prose this makes me happy. I love your product but the current models are very hard to work with if you need to do a lot of context switching. Brevity is key.
Re: Claude Fable 5.1 and Claude Mythos 5.1
#565What I don't see in the comments: "I had a specific problem I couldn't solve with the previous version of this LLM. But the improvements in this version unlocked the solution for me." What I do see in the comments: subjective improvement in text generation, possibly lower cost, some optimism about code generation, but some skepticism too. I use coding agents. To me they are very useful. But what I spend on them isn't…
I get your point, but we can only have groundbreaking leaps once in a blue moon. That doesn’t mean incremental improvements aren’t useful.
Re: Claude Fable 5.1 and Claude Mythos 5.1
#566Earlier quoted context omitted.
As someone working in science, this belief confuses me. How (by what means) do you think Fable 5.1 will be able to make further progress in scientific domains? The problem with science is that there is no agentic harness. The agent can't test things. At best it can hallucinate something and ask if that hallucination "makes sense", but this doesn't work in science.
Great news, then! TFA: "Last week, we previewed the Model Hardware Standard, which allows Claude to directly and safely operate laboratory equipment."
Re: Claude Fable 5.1 and Claude Mythos 5.1
#567Earlier quoted context omitted.
While I can't speak for everyone in academia, I personally don't feel comfortable in putting my research questions and outputs to a private website, before the idea is at least arxived. Especially as all the Fable/Mythos prompts are said to be human reviewed. So I believe that, at least in the short run, we might be seeing breakthroughs in hard open problems or in low hanging problems which are not that interesting t…
Isn't it showing a problem with an academia? "I don't want to live in a world where someone else makes the world a better place than we do."
Re: Claude Fable 5.1 and Claude Mythos 5.1
#568I cancelled my pro max 20x subscription, tired of Opus stopping the work from time to time, or saying "this is 2 months of work"
Re: Claude Fable 5.1 and Claude Mythos 5.1
#569Looks like all three breaking changes are patches for inadvertent chain of thought disclosure. Someone found out (don't have the tweet handy) that if you created a bogus "think_deeply" tool and then forced the model to use it, it would output what is believed to be its raw thinking there - I believe the first breaking change stops this. The second two are aimed at people getting Haiku to repeat thinking blocks from o…
These draconian "Preserved Thinking" measures they're taking are going to be an absolute pain in the ass. This alone is enough for me to move our API use off their platform entirely. It's a HUGE breaking change that they're trying to dampen by having it not affecting current customers until "in the future", see: https://platform.claude.com/docs/en/build-with-claude/preser... You're no longer allowed to edit the conte…
Hm, aiui you can support both of these via mid-conversation system turns https://platform.claude.com/docs/en/build-with-claude/mid-co... - and in general you'd want to to preserve the cache and recency of the instruction anyways rather than frankensteining an off-distribution transcript. Not sure though.
Re: Claude Fable 5.1 and Claude Mythos 5.1
#570AI is really not "just software" anymore. It is able to discover facts and advance science. Hard to disagree that we're near or at the point where Artificial Intelligence has expanded reality into 4 quadrants: objects that are not alive: dust, rocks, water, wood, hats, lego, aluminum, etc. objects that are alive but not intelligent: trees, mold, staphylococcus, cancer, grapes, etc. objects that are alive and intellig…
SH is in the wrong bucket