Claude Sonnet 5
141–150 of 822 posts
Re: Claude Sonnet 5
#142Earlier quoted context omitted.
Why do you think they are bragging? Anthropic has long been the company to give us by far the most in-depth information about their models, both positive and negative. I read this as them just stating a fact about this model that users would want to know.
I'm absolutely certain that their marketing team has input on (if not owning) these announcements.
Re: Claude Sonnet 5
#143Earlier quoted context omitted.
I've been largely disappointed how much the Claude models ignore custom instructions, and sometimes even prompts on the chat interface. It sometimes feels like talking to a wall, or as if there was a third person in the chatroom whose messages I can't see. I can't help but feel this is intentional towards the 'Agentic' workflow.
> or as if there was a third person in the chatroom whose messages I can't see. If you set off a classifier, that's how it looks to Claude.
IMO, they were quite good with checklists even a year ago, and tried to tick off each one.
Re: Claude Sonnet 5
#144Earlier quoted context omitted.
Totally agreed. I sometimes wonder if they are making the model "lazy" with each iteration, it keeps getting better at avoiding work.
This is why Fable was so good. It followed instructions and it was in no way lazy.
Re: Claude Sonnet 5
#145> Evaluations also show that it has a much lower ability to perform cybersecurity tasks than our current Opus models. Why would they brag about something like this? It's like they know people want to use models to perform cybersecurity tasks yet knowingly deny them the ability. And Opus 4.8 is still cheaper for a higher pass rate (much less open weight models like GLM 5.2) so not sure why I'd use Sonnet except on the…
Re: Claude Sonnet 5
#146Earlier quoted context omitted.
"Lower ability to perform cybersecurity-related tasks" makes me super concerned it will leave my codebase like Swiss cheese for any American granny with access to Fable 5, when we non-American Brits, or rest-of-worlders, don't have access to it to clean our codebases.
That’s not even close to true. Unless you’re vibe coding trash that a better model might catch.
Re: Claude Sonnet 5
#147Earlier quoted context omitted.
"Lower ability to perform cybersecurity-related tasks" makes me super concerned it will leave my codebase like Swiss cheese for any American granny with access to Fable 5, when we non-American Brits, or rest-of-worlders, don't have access to it to clean our codebases.
I think they don’t understand that cybersecurity skills are what prevent bad code from making it into production. It’s like telling a chef to cook without a knife because knives can kill people. Dario and his lackeys at Anthropic aren’t visionaries.
Re: Claude Sonnet 5
#148Is it just me or is there a huge difference between how much one can accomplish in a 5-hour window with GPT 5.5 on xhigh versus any Claude model?