Live data from Hacker News

Degraded performance for multiple models

status.claude.com

91–100 of 149 posts

Re: Degraded performance for multiple models

#91

I was drafting a partnership document and Opus 5 decided that including my company's revenues, churn, assets would "make us appear a more legitimate counterparty". Thank God I read what it outputted or that could have been awkward. I cannot believe Opus 5 is a frontier level model after seeing that. I immediately cancelled my entire claude.ai subscription and am perfectly happy using a mixture of open weights + codex…

Opus5 decided it would be a good idea to include my email as part an HTTP request header's User Agent because Anthropic includes your email as part of the system prompt.

Fucking morons over there.

Re: Degraded performance for multiple models

#94
post #61

Earlier quoted context omitted.

I've been bouncing between the two for years now with great success. It's easy for me because I don't use any of the skills, agent.md, or sort of custom instructions. When it comes to most companies, there is no reward for loyalty.

skills and agent.md are very portable though? I figure at most, as the models get better, the only maintenance you need to do is pare them down to remove unnecessary context.

I mean, I wouldn't know how portable they are between models because I don't use them. I don't use them because, well, I don't need them. I mainly use LLMs as a StackOverflow replacement via the Web chatbots.

I only use CC or Codex for a quick MVP once every few months or so.

Re: Degraded performance for multiple models

#95
post #81

It's very interesting. I think Anthropic's early success in coding/tooling resulted in a lot of workflows using claude. I have started using every bit of my spare capacity to now move off these workflows. It's almost at a point now that if I use anything but Fable, the quality is subpar, Compared to alternatives (closed and open). The only reason I use Fable is because my harnesses still depend on claude code.

It’s so hard to be sure but opus feels like its been steadily declining since 4.6

Re: Degraded performance for multiple models

#99
post #87

I was drafting a partnership document and Opus 5 decided that including my company's revenues, churn, assets would "make us appear a more legitimate counterparty". Thank God I read what it outputted or that could have been awkward. I cannot believe Opus 5 is a frontier level model after seeing that. I immediately cancelled my entire claude.ai subscription and am perfectly happy using a mixture of open weights + codex…

"Thank God I read what it outputted" lmao

"Thank God i checked to see if the gun was loaded before i pointed it in a random direction and pulled the trigger". sheesh, my assumptions of general human intelligence continues to be wrong.

Re: Degraded performance for multiple models

#100
Monthly uptime:

Claude API - 99.27%

Claude Code - 99.16%

Claude.ai - 99.14%

At any large tech company these numbers would get entire teams of engineers fired. Anthropic, meanwhile, has been busy selling its “better than human engineers” AI while not managing to crack three 9s of availability.

Post reply on HN