Live data from Hacker News

Elevated error rate across multiple models

status.claude.com

151–160 of 293 posts

Re: Elevated error rate across multiple models

#151
post #38

Earlier quoted context omitted.

"curl -fsSL https://pi.dev/install.sh | sh" — seriously? That tells me a lot about the whole project, unfortunately.

Seriously, what is the threat model here?

There is no threat model that doesn't also apply to pretty much every other distribution method.

It's just people who have internalized "don't paste commands from the Internet into your terminal" and aren't thinking about exactly what makes pasting commands from the Internet into your terminal dangerous, and how that applies to this specific case.

Re: Elevated error rate across multiple models

#152
post #142

Earlier quoted context omitted.

Meh, this is the "must be the veganism" fallacy: if someone knows you're vegan, then any ailment you might have, no matter how ubiquitous in the population, must be somehow due to your vegan diet and no more details are required. Except now it's the "AI did it" fallacy where if you know a company uses AI, even infra scaling issues must be due to AI, and if you had just used less or no AI, you would have been spared e…

This is not like that. This is literally they saying they are letting their LLM run wild(ish) and seeing the status.claude.com we can see the result. This is a case where the outcome is the direct result of the engineering practices like the ones they describe. PS: Yes I use Claude, Coded, Amp and Cursor agents every day so I am not saying here LLMs are not valuable. LE: They did not made claims that "AI is good" the…

But it is like that. You have zero insight into the infrastructure issue. And the person quoted above is a Claude Code developer. So because this guy uses Claude generously to build Claude Code, then Anthropic's API scaling issues must necessarily be caused by his agent loops even though scaling issues plague every tech company, no less often pre-AI.

The issue is that it's a thought-terminating cliche, and it would be nice to have one place on the internet that isn't just who can post one the fastest with the most glee to the giddy seal-clapping of the audience.

Re: Elevated error rate across multiple models

#153
post #142

Earlier quoted context omitted.

Meh, this is the "must be the veganism" fallacy: if someone knows you're vegan, then any ailment you might have, no matter how ubiquitous in the population, must be somehow due to your vegan diet and no more details are required. Except now it's the "AI did it" fallacy where if you know a company uses AI, even infra scaling issues must be due to AI, and if you had just used less or no AI, you would have been spared e…

This is not like that. This is literally they saying they are letting their LLM run wild(ish) and seeing the status.claude.com we can see the result. This is a case where the outcome is the direct result of the engineering practices like the ones they describe. PS: Yes I use Claude, Coded, Amp and Cursor agents every day so I am not saying here LLMs are not valuable. LE: They did not made claims that "AI is good" the…

Another data point: GitHub is extremely insistent its employees maximally use AI for internal development [0], and we’ve concomitantly seen its reliability fall off a cliff in the last year or so.

[0] https://github.com/resources/insights/ai-powered-workforce-p...

Re: Elevated error rate across multiple models

#154

I don’t prompt Claude anymore. I have loops running that prompt Claude and figuring out what to do. My job is to write loops. — Boris Cherny, head of Claude Code Reliability is a direct reflection of the quality of the underlying infrastructural code. If even Anthropic, the company with the world's best agentic vibecoders, has horribly unreliable infrastructure, it really says something about the quality of the world…

Meh, this is the "must be the veganism" fallacy: if someone knows you're vegan, then any ailment you might have, no matter how ubiquitous in the population, must be somehow due to your vegan diet and no more details are required. Except now it's the "AI did it" fallacy where if you know a company uses AI, even infra scaling issues must be due to AI, and if you had just used less or no AI, you would have been spared e…

There’s a difference between having normal levels of difficulty and bad luck, and having people blame those on the wrong thing, vs having extraordinarily miserable quality and having people find the obvious difference. Potentially yes, they might have terrible wiring in their office or a crippling fondness for vim. But if I were their PR department I’d be talking about that if it was the problem.

Re: Elevated error rate across multiple models

#155

I don’t prompt Claude anymore. I have loops running that prompt Claude and figuring out what to do. My job is to write loops. — Boris Cherny, head of Claude Code Reliability is a direct reflection of the quality of the underlying infrastructural code. If even Anthropic, the company with the world's best agentic vibecoders, has horribly unreliable infrastructure, it really says something about the quality of the world…

I wonder how they fix things when Claude is down.

"Gemini, fix my Claude infra"

Re: Elevated error rate across multiple models

#156
post #138
post #18

I suppose it's a good time to encourage people trying out pi[1] with any cheap model from the openrouter rankings page[1]. [1] https://pi.dev/ [2] https://openrouter.ai/rankings

Is pi better than opencode?

I like it.

One caveat is that it doesn't do MCP tools, but can wire them up with bash (or use CLIs if those are available).

Re: Elevated error rate across multiple models

#159

I don’t prompt Claude anymore. I have loops running that prompt Claude and figuring out what to do. My job is to write loops. — Boris Cherny, head of Claude Code Reliability is a direct reflection of the quality of the underlying infrastructural code. If even Anthropic, the company with the world's best agentic vibecoders, has horribly unreliable infrastructure, it really says something about the quality of the world…

Is there any indication these errors are related to Anthropic-written code as opposed to operational issues from the fastest-growing infra buildout ever? Layer-wise, the app is pretty far removed from request routing to GPU pools.

This is almost certainly a software issue, though. Even if it's due to scaling, they still built a system that failed catastrophically rather than degrading gracefully.

Re: Elevated error rate across multiple models

#160

Actual 90d uptime: 97.6838% (calculated by Codex from live data) Computed from the page’s own data for 2026-03-26 through 2026-06-23: - Partial outage: 43h 15m 1s - Major outage: 6h 46m 48s - Total affected time: 50h 1m 49s - Major-only uptime: 99.6861% So, only one 9 for 10x vibes.

I want uptime modulo in my timezone/work hours. I don't give a shit about any 9's earned while I'm sleeping.
Post reply on HN