Live data from Hacker News

Elevated error rate across multiple models

status.claude.com

171–180 of 293 posts

Re: Elevated error rate across multiple models

#171
post #142

Earlier quoted context omitted.

This is not like that. This is literally they saying they are letting their LLM run wild(ish) and seeing the status.claude.com we can see the result. This is a case where the outcome is the direct result of the engineering practices like the ones they describe. PS: Yes I use Claude, Coded, Amp and Cursor agents every day so I am not saying here LLMs are not valuable. LE: They did not made claims that "AI is good" the…

Another data point: GitHub is extremely insistent its employees maximally use AI for internal development [0], and we’ve concomitantly seen its reliability fall off a cliff in the last year or so. [0] https://github.com/resources/insights/ai-powered-workforce-p...

>GitHub is extremely insistent its employees maximally use AI for internal development

Or it could be that GitHub saw a 14x increase in commit volume last year[0], and we've concomitantly seen its reliability fall of a cliff in the last year or so. Given that Microsoft is leasing additional space on AWS(!)[1] to handle the additional commit volume, my personal money is on commit volume growth being a bigger issue than internal use of AI.

Internal use of AI may have been an issue. Commit volume growth may have been an issue. Unless one has direct knowledge of their infrastructure issues, claiming to know is quite literally making exactly the "they are vegan, their illness must be caused by their veganism" argument the GP commenter was talking about.

[0]https://daringfireball.net/linked/2026/05/04/commits-on-gith...

[1]https://www.businessinsider.com/microsoft-github-amazon-ai-c...

Re: Elevated error rate across multiple models

#172
post #142

Earlier quoted context omitted.

This is not like that. This is literally they saying they are letting their LLM run wild(ish) and seeing the status.claude.com we can see the result. This is a case where the outcome is the direct result of the engineering practices like the ones they describe. PS: Yes I use Claude, Coded, Amp and Cursor agents every day so I am not saying here LLMs are not valuable. LE: They did not made claims that "AI is good" the…

You just said "it's not the 'must be veganism' thing, it's the 'must be veganism thing'" Unless you have inside knowledge of their infra ops and management tools, it is just guessing and blaming veganism. For all we know it could be tools from Nvidia or anyone else failing under massive load. It could be the veganism. Some things are. Leaping to it as the only possible explanation for every ailment is exactly the fal…

No. We dont need metaphors like that with veganism (which touches ideologies also) when talking about engineering and a company that promotes out loud that engineering is done.

I have not stated anything. I just replied to a metaphor which is not needed cause here we talk about engineering problems handled by engineers in a tech company. I give you something else where this line of thought could be wrong: culture beats (and destroys) engineering practices unless regulated by law. In this case yes this is not because of LLMs but because of company culture.

Still hard to know where the line draws because Anthropic talks about solving computer science for good as in humans need not apply.

Re: Elevated error rate across multiple models

#174
post #100

Earlier quoted context omitted.

As long as I'm alive (and not senile) there will always be at least one developer who can code I'm not using AI coding tools yet, and even if they force me at gunpoint to use them at work no one can force me to in my spare time I'm not too worried about the case where no one can code anymore because that will be after I'm dead

That wasn't the premise of the question.

My answer to your question is "I don't care, because I'll be dead"

Re: Elevated error rate across multiple models

#176

Earlier quoted context omitted.

>> What can your company do? Hire some Developers?

Developers who can code without LLMs will go extinct in couple years and there will be legends about them, you should at least have some decent open weight model as a backup

I don't plan on using LLMs for programming any time soon.

And I know like one guy who does use them. He's not a developer by trade, he just has to write programs sometimes.

Re: Elevated error rate across multiple models

#177

Earlier quoted context omitted.

Meh, this is the "must be the veganism" fallacy: if someone knows you're vegan, then any ailment you might have, no matter how ubiquitous in the population, must be somehow due to your vegan diet and no more details are required. Except now it's the "AI did it" fallacy where if you know a company uses AI, even infra scaling issues must be due to AI, and if you had just used less or no AI, you would have been spared e…

Ahem... "Vegans and vegetarians may have higher stroke risk" - https://www.bbc.com/news/health-49579820 "Vegans had a 43% higher risk of fractures overall compared to nonvegetarians, as well as higher risks of hip, leg, and vertebral fractures." - https://sniglobal.org/plant-based-diets-and-fracture-risk/ "The Impact of a Vegan Diet on Many Aspects of Health: The Overlooked Side of Veganism" - https://www.cureus.com/…

There are a lot of nutritional blind spots in vegan diets. It is a diet that requires exceptional planning and intentionality to be at a baseline of health similar to a balanced omnivorous diet.

So indeed, the "it must be veganism" is not an unfounded concern when health complications arise, in a very similar way to "it must be the AI" is a valid concern when software issues arise.

Re: Elevated error rate across multiple models

#178
post #127
post #106

Earlier quoted context omitted.

Both of them provide that option. I've never installed rust without a package manager. Why would I?

> Why would I? Because then you can install it without depending on a package manager?

Yeah, from source in that case. Or using a verified binary if I absolutely had to.

Re: Elevated error rate across multiple models

#179

I hear that 100% of code at Anthropic is coded by Claude, so this was caused by Claude. And also, no one but Claude can fix Claude

>> I hear that 100% of code at Anthropic is coded by Claude, so this was caused by Claude. And also, no one but Claude can fix Claude Claude is down....

Don't worry they've got a dude named Claude they keep in the back just for occasions like this

Re: Elevated error rate across multiple models

#180
post #178
post #127

Earlier quoted context omitted.

> Why would I? Because then you can install it without depending on a package manager?

Yeah, from source in that case. Or using a verified binary if I absolutely had to.

Yes, if you want to, you can do that.

Understand that 99% are comfortable trusting downloads. They know that it's just as easy to sneak backdoors into source code as it is to sneak backdoors into executables.

See also: XZ hack.

Post reply on HN