Yet https://marginlab.ai/trackers/claude-code/ says no issue. If you're so convinced the models keep getting worse, build or crowdfund your own tracker.
Issue: Claude Code is unusable for complex engineering tasks with Feb updates
221–230 of 829 posts
Re: Issue: Claude Code is unusable for complex engineering tasks with Feb updates
#222Until there is either more capacity or some efficiency breakthroughs the only way for providers to cut costs is to make the product worse.
Re: Issue: Claude Code is unusable for complex engineering tasks with Feb updates
#223Earlier quoted context omitted.
There's been more going on than just the default to medium level thinking - I'll echo what others are saying, even on high effort there's been a very significant increase in "rush to completion" behavior.
Thanks for the feedback. To make it actionable, would you mind running /bug the next time you see it and posting the feedback id here? That way we can debug and see if there's an issue, or if it's within variance.
Re: Issue: Claude Code is unusable for complex engineering tasks with Feb updates
#224Hey all, Boris from the Claude Code team here. I just responded on the issue, and cross-posting here for input. --- Hi, thanks for the detailed analysis. Before I keep going, I wanted to say I appreciate the depth of thinking & care that went into this. There's a lot here, I will try to break it down a bit. These are the two core things happening: > `redact-thinking-2026-02-12` This beta header hides thinking from th…
I look at it, and I am very upset that I no longer see it.
Re: Issue: Claude Code is unusable for complex engineering tasks with Feb updates
#225Re: Issue: Claude Code is unusable for complex engineering tasks with Feb updates
#226Not claude code specific, but I've been noticing this on Opus 4.6 models through Copilot and others as well. Whenever the phrase "simplest fix" appears, it's time to pull the emergency break. This has gotten much, much worse over the past few weeks. It will produce completely useless code, knowingly (because up to that phrase the reasoning was correct) breaking things. Today another thing started happening which are…
The cope is hard. Just at this point admit that the LLM tech is doomed and sucks.
Re: Issue: Claude Code is unusable for complex engineering tasks with Feb updates
#227Hey all, Boris from the Claude Code team here. I just responded on the issue, and cross-posting here for input. --- Hi, thanks for the detailed analysis. Before I keep going, I wanted to say I appreciate the depth of thinking & care that went into this. There's a lot here, I will try to break it down a bit. These are the two core things happening: > `redact-thinking-2026-02-12` This beta header hides thinking from th…
The irony lol. The whole ticket is just AI-generated. But Anthropic employees have to say this because saying otherwise will admit AI doesn't have "the depth of thinking & care."
Re: Issue: Claude Code is unusable for complex engineering tasks with Feb updates
#228Earlier quoted context omitted.
I've noticed a strong degradation as its started doing more skill like things and writing more one off python scripts rather than using tools. the agent has a set of scripts that are well tested, but instead it chooses to write a new bespoke script everytime it needs to do something, and as a result writes both the same bugs over and over again, and also unique new bugs every time as well.
I'm going absolutely insane with this. Nearly all of my "agent engineering" effort is now figuring out how to keep Opus from YOLO'ing is own implementation of everything. I've lost track of the number of times it's started a task by building it's own tools, I remind it that it has a tool for doing that exact task, then it proceeds to build it's own tools anyways. This wasn't happening 2 months ago.
Re: Issue: Claude Code is unusable for complex engineering tasks with Feb updates
#229Called it 10 days ago: https://news.ycombinator.com/item?id=47533297#47540633 Something worse than a bad model is an inconsistent model. One can't gauge to what extent to trust the output, even for the simplest instructions, hence everything must be reviewed with intensity which is exhausting. I jumped on Max because it was worth it but I guess I'll have to cancel this garbage.
I don't see how this can be the future of software engineering when we have to put all our eggs in Anthropic's basket.