Live data from Hacker News

AMD AI director says Claude Code is becoming dumber and lazier since update

theregister.com

1–10 of 19 posts

Re: AMD AI director says Claude Code is becoming dumber and lazier since update

#4
Ok, I thought I was going insane. The last two larger coding tasks I gave Claude Code it left about 35% of my request completely undone or done sloppily.

I because of this, the next task I gave it on the larger side, I ran its work through Codex which identified 7 glaring unfinished parts of the task.

The trend was starting the part of the task but then leaving a "skeleton" of what I has requested without any of the actual working parts.

The way I would describe it is a kid cramming his 3 month project into a Sunday evening for Monday's due date.

Re: AMD AI director says Claude Code is becoming dumber and lazier since update

#5
post #3

Boris from the Claude Code team explained this on HN 2 days ago https://news.ycombinator.com/item?id=47664442

And as the person who raised the issue said

> The frustrating part is that it's not a workflow _or_ model issue, but a silently-introduced limitation of the subscription plan. They switched thinking to be variable by load, redacted the thinking so no one could notice, and then have been running it at ~1/10th the thinking depth nearly 24/7 for a month. That's with max effort on, adaptive thinking disabled, high max thinking tokens, etc etc.

So Boris' explanation isn't really an explanation.

Re: AMD AI director says Claude Code is becoming dumber and lazier since update

#6
i run claude code pretty heavily for overnight sessions and yeah the inconsistency b/w runs is noticeable. same prompt, same codebase, wildly different quality depending on the day. the frustrating part is when it half-finishes something and you come back to a mess you now have to untangle. still the most capable coding agent i've used but the variance is real.

Re: AMD AI director says Claude Code is becoming dumber and lazier since update

#7
post #2

Lol OAI and AMD did a deal together so whatever. In reality as they scale up, the models lose nuance and become noisier. The boosters do not want to admit this. We need highly-specialised models/interfaces. Not one thing and trying to force-fit it.

  > Not one thing and trying to force-fit it.
agree, but then they become glorified ide plugins and can't justify the huge valuations that a magic box that does and knows everything can justify...

Re: AMD AI director says Claude Code is becoming dumber and lazier since update

#9
post #5
post #3

Boris from the Claude Code team explained this on HN 2 days ago https://news.ycombinator.com/item?id=47664442

And as the person who raised the issue said > The frustrating part is that it's not a workflow _or_ model issue, but a silently-introduced limitation of the subscription plan. They switched thinking to be variable by load, redacted the thinking so no one could notice, and then have been running it at ~1/10th the thinking depth nearly 24/7 for a month. That's with max effort on, adaptive thinking disabled, high max th…

> ~1/10th the thinking depth

While simultaneously drastically reducing the amount of work you can get done even at $200 a month. I've cancelled my subscription, it's not worth it anymore.

Re: AMD AI director says Claude Code is becoming dumber and lazier since update

#10

Ok, I thought I was going insane. The last two larger coding tasks I gave Claude Code it left about 35% of my request completely undone or done sloppily. I because of this, the next task I gave it on the larger side, I ran its work through Codex which identified 7 glaring unfinished parts of the task. The trend was starting the part of the task but then leaving a "skeleton" of what I has requested without any of the…

Today Claude asked if I "wanted to leave this until tomorrow" as it was a "big rework", then stopped, requiring me to tell it to continue multiple times - that seemed kinda weird to me, it doesn't have the context of time of working day or similar (I'd only just started for one).

I have no idea what link it made to ask that, what in its training data or prompts, but it's very much "not a useful result".

I don't remember seeing anything similar, but have only been using Claude on and off for 6 months or so.

Post reply on HN