Live data from Hacker News

I used to love Claude, but the latest models are slowly ruining it

androidauthority.com

31–40 of 63 posts

Re: I used to love Claude, but the latest models are slowly ruining it

#31
post #10

It's quite obnoxious. I asked if brown rice left in the fridge for a couple days–originally put in for use in fried rices–was still safe. Fable decided I'm trying to produce biotoxins. Which, ironically, prompted me to learn how to produce Bacillus cereus at home [1]. I paid for a year but am going back to Kagi's multi-model system [2]. [1] https://pmc.ncbi.nlm.nih.gov/articles/PMC7913059/ Don't Do It [2] https://ass…

"Beware of he who would deny you access to information, for in his heart he dreams himself your master." (he, in this case, would not be the llm but the people over it) I find those kind of limitation very dystopian and way more dangerous than the threat they claim to fight against.

Alpha Centauri was such a fantastic game. And even though it encourages you to move forward with technology, the tone has this unease about what you're doing. There's a lot to be learned from it.

Re: I used to love Claude, but the latest models are slowly ruining it

#32
Opus 4.5 was the high point for me. It was like a mind-reader, it just got it and did pretty much exactly what I wanted.

Since then, I've been less impressed and I agree it feels a bit downhill. At work we are "stuck" on Opus 4.6 which is okay but I feel like that was when the deviant opinionated behaviour started to creep in.

It's a tool, I don't want my hammer to refuse to hammer a screw if I decide that is what I want to do today. I know it's wrong, but I'm the fucking boss.

Re: I used to love Claude, but the latest models are slowly ruining it

#33
post #8

Claude has never been the best Chat agent. GPT and Gemini have the lead there. But Claude chat is still perfectly serviceable if you don’t want to pay for two.

> Claude has never been the best

Hard disagree. It wasn’t that long ago that Gpt was clearly falling behind, and Gemini was like the “and also in the room”.

Re: I used to love Claude, but the latest models are slowly ruining it

#34

From my point of view, a lot of frustration is tied to LLMs not evolving at the same rate/manner as people around you do. We're using human language against a system that produces human-like output, which tricks our brain into having similar expectations.

I don’t think that this is the case. I’m using Claude the same way I did at the beginning of the year, and results are increasingly more untrustworthy and I don’t know how to say it, but they are “less” for same effort from my side.

It is super anecdotal, but I’m convinced that Claude peaked in January/April period and since then is on steady decline. And I don’t know what they do in background, but the older models (Opus 4.6 in particular) have degraded too. Same for codex but I use it less overall.

“You were right to push back on this”, “my previous claim was too broad” - this is super annoying

Re: I used to love Claude, but the latest models are slowly ruining it

#35

It's quite obnoxious. I asked if brown rice left in the fridge for a couple days–originally put in for use in fried rices–was still safe. Fable decided I'm trying to produce biotoxins. Which, ironically, prompted me to learn how to produce Bacillus cereus at home [1]. I paid for a year but am going back to Kagi's multi-model system [2]. [1] https://pmc.ncbi.nlm.nih.gov/articles/PMC7913059/ Don't Do It [2] https://ass…

> I paid for a year but am going back to Kagi's multi-model system [2]. I've been using DuckDuckGo's multi-model service for my "ask an AI random questions" needs. I was already paying them and discovered this LLM thing is part of my subscription. Works pretty well and has privacy guarantees I'd expect out of DDG, though I think they've been tightening the usage you can get out of it recently. I'll have to try Kagi i…

I use them frequently but I'm skeptical of their "private" claims.

I asked one of the DDG ChatGPT bots about some error output from a program and apparently forgot to redact a URL. I found the DDG bot trying to access it about half an hour later. It was a long and randomly generated alpha-numeric subdomain and the conversation never discussed the URL or anything related to it- it was simply buried in the pasted error output. I think this was even before the DGG bots had web search capability. I was thoroughly spooked.

Re: I used to love Claude, but the latest models are slowly ruining it

#36
This is a serious suggestion, not a joke: Have you tried being nice to the model?

There are so many criticisms here that I just don't see myself.

If the models have been trained on human responses, then it's plausible that they will prefer to become less helpful to requests which are blunt or even rude, because that's what humans do too.

Re: I used to love Claude, but the latest models are slowly ruining it

#37

This is a serious suggestion, not a joke: Have you tried being nice to the model? There are so many criticisms here that I just don't see myself. If the models have been trained on human responses, then it's plausible that they will prefer to become less helpful to requests which are blunt or even rude, because that's what humans do too.

I am always very polite with my chat bots. Mostly because the interface does mimic human conversation and I’m afraid that if I’m rude to the bot, I’ll get used to being rude and then I’ll start being rude to people.

Re: I used to love Claude, but the latest models are slowly ruining it

#38
post #10

Earlier quoted context omitted.

"Beware of he who would deny you access to information, for in his heart he dreams himself your master." (he, in this case, would not be the llm but the people over it) I find those kind of limitation very dystopian and way more dangerous than the threat they claim to fight against.

Alpha Centauri was such a fantastic game. And even though it encourages you to move forward with technology, the tone has this unease about what you're doing. There's a lot to be learned from it.

Between Yang, Zakharov and Miriam the whole unease aspect of tech progress goes to 11. Absolute gem of a game.

Re: I used to love Claude, but the latest models are slowly ruining it

#39

It's quite obnoxious. I asked if brown rice left in the fridge for a couple days–originally put in for use in fried rices–was still safe. Fable decided I'm trying to produce biotoxins. Which, ironically, prompted me to learn how to produce Bacillus cereus at home [1]. I paid for a year but am going back to Kagi's multi-model system [2]. [1] https://pmc.ncbi.nlm.nih.gov/articles/PMC7913059/ Don't Do It [2] https://ass…

I cancelled my subscription because of the obnoxious "safety classifers" as well. Will switch to OpenAI next month. I don't ever want to hear about Anthropic's patronizing safety nonsense ever again.

Re: I used to love Claude, but the latest models are slowly ruining it

#40

This is a serious suggestion, not a joke: Have you tried being nice to the model? There are so many criticisms here that I just don't see myself. If the models have been trained on human responses, then it's plausible that they will prefer to become less helpful to requests which are blunt or even rude, because that's what humans do too.

Yeah I'm in a similar boat - these experiences are wildly outside of my experience working on a moderately large, professional codebase on a daily basis.
Post reply on HN