Live data from Hacker News

I cancelled Claude: Token issues, declining quality, and poor support

nickyreinert.de

111–120 of 604 posts

Re: I cancelled Claude: Token issues, declining quality, and poor support

#111
post #68

Earlier quoted context omitted.

It feels more and more like OpenAI/Anthoropic aren't the future but Qwen, Kimi, or Deepseek are. You can run them locally, but that isn't really the point, it is about democratization of service providers. You can run any of them on a dozen providers with different trade-offs/offerings OR locally. They won't ever be SOTA due to money, but "last year's SOTA" when it costs 1/4 or less, may be good enough . More quantit…

Open Source isn't even within 50% of what the SOTA models are. Benchmarks are toys, real world use is vastly different, and that's where they seriously lag. Why should anyone waste time on poorer results? I'd rather pay my $200/mo because my time matters. I'm not a poor college student anymore, and I need more return on my time. I'm not shitting on open weights here - I want open source to win. I just don't see how t…

> Open Source isn't even within 50% of what the SOTA models are

Who said so? GLM 5.1 is 90% Opus, at least. Some people quite happy with Kimi 2.6 too. I did not try Deepseek 4 yet but also hearing it is as good as Opus. You might be confusing open source models with local models. It is not easy to run a 1.6T model locally, but they are not 50% of SOTA models.

Re: I cancelled Claude: Token issues, declining quality, and poor support

#112

Earlier quoted context omitted.

Luckily local AI is becoming more feasible every day.

Maybe for folks who are deep into this, but it’s not exactly accessible. I tried reading up on it a couple of months ago, but parsing through what hardware I needed, the model and how to configure it (model size vs quantization), how I’d get access to the hardware (which for decent results in coding, new hardware runs $4k-$10k last I checked)—it had a non trivial barrier of entry. I was trying to do this over a long…

> new hardware runs $4k-$10k last I checked

Starting closer to 40k if you want something that's practical. 10k can't run anything worthwhile for SDLC at useful speeds.

Re: I cancelled Claude: Token issues, declining quality, and poor support

#113
i ran prompts used up a ton of usage, and got no return just showed error.

Asked support hey i got nothing back i tried prompting several times used a ton of usage and it gave no response. I'd just like usage back. What I payed for I never got.

Just bot response we don't do refunds no exceptions. Even in the case they don't serve you what your plan should give you.

Re: I cancelled Claude: Token issues, declining quality, and poor support

#114

Earlier quoted context omitted.

Honestly, it sounds like, assuming you have no ethical qualms, you could get by with a Mac or AMD 395+ and the newest models, specifically QWEN3.5-Coder-Next. It does exactly as you describe. It maxes out around 85k context, which if you do a good job providing guard rails, etc, is the length of a small-medium project. It does seem like the sweet spot between WallE and the destroyed earth in WallE.

Sorry, out of the loop. Which ethical qualms are you referring to?

Using a Mac, obviously.

Re: I cancelled Claude: Token issues, declining quality, and poor support

#115
post #68

Earlier quoted context omitted.

It feels more and more like OpenAI/Anthoropic aren't the future but Qwen, Kimi, or Deepseek are. You can run them locally, but that isn't really the point, it is about democratization of service providers. You can run any of them on a dozen providers with different trade-offs/offerings OR locally. They won't ever be SOTA due to money, but "last year's SOTA" when it costs 1/4 or less, may be good enough . More quantit…

Open Source isn't even within 50% of what the SOTA models are. Benchmarks are toys, real world use is vastly different, and that's where they seriously lag. Why should anyone waste time on poorer results? I'd rather pay my $200/mo because my time matters. I'm not a poor college student anymore, and I need more return on my time. I'm not shitting on open weights here - I want open source to win. I just don't see how t…

> Why should anyone waste time on poorer results?

Because in almost no real-world project is "programming time" the limiting factor?

Re: I cancelled Claude: Token issues, declining quality, and poor support

#116
post #88

Earlier quoted context omitted.

I also use it this way and I'm overall pretty happy with it, but it feels like they really want us to use it in "autopilot" mode. It's like they have two conflicting priorities of "make people use more tokens so we can bill them more" and "people are using more tokens than expected, our pricing structure is no longer sustainable" (but I guess they're not really conflicting, if the "solution" involves upgrading to a h…

I feel like they are making it harder to use it this way. Encouraging autonomous is one thing, but it really feels more like they are handicapping engaged use. I suspect it reflects their own development practices and needs.

[deleted]

Re: I cancelled Claude: Token issues, declining quality, and poor support

#117
post #61

Earlier quoted context omitted.

Luckily local AI is becoming more feasible every day.

Indeed, I feel like we are in the early computer equivalent phase of AI, where giant expensive hardware is still required for frontier models. In 5 years I bet there will be fully open models we'll be able to run on a few $1000 of consumer hardware with equivalent performance to opus 4.7/4.6.

You'll never have the power of what they have though. Cloud capital is insane.

So you can run 1 agent locally on $1k to $3k hardware

They can run a fleet of thousands

Re: I cancelled Claude: Token issues, declining quality, and poor support

#118
post #99

If all Claude does is automate mundane code, why not just make a "meta library" of said common mundane code snippets?

Like Stack Overflow?

you still have to search stack overflow and sift, but I'm surprised someone doesn't just make a TLDR style or shortcuts-expanding product out of it that you can just pop in code for this use case or that. vs. the current product spending a few datacenter's worth of energy to give you the answer that's only correct x% of the time

Re: I cancelled Claude: Token issues, declining quality, and poor support

#119
Maybe this is an unpopular opinion, but I think choosing which companies to support during this period of pre-alignment is one way to vote which direction this all goes. I'm happy to accept a slightly worse coding agent if it means I don't get exterminated someday.

Re: I cancelled Claude: Token issues, declining quality, and poor support

#120

Claude with Sonnet medium effort just used 100% of my session limit, some extra dollars, thought for 53 minutes, and said: API Error: Claude's response exceeded the 32000 output token maximum. To configure this behavior, set the CLAUDE_CODE_MAX_OUTPUT_TOKENS environment variable.

Just copy and past the error back to Claude and you will be able to continue. I have seen this many times over the past few months. I thought it was related to AWS bedrock that I have been using - but probably not.
Post reply on HN