Live data from Hacker News

Claude.ai unavailable and elevated errors on the API

status.claude.com

231–240 of 277 posts

Re: Claude.ai unavailable and elevated errors on the API

#231
post #197

We've been running our 10 dev org on 8 H100s on open models (with some tweaks). Sure they aren't as good as the big providers but they 1. don't go down 2. have pretty damn high tok/s. It pays for itself. Posting with a fresh account because I'm not supposed to share these details for obvious reason. If you want help on setting this up, just reply with a way to reach you.

We're planning to do the same thing - buy something like 8xH100 and run all coding there. The CTO almost agreed to find the budget for it but I need to make sure there are no risks before we buy (i.e. it's a viable/usable setup for professional AI-assisted coding) Can you share what models you run and find best performing for this setup? That would help a lot. I already run a smaller AI server in the office but only…

Deepseek, GLM, Minimax or Kimi are the most likely contenders.

Re: Claude.ai unavailable and elevated errors on the API

#232

Earlier quoted context omitted.

> Sure they aren't as good as the big providers If you haven't done so already, finetune the model on all your company's code that you can get your hands on. This is one of the great advantages that you get when running local models. I like the style of the generated code much better now, I have to rewrite much less, and my prompts can be shorter too. But maybe these already are the "tweaks" that you mentioned.

How would they do that? Would it be as easy as telling a model "Hey, review all this code, identify patterns, and then write in this style going forward"? Sorry if this is a stupid question, I've never finetuned or trained a LLM.

Unsloth has consumer accessible stuff on fine tuning models

Re: Claude.ai unavailable and elevated errors on the API

#234

Earlier quoted context omitted.

> Would I theoretically have a more stable harness backing my usage? If you don’t mind an opinionated harness that asks for a pretty specific workflow, but one that works well, use OpenCode. If you want to spread your wings and feel the sweet kiss of freedom, use Pi.

Im looking at moving to Pi and I like the minimal nature, but I disagree with a handful of decisions they make. So Id likely need to maintain a fork which is less than ideal.

check out my pi forks.

Re: Claude.ai unavailable and elevated errors on the API

#235

Earlier quoted context omitted.

Im looking at moving to Pi and I like the minimal nature, but I disagree with a handful of decisions they make. So Id likely need to maintain a fork which is less than ideal.

check out my pi forks.

Ummmmmm, how?

Re: Claude.ai unavailable and elevated errors on the API

#237

Earlier quoted context omitted.

I led the team that did the math and analysis for determining our direction in selecting Anthropic. We initially assumed this was where we would end up, but after some investment exploring our options we found it not worth the trouble. Local models sound great until you realize you dont get alot of the features that we implicitly expect from hosted models. Many things would require additional investment into the oper…

To start, I want to be clear I am trying to understand not criticizing, and mistakes are how institutional knowledge grows. Your last paragraph hints at retention struggles which complicates the issue. But was vendor mitigation not part of the evaluation? I get that most companies view governance and compliance as a pay to play issue, but there has always been an issue with rapidly changing areas and single source su…

it's an SMB, if you need redundancy on every 3rd party dependency your business will die anyway

better to take the risk for most things. if the worst case happens and you have to migrate, you migrate. otherwise you risk overengineering upfront and guaranteeing reduced productivity rather than risking it

Re: Claude.ai unavailable and elevated errors on the API

#238

Earlier quoted context omitted.

Have you guys considered running your own local models? 200k a month is a ton of money and puts all your eggs in one basket. Or is it easier to just be able to run away from it all if you are done with it or something changes?

Local models perform objectively worse than SotA SaaS models. Your employees will hate this decision.

Only if you're vibe coding, with ambiguous prompts that require the model to fill in a huge number of gaps and basically write the software for you.

The people who don't really know what they're doing (or don't care) need the full power of the SOTA models, those with experience can provide enough context and instruction to make even small local models work.

Re: Claude.ai unavailable and elevated errors on the API

#239
As a long-term 20x user, Claude has recently felt a lot like using AI for coding a year or so ago. It can't reliably handle basic tasks. I ask for something straightforward and get something subtly wrong, incomplete, or just not workable. I always use the best model available and effort levels maxed, but with all their changes I have to relearn how to make the model perform at best every day, and it seems I can't keep up. It’s not that Claude can’t do impressive things, it clearly can, but the inconsistency on simple, expected behavior makes it hard to use. The downtime is annoying but hasn't been the deciding factor. I’m not waiting it out this time. I’m switching over to Codex, and based on my usage today it looks like I’ll be fine on the 5x plan, so I can drop down and save about $100 a month which is nice. I didn't quite have a grasp on how quickly companies can change for better or worse until Anthropic showed me. I'm surprised at how quickly they brought me from a happily paying max user to not even wanting the lowest paid tiers.

Re: Claude.ai unavailable and elevated errors on the API

#240
post #216

Earlier quoted context omitted.

No, we can literally buy our own hardware for what we spend in a month and host our own local LLMs for company usage.

> and host our own local LLMs for company usage. What local alternative could replace your Anthropic use? I have found none. I don't think many have, which is why most of us pay Anthropic, rather than using one of the numerous, far cheaper, cloud services that host "local" class models. Most of us are paying for access to proprietary SOTA models, rather than hosting.

[deleted]
Post reply on HN