Live data from Hacker News

Claude now has access to a server-side container environment

anthropic.com

101–110 of 363 posts

Re: Claude now has access to a server-side container environment

#101
post #4
post #3

They need to focus on fixing reliability first. Their systems constantly go down and it appears they are having to quantise the models to keep up with demand, reducing intelligence significantly. New features like this feel pointless when the underlying model is becoming unusable.

This can't be understated. I started using it heavily earlier this summer and it felt like magic. Someone signing up now based on how I described my personal experiences with it then would think I was out of my mind . For technical tasks it has been a net negative for me for the last several weeks. (Speaking of both Claude Code and the desktop app, both Sonnet and Opus >=4, on the Max plan.)

I felt like the model degraded lately as well, I've been using Claude everyday for months now

Re: Claude now has access to a server-side container environment

#102
post #3

They need to focus on fixing reliability first. Their systems constantly go down and it appears they are having to quantise the models to keep up with demand, reducing intelligence significantly. New features like this feel pointless when the underlying model is becoming unusable.

Have you run into the bug where claude acts as if it updated the artifact, but it didn’t? You can see the changes in real time, but then suddenly it’s all deleted character by character as if the backspace was held down, you’re left with the previous version, but claude carries on as if everything is fine. If you point it out, it will acknowledge this, try again, and… same thing. The only reliable fix I’ve seen is to…

Yes, this was so frustrating.

I had to keep prompting it to generate new artifacts all the time.

Thankfuly that is mostly gone with Claude Code.

Re: Claude now has access to a server-side container environment

#103
post #91

Earlier quoted context omitted.

Some of this has gotta be people asking more of it than they did before, and some has gotta be people who happened to use it for things it's good at to begin with and are now asking it things it's bad at (not necessarily harder things, just harder for the model). However there have been some bugs causing performance degradation acknowledged by Anthropic as well (and fixed) and so I would guess there's a good amount o…

What makes it particularly tricky to evaluate is that there could still be other bugs given how long these went without even acknowledgement until now, and they did state they are still looking into potential Opus issues. I'll probably come back and try a Claude Code subscription again, but I'm good for the time being with the alternative I found. I also kind of suspect the subscription model isn't going to work for…

Benchmarks are too expensive for ordinary users to run, but it would be useful if they could publish their benchmarks using prod over time, that would expose degradations in a more objective manner.

Of course there’s always the problem of teaching to the test and out of test degradations, but presumably bugs would be independent of that.

Re: Claude now has access to a server-side container environment

#104

Is it able to process a prompt on each file in a folder-full of files and then return the collated results? That's the functionality which I could use for my day job, but I'm not finding an LLM which directly affords that capability (without programming or other steps which are difficult on my work computer).

I bet you could easily get an LLM to write a python script that would do that for you.

Can the LLM also convince IT to allow me to run Python?

I'd like an all-in-one tool of an LLM front-end which can access multiple files since that is more easily explained/permission granted for.

Re: Claude now has access to a server-side container environment

#105
post #91

Earlier quoted context omitted.

What makes it particularly tricky to evaluate is that there could still be other bugs given how long these went without even acknowledgement until now, and they did state they are still looking into potential Opus issues. I'll probably come back and try a Claude Code subscription again, but I'm good for the time being with the alternative I found. I also kind of suspect the subscription model isn't going to work for…

Benchmarks are too expensive for ordinary users to run, but it would be useful if they could publish their benchmarks using prod over time, that would expose degradations in a more objective manner. Of course there’s always the problem of teaching to the test and out of test degradations, but presumably bugs would be independent of that.

A few weeks ago reddit was on fire with outages and timeouts and yet the Anthropic Jira status page was showing everything as green. So even if they had benchmarks, I'm not sure they'd be transparent with them.

Re: Claude now has access to a server-side container environment

#106
post #71

Earlier quoted context omitted.

This, so much this... I signed up for Claude over a week ago and I totally regret it! Previously I was using it and some ChatGPT here and there (also had a subscription in the past) and I felt like Claude added some more value. But it's getting so unstable. It generates code, I see it doing that, and then it throws the code away and gives me the previous version of something 1:1 as a new version. And then I have to w…

> But it's getting so unstable. It generates code, I see it doing that, and then it throws the code away and gives me the previous version of something 1:1 as a new version. I've had the same experience. Totally unreliable.

I regularly have this happen:

1. Ask Claude to fix something

2. It fails to fix the issue

3. I tell it that the fix didn’t work

4. It reverts its failed fix and tells me everything is working now.

This is like finding a decapitated body, trying to bring it back to life by smooshing the severed head against the neck, realizing that didn’t bring them back to life, dropping the head back on the ground, and saying, “There; I’ve saved them now.”

Re: Claude now has access to a server-side container environment

#108

Earlier quoted context omitted.

Have you considered perhaps that you are, indeed, out of your mind? Or more precisely, that you could be rationalizing what is essentially a random process? Based on the discussions here it seems that every model is either about to be great or was great in the past but now is not. Sucks for those of us who are stuck in the now , though.

Anthropic literally stated yesterday that they suffered degraded model performance over the last month due to bugs: https://status.anthropic.com/incidents/72f99lh1cj2c Suggesting people are "out of their mind" is not really appropriate on this forum, especially so in this circumstance.

The first comment claims that Anthropic "are having to quantise the models to keep up with demand", to which the parent comment agrees with "This can't be understated". So based on this discussion so far Anthropic has [1] great models, [2] models that used to be great but now aren't due to quantization, [3] models that used to be great but now aren't due to a bug, and [4] models that constantly feel like a "bait and switch".

This most definitely feels like people analyzing the output of a random process - at this point I am feeling like I'm losing my mind.

(As for the phrasing I was quoting the OP, who I believe took it in the spirit in which it was meant)

[1] https://news.ycombinator.com/item?id=45183587

[2] https://news.ycombinator.com/item?id=45182714

[3] https://news.ycombinator.com/item?id=45183820

[4] https://news.ycombinator.com/item?id=45183281

Re: Claude now has access to a server-side container environment

#109
post #47

Maybe one day Claude can rewrite its interface to be more accessible to blind people like me.

Curious what a11y issues you see with Claude? I use it a remarkable amount and haven't found any showstoppers. Web interface and Claude Code.

Claude has no TTS while most LLMs have it. It makes the text more accessible.
Post reply on HN