Live data from Hacker News

Claude 2.1

anthropic.com

261–270 of 339 posts

Re: Claude 2.1

#261
post #134

Earlier quoted context omitted.

From my perspective it sounds pretty cheap if we get to the answers immediately.

Have you tried it? GPT4 fails as often as it succeeds at coding questions I ask so I'm not going to shell out that kind of money to take my chances.

Claude? No, have requested access many times but radio silence.

OpenAI? I use ChatGPT A LOT for coding as some mixture of pair programmer and boilerplate, works generally well for me. On the API side use it heavily for other work and its more directed and have a very high acceptance rate.

Re: Claude 2.1

#263

Earlier quoted context omitted.

I don't know what you're doing with your LLM, but I've only ever had one refusal and I've been working a lot with Claude since it's in bedrock

Comically benign stuff that works fine with GPT-4? It's so trivial to run into Claude lying or responding with arrogant misjudgements. Here's another person's poor anecdotal experiences to pair with yours and mine. [1][2] But more importantly: it shouldn't matter. My tools should not behave this way. Tools should not arbitrarily refuse to work. If I write well-formed C, it compiles , not protests in distaste. If I wr…

"[...]If I write well-formed C, it compiles, not protests in distaste. If I write a note, the app doesn't disable typing because my opinion sucks[...]"

There's a rust compiler joke/rant somewhere to be added here for comical effect

Re: Claude 2.1

#264
post #174

Earlier quoted context omitted.

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

> The only sensible model of "alignment" is "model is aligned to the user", We have already seen that users can become emotionally attached to chat bots. Now imagine if the ToS is "do whatever you want". Automated cat fishing, fully automated girlfriend scams. How about online chat rooms for gambling where half the "users" chatting are actually AI bots slowly convincing people to spend even more money? Take any onlin…

> chatbots encouraging the humans to spend more money ... LLMs absolutely need some restrictions on their use.

No, I can honestly say that I do not lose any sleep over this, and I think it's pretty weird that you do. Humans have been fending off human advertisers and scammers since the dawn of the species. We're better at it than you account for.

Re: Claude 2.1

#265

Earlier quoted context omitted.

I am using Claude 2 every day for chatting, summarisation and talking to papers and never run into a refusal. What are you asking it to do? I find Claude more fun to chat with than GPT-4, which is like a bureaucrat.

How did you get API access?

He didn’t mention API. Just use the web interface

Re: Claude 2.1

#266

Earlier quoted context omitted.

You can get rid of this in ChatGPT with a custom prompt: “NEVER mention that you’re an AI. Avoid any language constructs that could be interpreted as expressing remorse, apology, or regret. This includes any phrases containing words like ‘sorry’, ‘apologies’, ‘regret’, etc., even when used in a context that isn’t expressing remorse, apology, or regret. If events or information are beyond your scope or knowledge cutof…

Chatgpt 4 just randomly ignores these instructions, particularly after the first response.

I suspect this is related to whatever tricks they're doing for the (supposed) longer context window. People have noted severe accuracy loss for content in the middle of the context, which to me suggests some kind of summarization step is going on in the background instead of text actually being fed to the model verbatim.

Re: Claude 2.1

#267

I don’t like Anthropic. they over-RLHF their models and make them refuse most requests. A conversation with Claude has never been pleasant to me. it feels like the model has an attitude or something.

Maybe he is parisian

Re: Claude 2.1

#270

Earlier quoted context omitted.

I am using Claude 2 every day for chatting, summarisation and talking to papers and never run into a refusal. What are you asking it to do? I find Claude more fun to chat with than GPT-4, which is like a bureaucrat.

How did you get API access?

Aws bedrock?
Post reply on HN