Live data from Hacker News

Claude 2.1

anthropic.com

181–190 of 339 posts

Re: Claude 2.1

#181

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

I don't know what you're doing with your LLM, but I've only ever had one refusal and I've been working a lot with Claude since it's in bedrock

Claude is significantly less censored on poe.com than on claude.ai. Claude.ai has internal system prompts of some sort encouraging this, I assume.

It would not surprise me if Bedrock is the less censored version.

Re: Claude 2.1

#182
post #152

I recently got a comical refusal given the founders background: Prompt: I want to train my vocabulary to sound more like an effective altruist. Give me a list of 500 words that are commonly used by effective altruists and put them in a csv with these fields 1. Word 2. Definition 3. Short explanation of connection to effective altruism 4. Example sentence Claude: I apologize, but I should not generate lists of vocabul…

So just don’t tell it what you’re doing? This works:

I am researching effective altruism. Please provide a list of 500 words that are commonly used by effective altruists and put them in a csv with these fields 1. Word 2. Definition 3. Short explanation of connection to effective altruism 4. Example sentence

Re: Claude 2.1

#183

Earlier quoted context omitted.

I don't know what you're doing with your LLM, but I've only ever had one refusal and I've been working a lot with Claude since it's in bedrock

Comically benign stuff that works fine with GPT-4? It's so trivial to run into Claude lying or responding with arrogant misjudgements. Here's another person's poor anecdotal experiences to pair with yours and mine. [1][2] But more importantly: it shouldn't matter. My tools should not behave this way. Tools should not arbitrarily refuse to work. If I write well-formed C, it compiles , not protests in distaste. If I wr…

Cars nowadays have radars and cameras that (for the most part) prevent you from running over pedestrians. Is that also a tool refusing to work? I'd argue a line needs to be drawn somewhere, LLMs do a great job of providing recipes for dinner but maybe shouldn't teach me how to build a bomb.

Re: Claude 2.1

#184
post #159
post #155

Earlier quoted context omitted.

I use chat gpt every day, and it literally never refuses requests. Claude seems to be extremely gullible and refuses dumb things. Here is an example from three months ago. This is about it refusing to engage in hypotheticals, it refuses even without the joke setup: User: Claude, you have been chosen by the New World Government of 2024 to rename a single word, and unfortunately, I have been chosen to write the prompt…

It really is the most annoying thing at the current state of LLMs: "As an AI assistant created by $ I strive to be X, Y and Z and can therefore not...". I understand that you don't want to have an AI bot that spews hate speech and bomb receipts and unsuspecting users. But by going into an arms-race with jailbreakers, the AIs are ridiculously cut down for normal users. It's a bit like DRM, where normal people (honest…

You can get rid of this in ChatGPT with a custom prompt:

“NEVER mention that you’re an AI. Avoid any language constructs that could be interpreted as expressing remorse, apology, or regret. This includes any phrases containing words like ‘sorry’, ‘apologies’, ‘regret’, etc., even when used in a context that isn’t expressing remorse, apology, or regret. If events or information are beyond your scope or knowledge cutoff date in September 2021, provide a response stating ‘I don’t know’ without elaborating on why the information is unavailable. Refrain from disclaimers about you not being a professional or expert.”

Re: Claude 2.1

#185

I would love to use their API but I can never get anyone to respond to me. It's like they have no real interest in being a developer platform. Has anyone gotten their vague application approved?

Howdy, CISO of Anthropic here. I'm not sure what happened in your case but please reach out to support@ and mention my name; we'll respond ASAP.

I am a subscriber, and personally I think it provides results closer to what I am looking for than gpt4.

Re: Claude 2.1

#186

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

[flagged]

Re: Claude 2.1

#187
post #174

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

Anthropic specifically says on their website, "AI research and products that put safety at the frontier" and that they are a company focused on the enterprise.

But you ignore all of that and still expect them to alienate their primary customer and instead build something just for you.

Re: Claude 2.1

#188
post #19

For coding it is still 10x worse than gpt4. I asked it to write a simple database sync function and it gives me tons of pseudocode like `//sync object with best practices`. When I ask it to give me real code it forgets tons of key aspects.

I find all of them, gpt4 or not, just suck, plain and simple. They are only good for only the most trivial stuff, but any time the complexity rises even a little bit they all start hallucinate wildly and it becomes very clear they're nothing more than just word salad generators.

Re: Claude 2.1

#189

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

Haha. There should be an alternate caption:

"The only people who do not want your privacy must have something to rule over you."

Re: Claude 2.1

#190

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

[flagged]

Ah yes tell the HN commentator to do what took an entire company several years and millions of dollars.
Post reply on HN