Live data from Hacker News

Claude 2.1

anthropic.com

201–210 of 339 posts

Re: Claude 2.1

#201

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

[flagged]

Was the child insult really necessary?

Re: Claude 2.1

#202
post #174

Earlier quoted context omitted.

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

Anthropic specifically says on their website, "AI research and products that put safety at the frontier" and that they are a company focused on the enterprise. But you ignore all of that and still expect them to alienate their primary customer and instead build something just for you.

It has problems summarizing papers because it freaks out about copyright. I then need to put significant effort into crafting a prompt that both gaslights and educates the LLM into doing what I need. My specific issue is that it won't extract, format or generally "reproduce" bibliographic entries.

I damn near canceled my subscription.

Re: Claude 2.1

#203

I would love to use their API but I can never get anyone to respond to me. It's like they have no real interest in being a developer platform. Has anyone gotten their vague application approved?

I applied today; hopefully it will be a short wait. (and, hopefully, they won't hold my "I don't know what business I can build on this until after I try it" opinion against me)

Re: Claude 2.1

#204
post #174

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

What's the issue with including some amount of "model is aligned to the interests of humanity as whole"?

If someone asks the model how to create a pandemic I think it would be pretty bad if it expertly walked them through the steps (including how to trick biology-for-hire companies into doing the hard parts for them).

Re: Claude 2.1

#205

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

I'm not one to mind the guardrails - but what i hate is something you mentioned, fighting the tool. Eg "Do an X-like thing" where X is something it may not be allowed to do, gets rejected. But then i say "Well, of course - that's why i said X-like. Do what you can do in that direction, so that it is still okay". Why do i even have to say that? I get why, but still - just expressing my frustration. I'm not trying to p…

This reminds me of everyday interactions on StackOverflow. "Yes, I really really really do want to use the library and language I mentioned."

Re: Claude 2.1

#206
post #174

Earlier quoted context omitted.

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

Anthropic specifically says on their website, "AI research and products that put safety at the frontier" and that they are a company focused on the enterprise. But you ignore all of that and still expect them to alienate their primary customer and instead build something just for you.

[deleted]

Re: Claude 2.1

#207

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

I'm not one to mind the guardrails - but what i hate is something you mentioned, fighting the tool. Eg "Do an X-like thing" where X is something it may not be allowed to do, gets rejected. But then i say "Well, of course - that's why i said X-like. Do what you can do in that direction, so that it is still okay". Why do i even have to say that? I get why, but still - just expressing my frustration. I'm not trying to p…

This is a great point, and something that may be at least partially addressable with current methods (e.g. RLHF/SFT). Maybe (part of) what's missing is a tighter feedback loop between a) limitations experienced by the human users of models (e.g. "actually okay but just near the off limits stuff"), and b) model training signal.

Thank you for the insightful perspective!

Re: Claude 2.1

#208
post #204
post #174

Earlier quoted context omitted.

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

What's the issue with including some amount of "model is aligned to the interests of humanity as whole"? If someone asks the model how to create a pandemic I think it would be pretty bad if it expertly walked them through the steps (including how to trick biology-for-hire companies into doing the hard parts for them).

[deleted]

Re: Claude 2.1

#210
post #159

Earlier quoted context omitted.

It really is the most annoying thing at the current state of LLMs: "As an AI assistant created by $ I strive to be X, Y and Z and can therefore not...". I understand that you don't want to have an AI bot that spews hate speech and bomb receipts and unsuspecting users. But by going into an arms-race with jailbreakers, the AIs are ridiculously cut down for normal users. It's a bit like DRM, where normal people (honest…

You can get rid of this in ChatGPT with a custom prompt: “NEVER mention that you’re an AI. Avoid any language constructs that could be interpreted as expressing remorse, apology, or regret. This includes any phrases containing words like ‘sorry’, ‘apologies’, ‘regret’, etc., even when used in a context that isn’t expressing remorse, apology, or regret. If events or information are beyond your scope or knowledge cutof…

Chatgpt 4 just randomly ignores these instructions, particularly after the first response.
Post reply on HN