Live data from Hacker News

Claude 2.1

anthropic.com

321–330 of 339 posts

Re: Claude 2.1

#321

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

Which models do you prefer?

Re: Claude 2.1

#322
post #317

Claude refuses a lot. GPT4 also refuses a lot and one has to try several prompts to get out what you need. LLMs are trained on the entire internet and more. I want a model that just gives me the answer with whatever it knows instead of playing pseudoethics. Sure it can say this is dangerous “don’t do this at home” but let me be the judge of it.

But aren't you a small child, and doesn't the AI know so much more than you?

To be honest, what they view as ethical is actually unethical: this idea that the AI knows more than a human, in the human's situation, and can pass judgment on that human.

Re: Claude 2.1

#323
I can't even register because it requires phone verification and myy country Czechia is not on the list. I don'teven think that phone verification should be necessary. I expect it to be highly censored thus useless anyway. I will stick with opensource models. <3

Re: Claude 2.1

#324
post #174

Earlier quoted context omitted.

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

Anthropic specifically says on their website, "AI research and products that put safety at the frontier" and that they are a company focused on the enterprise. But you ignore all of that and still expect them to alienate their primary customer and instead build something just for you.

No, I mean any user, including enterprise.

With some model (not relevant which one, might or might not be Anthropic's), we got safety-limited after asking the "weight of an object" because of fat shaming (i.e. woke sensibilities).

That's just absurd.

Re: Claude 2.1

#325
post #236

Earlier quoted context omitted.

It is very unlikely that the development team will be able to build features that actually cause the model to act in the best interests of humanity on every inference. What is far more likely is that the development team will build a model that often mistakes legitimate use for nefarious intent while at the same time failing to prevent a tenacious nefarious user from getting the model to do what they want.

I think the current level of caution in LLMs is pretty silly: while there are a few things I really don't want LLMs doing (telling people how to make pandemics is a big one) I don't think keeping people from learning how to hotwire a car (where the first google result is https://www.wikihow.com/Hotwire-a-Car ) is worth the collateral censorship. One thing that has me a bit nervous about current approaches to "AI safe…

That reminds me of my last query to ChatGPT. A colleague of mine usually writes "Mop Programming" when referencing out "Mob programming" sessions. So as a joke I asked ChatGPT to render an image of a software engineer using a mop trying to clean up some messy code that spills out of a computer screen. It told me that it would not do this because this would display someone in a derogatory manner.

Another time I tried to let it generate a very specific Sci-fi helmet which covers the nose but not the mouth. When it continusly left the nose visible, I tried to tell it to make this particular section similar to Robocop, which caused it again to deny to render because it was immediately concerned about copyright. While I at least partially understand the concern for the last request, this all adds up to making this software very frustrating to use.

Re: Claude 2.1

#326
post #172

Earlier quoted context omitted.

I’ve had some really absurd ChatGPT refusals. I wanted some invalid UTF-8 strings, and ChatGPT was utterly convinced that this was against its alignment and refused (politely) to help.

That's not absurd, you absolutely don't want invalid strings being created within then passed between layers of a text-parsing model. I don't know what would happen but I doubt it would be ideal. 'hey ai, can you crash yourself' lol

Huh? The LLMs (mostly) use strings of tokens internally, not bytes that might be invalid UTF-8. (And they use vectors between layers. There’s no “invalid” in this sense.)

But I didn’t ask for that at all. I asked for a sequence of bytes (like “0xff” etc) or a C string that was not valid as UTF-8. I have no idea whether ChatGPT is capable of computing such a thing, but it was not willing to try for me.

Re: Claude 2.1

#327
post #326

Earlier quoted context omitted.

That's not absurd, you absolutely don't want invalid strings being created within then passed between layers of a text-parsing model. I don't know what would happen but I doubt it would be ideal. 'hey ai, can you crash yourself' lol

Huh? The LLMs (mostly) use strings of tokens internally, not bytes that might be invalid UTF-8. (And they use vectors between layers. There’s no “invalid” in this sense.) But I didn’t ask for that at all. I asked for a sequence of bytes (like “0xff” etc) or a C string that was not valid as UTF-8. I have no idea whether ChatGPT is capable of computing such a thing, but it was not willing to try for me.

You can understand why, though, can't you?

Re: Claude 2.1

#328
post #46

I don’t like Anthropic. they over-RLHF their models and make them refuse most requests. A conversation with Claude has never been pleasant to me. it feels like the model has an attitude or something.

It's awful. 9/10 of things I ask Claud, I get denied because it crosses some kind of imaginary ethical boundary that's completely irrelevant.

[flagged]

Re: Claude 2.1

#329
post #174

Earlier quoted context omitted.

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

Anthropic specifically says on their website, "AI research and products that put safety at the frontier" and that they are a company focused on the enterprise. But you ignore all of that and still expect them to alienate their primary customer and instead build something just for you.

Well it's nice that it has one person who finds it useful.

Re: Claude 2.1

#330

Earlier quoted context omitted.

I am a subscriber, and personally I think it provides results closer to what I am looking for than gpt4.

That’s hard to believe but I’m open to the possibility. Cam you share a few examples that might demonstrate this?

Not really at the moment. I was asking it to help write some professional (but boring) letters of interest that were academic in nature. I found the style of writing to be closer to where I wanted it..so a very subjective.opinion.
Post reply on HN