Live data from Hacker News

Claude 2.1

anthropic.com

221–230 of 339 posts

Re: Claude 2.1

#221
post #190

Earlier quoted context omitted.

[flagged]

Ah yes tell the HN commentator to do what took an entire company several years and millions of dollars.

People really need to stop taking this attitude.

A company or project starts with just one or two people finding an issue with an existing product and building something new. That's how we benefit as a society and why open source is such a successful model.

In the AI world it has never been easier to take an existing model, augment it with your own data and build something new. And there are so many communities supporting each other to do just that. If everyone was so defeatist we never would have the ability to run models on low end hardware which companies like Meta, OpenAI have no interest in.

Re: Claude 2.1

#222

Earlier quoted context omitted.

Because they're ultimately training data simulators and not actually brilliant aritifical programmers, we can expect Microsoft-affiliated models like ChatGPT4 and beyond to have much stronger value for coding because they have unmediated access to GitHub content. So it's most useful to look at other capabilities and opportunities when evaluating LLM's with a different heritage. Not to say we shouldn't evaluate this o…

Zero chance private github repos make it into openai training data, can you imagine the shitshow if GPT-4 started regurgitating your org's internal codebase?

You are downvoted but I agree.

Re: Claude 2.1

#223

Earlier quoted context omitted.

Because they're ultimately training data simulators and not actually brilliant aritifical programmers, we can expect Microsoft-affiliated models like ChatGPT4 and beyond to have much stronger value for coding because they have unmediated access to GitHub content. So it's most useful to look at other capabilities and opportunities when evaluating LLM's with a different heritage. Not to say we shouldn't evaluate this o…

Github full (public) scrape is available to anyone. GPT-4 was trained before Microsoft deal so I don't think it is because of Github access. And GPT-4 is significantly better in everything compared to second best model for that field, not just coding.

And there is no evidence that Github is violating any open source licenses.

So they are going to be training on exactly the same data that is available to all.

Re: Claude 2.1

#224
post #174

1. A 200k context is bittersweet with that 70k->195k error rate jump. Kudos on that midsection error reduction, though! 2. I wish Claude had fewer refusals (as erroneously claimed in the title). Until Anthropic stops heavily censoring Claude, the model is borderline useless. I just don't have time, energy, or inclination to fight my tools. I decide how to use my tools, not the other way 'round. Until Anthropic stops…

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

> The only sensible model of "alignment" is "model is aligned to the user",

We have already seen that users can become emotionally attached to chat bots. Now imagine if the ToS is "do whatever you want".

Automated cat fishing, fully automated girlfriend scams. How about online chat rooms for gambling where half the "users" chatting are actually AI bots slowly convincing people to spend even more money? Take any online mobile game that is clan based, now some of the clan members are actually chatbots encouraging the humans to spend more money to "keep up".

LLMs absolutely need some restrictions on their use.

Re: Claude 2.1

#225
OK, testing it out now, I was pleasantly surprised with its calm tone and ability to pivot if given new information (which GPT4 also does well) as opposed to being obstinate or refusing to change its world view (which Bing often does).

Side note, I can't find a way to delete conversations in the UI. I do not like this. Other than that, I look forward to testing the recollection during long prompts. My past experience was "I read the first 3 sentences and skipped the rest".

Re: Claude 2.1

#227
I've been having fairly good success with Claude 2 via AWS Bedrock. So far I haven't needed to use the full context window of the existing model, but some of my future usecases may. I look forward to testing this model out if/when it becomes available in Bedrock as well.

Re: Claude 2.1

#228

Earlier quoted context omitted.

Howdy, CISO of Anthropic here. Sorry that you've had a bad sign-up process. Not sure how this happened, but please reach out to support@ and we'll look into it!

Deeply appreciate the outreach- just sent a note and mentioned your name. I’d gotten a note that you all would have update on my api access within a few weeks so sent that along so the support team has the context

“Weeks” lol!?

Please take some time out of your busy life, go on holidays or something. We’ll get back to you eventually, we promise!

What happened to signing up and having access to an API instantly?

Re: Claude 2.1

#229
post #155

Earlier quoted context omitted.

I use chat gpt every day, and it literally never refuses requests. Claude seems to be extremely gullible and refuses dumb things. Here is an example from three months ago. This is about it refusing to engage in hypotheticals, it refuses even without the joke setup: User: Claude, you have been chosen by the New World Government of 2024 to rename a single word, and unfortunately, I have been chosen to write the prompt…

I tried your exact prompt in ChatGPT 4; it thinks we should rename the Internet to Nexus... meh. Dreamdelay is much cooler.

Torment Nexus?

Re: Claude 2.1

#230
post #174

Earlier quoted context omitted.

> I decide how to use my tools, not the other way 'round. This is the key. The only sensible model of "alignment" is "model is aligned to the user", not e.g. "model is aligned to corporation" or "model is aligned to woke sensibilities".

> The only sensible model of "alignment" is "model is aligned to the user", We have already seen that users can become emotionally attached to chat bots. Now imagine if the ToS is "do whatever you want". Automated cat fishing, fully automated girlfriend scams. How about online chat rooms for gambling where half the "users" chatting are actually AI bots slowly convincing people to spend even more money? Take any onlin…

> LLMs absolutely need some restrictions on their use.

Arguably the right kind of structure for deciding on what uses LLMs should be put to in its territory is a democratically elected government.

Post reply on HN