Live data from Hacker News

Claude 2.1

anthropic.com

41–50 of 339 posts

Re: Claude 2.1

#41
I know you guys from Anthropic are reading this. Love you guys, but PLEASE open access in EU - even if it means developer preview no strings attached or whatever. If you don't, you're going to make us talk to your board on Friday. Please.

Re: Claude 2.1

#42
post #19

For coding it is still 10x worse than gpt4. I asked it to write a simple database sync function and it gives me tons of pseudocode like `//sync object with best practices`. When I ask it to give me real code it forgets tons of key aspects.

Agreed, but I do find gpt4 has been increasing the amount of pseudo code recently. I think they are a/b testing me. I find myself asking if how much energy it wasted giving me replies that I then have to tell it to fix.. Which is of course a silly thing to do, but maybe someone at oAI is listening?

If you mean through the user friendly chat GPT website, they're probably making it output as few tokens as possible to cut costs

Re: Claude 2.1

#44
post #26

This is where OpenAI/MSFT loses. Chaos in OpenAI/MSFT will lead to Anthropic overtaking them. They've already been ahead in many areas, dead locked in others, but with OpenAI facing a crisis, they'll likely gain significant headway if they execute well .. at least for the risk-adverse enterprise use-cases. I still am not a fan of either due to restrictions and 'safety' training wheels that treat me like a child

From what I see they still suck bad

But at least there are heads down and focused on their product /their company (employees) and not all about themselves & their egos. Employees who arent being used as pawns .. if Altman didn't flail around and did just that (moved all into new company backed or under Microsoft) they'd not look like pawns rather following a strong leader who demands self respect first / foremost.

Re: Claude 2.1

#45
post #19

For coding it is still 10x worse than gpt4. I asked it to write a simple database sync function and it gives me tons of pseudocode like `//sync object with best practices`. When I ask it to give me real code it forgets tons of key aspects.

Except: you can feed it an entire programming language manual, all the docs for all the modules you want to use, and _then_ it's stunningly good, whipping chatgpt4 that same 10x.

Re: Claude 2.1

#46

I don’t like Anthropic. they over-RLHF their models and make them refuse most requests. A conversation with Claude has never been pleasant to me. it feels like the model has an attitude or something.

It's awful. 9/10 of things I ask Claud, I get denied because it crosses some kind of imaginary ethical boundary that's completely irrelevant.

Re: Claude 2.1

#47

This is where OpenAI/MSFT loses. Chaos in OpenAI/MSFT will lead to Anthropic overtaking them. They've already been ahead in many areas, dead locked in others, but with OpenAI facing a crisis, they'll likely gain significant headway if they execute well .. at least for the risk-adverse enterprise use-cases. I still am not a fan of either due to restrictions and 'safety' training wheels that treat me like a child

I mean, that would be predicated on it actually being possible to get access to and use their models...which in my experience is basically a limitless void. Meanwhile I spend hundreds of dollars a month with msft/oai.

AWS Bedrock has Claude. It took 30 mins for approval.

Re: Claude 2.1

#48
That 200k context needs some proper testing. GPT-4-Turbo advertises 128k but the quality of output there goes down significantly after ~32k tokens.

Re: Claude 2.1

#49
post #19

For coding it is still 10x worse than gpt4. I asked it to write a simple database sync function and it gives me tons of pseudocode like `//sync object with best practices`. When I ask it to give me real code it forgets tons of key aspects.

Except: you can feed it an entire programming language manual, all the docs for all the modules you want to use, and _then_ it's stunningly good, whipping chatgpt4 that same 10x.

I honestly don’t have time for that level of prompt engineering. So, chatGPT wins (for me)
Post reply on HN