Live data from Hacker News

Claude 3.5 Sonnet

anthropic.com

161–170 of 287 posts

Re: Claude 3.5 Sonnet

#161

Unfortunately still thinks "There are two 'r's in the word "raspberry"." The only one that got it right was the basic version of Gemini "There are actually three "r"s in the word "strawberry". It's a bit tricky because the double "r" sounds like one sound, but there are still two separate letters 'r' next to each other." The paid Gemini advanced had "There are two Rs in the word "strawberry"."

One theory I heard about this type of problem is because these algorithms tokenize the text early, and each token can be multiple characters.

Re: Claude 3.5 Sonnet

#162
So far it isn’t doing better or worse than gpt4o, if you want someone to switch, it better be way better or way better price. The price is exactly what OpenAI is charging to the cent. So no, you won’t get someone to switch because the differentiator is just the UI

Re: Claude 3.5 Sonnet

#163

Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…

What I understand is that it's GPT 6 that just went into training, and that GPT 5 is complete and being delayed until after the U.S. election.

I also believe that gpt-4o was originally called gpt-5. If you look at the image generation on their website from gpt-4o which has not been released, I believe that along with the voice caused Ilya to declare mission accomplished (AGI) and that is why there was a coup. The coup failed because no one wanted to wrap up the company or change the way it operated because they would lose a lot of money.

The reason the name was changed was because there was a big public scare about gpt-5 taking over and so Altman had to promise not to release gpt-5 soon. So they changed the name to gpt-4o (omni). Which is A) obviously dramatically a different architecture, B) a huge step up in capabilities (most still unreleased) C) very general purpose. Because of A) and B), this should obviously be a new major version (5).

Yes, this is speculation, but it's very obvious speculation to me. It's weird for me that most people not only don't share this view but seem to absolutely hate when I say it.

Re: Claude 3.5 Sonnet

#164

Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…

Both GPT-4 and 4o have been completely useless for coding in the past couple of weeks for me - constant errors, and not just your typical LLM inaccuracies but incapable of producing a few lines of self-consistent code e.g. defines variables foo on one line and refers to it as bar on the next, or it misspells it as foox.

It's the same model though. Maybe your perception has changed.

Re: Claude 3.5 Sonnet

#165

Using this is the first time since GPT-4 where I've been shocked at how good a model is. It's helped by how smooth the 'artifact' UI is for iterating on html pages, but I've been instructing it to make a simple web app one bit of functionality at a time and it's basically perfect (and even quite fast). I'm sure it will be like GPT-4 and the honeymoon period will wear off to reveal big flaws but honestly I'd take this…

i don't think the point of an intern is to have them do this kind of work. to me, it's just a side effect if they accomplish anything at all. if we take this to its logical conclusion, without the kind of basic training that comes from internships, where will we be in 5 years?

The only hope is that intern level will also increase significantly with this, helping them to catch up with the fundamentals super quickly.

Re: Claude 3.5 Sonnet

#166
I gave it a fairly simple coding questions and it failed pretty severely, to be fair ChatGPT 4o also failed that. Just saying it ain’t all that given the hype

Re: Claude 3.5 Sonnet

#167

Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…

I've had way better success with GPT-4o than claude. I wonder why

Re: Claude 3.5 Sonnet

#168
post #90
post #11

Earlier quoted context omitted.

a Kagi Ultimate subscription gets you access to both (plus others) for $25/mo

This is only via API though. There is a level of magic that Claude.ai and ChatGPT bring to the table that makes it worthwhile.

I can't speak to any new features announced today but the API version of Claude has been superior in every way when paired with a more feature rich front end.

Re: Claude 3.5 Sonnet

#169
post #25

Just tried it. This is the first model that immediately gives me the correct answer to my test prompt: "Hi , can you give me an exact solution to pi in python?". All other models I've tried first give an approximation, taking several prompts to come to the correct conclusion: it's impossible.

Couldn't it output "use a symbolic math library in Python to get an exact solution to pi" and technically be correct?

Re: Claude 3.5 Sonnet

#170

Opus remained better than GPT for me, even after the release of GPT-4o. VERY happy to see an even further improvement beyond that, Claude is a terrific product and given the news that GPT-5 only began its training several weeks ago I don't see any situation where Anthropic is dethroned in the near term. There are only two parts of Anthropic's offering I'm not a fan of: - Lack of conversation sharing: I had a conversa…

I've had way better success with GPT-4o than claude. I wonder why

Have you tried 3 Opus or 3.5 Sonnet? Are you using it for programming, or something else?
Post reply on HN