Live data from Hacker News

Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

anthropic.com

171–180 of 758 posts

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#171
post #84

Why not rev the numbers? "3.5" vs. "3.5 New" feels weird -- is there a particular reason why Anthropic doesn't want to call this 3.6 (or even 3.5.1)?

The confusing choice they seem to have made is that "Claude 3.5 Sonnet" is a name, rather than 3.5 being a version. In their view, the model "version" is now `claude-3-5-sonnet-20241022` (and was previously `claude-3-5-sonnet-20240620`). https://docs.anthropic.com/en/docs/about-claude/models

OpenAI does exactly the same thing, by the way; the named models also have dated versions. For instance, there current models include (only listing versions with more than one dated version for the same "name" version):

  gpt-4o-2024-08-06 
  gpt-4o-2024-05-13
  gpt-4-0125-preview
  gpt-4-1106-preview
  gpt-4-0613
  gpt-4-0314
  gpt-3.5-turbo-0125
  gpt-3.5-turbo-1106

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#172

I have been a paying ChatGPT customer for a long time (since the very beginning). Last week I've compared ChatGPT to Claude and the results (to my eye) were better, the output better structured and the canvas works better. I'm on the edge of jumping ship.

Anthropic's rate limit are very low sadly, even for paid customers. You can use the API of course but it's not as convenient and may be more expensive.

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#174

Is there an easy way to use Claude as a Co-Pilot in VS Code? If it is better at coding, it would be great to have it integrated.

Cody by Sourcegraph has unlimited code completions for Claude & a very generous monthly message limit. They don't have this new version I think but they roll these out very fast.

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#175
post #15

Why not rev the numbers? "3.5" vs. "3.5 New" feels weird -- is there a particular reason why Anthropic doesn't want to call this 3.6 (or even 3.5.1)?

For a company selling intelligence, that's a pretty stupid way of labelling a new product.

Every major AI vendor seems to do it with hosted models; within "named" major versions of hosted models, there are also "dated" minor versions. OpenAI does it. Google does it (although for Google Gemini models, the dated instead of numbered minor versions seem to be only for experimental versions like gemini-1.5-pro-exp-0827, stabled minor versions get additional numbers like gemini-1.5-pro-002.)

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#176
post #173

This "Computer use" demo: https://www.youtube.com/watch?v=jqx18KgIzAE shows Sonnet 3.5 using the Google web UI in an automated fashion. Do Google's terms really permit this? Will Google permit this when it is happening at scale?

I wonder how they could combat it if they choose to disallow AI access through human interfaces. Maybe more captchas, anti-AI design language, or even more tracking of the user's movements?

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#177
post #9

I still feel like the difference between Sonnet and Opus is a bit unclear. Somewhere on Anthropic's website it says that Opus is the most advanced, but on other parts it says Sonnet is the most advanced and also the fastest. The UI doesn't make the distinction clear either. Then on Perplexity, Perplexity says that Opus is the most advanced, compared to Sonnet. And finally, in the table in the blogpost, Opus isn't eve…

Anthropic use the names Haiku/Sonnet/Opus for the small/medium/large versions of each generation of their models, so within-generation that is also their performance (& cost) order. Evidentially Sonnet 3.5 outperforms Opus 3.0 on at least some tasks, but that is not a same-generation comparison.

I'm wondering at this point if they are going to release Opus 3.5 at all, or maybe skip it and go straight to 4.0. It's possible that Haiku 3.5 is a distillation of Opus 3.5.

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#178
post #148

Why on god's green earth is it not just called Claude 3.6 Sonnet. Or Claude 4 Sonnet. I don't actually care what the answer is. There's no answer that will make it make sense to me.

The best answer I've seen so far is that "Claude 3.5 Sonnet" is a brand name rather than a specific version. Not saying I agree, just a way to visualize how the team is coming up with marketing.

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#179
post #116

Earlier quoted context omitted.

As I understand Cursor tab autocomplete uses their own model. Only chat has Sonnet and co.

Ah, i thought it used the model selected for your prompts, either way, it seems to work very well

I originally thought that too but learned yesterday they have their own model. Definitely explains how its so fast and accurate!

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#180

Is there an easy way to use Claude as a Co-Pilot in VS Code? If it is better at coding, it would be great to have it integrated.

Cody by Sourcegraph has unlimited code completions for Claude & a very generous monthly message limit. They don't have this new version I think but they roll these out very fast.

Cody (https://cody.dev) will have support for the new Claude 3.5 Sonnet on all tiers (including the free tier) asap. We will reply back here when it's up.
Post reply on HN