Did I miss something? Did they have to make changes to the model for this?
Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
71–80 of 758 posts
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#72Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#73Earlier quoted context omitted.
Because its a finetune of 3.5 optimized for the use case of computer use. Its actually accurate and its not a 3.6.
I don't think that's correct. This looks like a new model. Significant jump in math and gpqa scores.
What if it isn't even a re-training from scratch but a fine-tune of an existing model/weights release, is it a new version then? Would be more like a iteration, or even a fork I suppose.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#74Great progress from Anthropic! They really shouldn't change models from under the hood, however. A name should refer to a specific set of model weights, more or less. On the other hand, as long as its actually advancing the Pareto frontier of capability, re-using the same name means everyone gets an upgrade with no switching costs. Though, all said, Claude still seems to be somewhat of an insider secret. "ChatGPT" ha…
In the API (https://docs.anthropic.com/en/docs/about-claude/models) they have proper naming you can rely on. I think the shorthand of "Sonnet 3.5" is just the "consumer friendly" name user-facing things will use. The new model in API parlance would be "claude-3-5-sonnet-20241022" whereas the previous one's full name is "claude-3-5-sonnet-20240620"
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#75I have been a paying ChatGPT customer for a long time (since the very beginning). Last week I've compared ChatGPT to Claude and the results (to my eye) were better, the output better structured and the canvas works better. I'm on the edge of jumping ship.
Yeah I think I might also jump ship. It’s just that chatGPT now kinda knows who I am and what I like and I’m afraid of losing that. It’s probably not a big deal though.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#76Why not rev the numbers? "3.5" vs. "3.5 New" feels weird -- is there a particular reason why Anthropic doesn't want to call this 3.6 (or even 3.5.1)?
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#77Is there an easy way to use Claude as a Co-Pilot in VS Code? If it is better at coding, it would be great to have it integrated.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#78Why not rev the numbers? "3.5" vs. "3.5 New" feels weird -- is there a particular reason why Anthropic doesn't want to call this 3.6 (or even 3.5.1)?
exactly my thought too, go up with the version number! Some negative examples: Claude Sonnet 3.5 for Workstations, Claude Sonnet 3.5 XP, Claude Sonnet 3.5 Max Pro, Claude Sonnet 3.5 Elite, Claude Sonnet 3.5 Ultra
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#79Fascinating. Though I expect people to be concerned about privacy implications of sending screenshots of the desktop, similar to the backlash Microsoft has received about their AI products. Giving the remote service actual control of the mouse and keyboard is a whole another level! But I am very excited about this in the context of accessibility. Screen readers and screen control software is hard to develop and hard…
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#80- AI Labs will eat some of the wrappers on top of their APIs - even complex ones like this. There are whole startups that are trying to build computer use.
- AI is fitting _some_ scaling law - the best models are getting better and the "previously-state-of-the-art" models are fractions of what they cost a couple years ago. Though it remains to be seen if it's like Moore's Law or if incremental improvements get harder and harder to make.