Completely irrelevant, and it might just be me, but I really like Anthropic's understated branding. OpenAI's branding isn't exactly screaming in your face either, but for something that's generated as much public fear/scaremongering/outrage as LLMs have over the last couple of years, Anthropic's presentation has a much "cosier" veneer to my eyes. This isn't the Skynet Terminator wipe-us-all-out AI, it's the adorable…
Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
121–130 of 758 posts
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#122Is there an easy way to use Claude as a Co-Pilot in VS Code? If it is better at coding, it would be great to have it integrated.
You can use it in Cursor - called "Cursor Tab" IMO Cursor Tab performs much better than Co-Pilot, easily works through things that would cause Co-Pilot to get stuck, you should give it a try
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#123From the computer use video demo, that's a lot of API calls. Even though Claude 3.5 Sonnet is relatively cheap for its performance, I suspect computer use won't be. It's a very good idea that Anthropic upfront that it isn't perfect. And it's guaranteed that there will be a viral story where Claude will accidentally delete something important with it. I'm more interested in Claude 3.5 Haiku, particularly if it is inde…
It's just bizarre to force a computer to go through a GUI to use another computer. Of course it's going to be expensive.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#124Earlier quoted context omitted.
Because its a finetune of 3.5 optimized for the use case of computer use. Its actually accurate and its not a 3.6.
So 3.5.1 ?
Stands out for me as I once replaced a 2.3 Turbo in a TurboCoupe with a 351 Windsor ))
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#125Earlier quoted context omitted.
With UIPath, Appian, etc. the whole field of RPA (robotic process automation) is a $XX billion industry that is built on that exact premise (that it's more feasible to do automation via GUIs than badly built/non-existing APIs). Depending on how many GUI actions correspond to one equivalent AI orchestrated API call, this might also not be too bad in terms of efficiency.
Most of the GUIs are Web pages, though, so you could just interact directly with an HTTP server and not actually render the screen. Or you could teach it to hack into the backend and add an API... Oh, and on edit, "bizarre" and "multi-billion-dollar-industry" are well known not to be mutually exclusive.
Most GUIs are in fact not web pages, that's a relatively newer development in the Enterprise side. So while some of them may be a web page, the goal is to be able to touch everything a user is doing in the workflow which very likely includes local apps.
This iteration from Anthropic is still engineering focused but you can see the future of this kind of tooling bypassing engineering/it teams entirely.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#126Earlier quoted context omitted.
So Opus that costs $15.00/$75.00 for 1mil tokens (input/output) is now worse than the model that costs $3.00/$15.00? That's according to https://docs.anthropic.com/en/docs/about-claude/models which has "claude-3-5-sonnet-20241022" as the latest model (today's date)
Yes, you will find similar things at essentially all other model providers. The older/bigger GPT4 runs at $30/$60 and peforms about on par with GPT4o-mini which costs only $0.15/$0.60. If you are currently, or have been integrating AI models in the past ~2 years, you should definitely keep up with model capability/pricing development. If you are staying on old models you are certainly overpaying/leaving performance o…
I don't think GPT-4o Mini has comparable performance to GPT-4 at all, where are you finding the benchmarks claiming this?
Everywhere I look says GPT-4 is more powerful, but GPT-4o Mini is most cost-effective, if you're OK with worse performance.
Even OpenAI themselves about GPT-4o Mini:
> Our affordable and intelligent small model for fast, lightweight tasks. GPT-4o mini is cheaper and more capable than GPT-3.5 Turbo.
If it was "on par" with GPT-4 they would surely say this.
> should definitely keep up with model capability/pricing development
Yeah, I mean that's why we're both here and why we're discussing this very topic, right? :D
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#127Earlier quoted context omitted.
For a company selling intelligence, that's a pretty stupid way of labelling a new product.
"computer use" is also as bad a marketing choice as possible for something that actually seems pretty cool.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#128Completely irrelevant, and it might just be me, but I really like Anthropic's understated branding. OpenAI's branding isn't exactly screaming in your face either, but for something that's generated as much public fear/scaremongering/outrage as LLMs have over the last couple of years, Anthropic's presentation has a much "cosier" veneer to my eyes. This isn't the Skynet Terminator wipe-us-all-out AI, it's the adorable…
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#129I have been a paying ChatGPT customer for a long time (since the very beginning). Last week I've compared ChatGPT to Claude and the results (to my eye) were better, the output better structured and the canvas works better. I'm on the edge of jumping ship.
> I'm on the edge of jumping ship. Yeah I think I might also jump ship. It’s just that chatGPT now kinda knows who I am and what I like and I’m afraid of losing that. It’s probably not a big deal though.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#130They need to get the price of 3.5 Haiku down. It's about 2x 4o-mini.