Live data from Hacker News

Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

anthropic.com

651–660 of 758 posts

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#651

Earlier quoted context omitted.

What if a user identifies as Claude too?

* Implemented inverse CAPTCHA using invisible Unicode characters and alpha-channel encoded image data to tell models and human impostors apart.

* The end state here is that SMBC comic: https://www.smbc-comics.com/comic/captcha

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#652

Earlier quoted context omitted.

It is. CTO of healthcare org here. I just put a hold on a new RPA project to keep an eye on this and see how it develops. According to their docs, Anthropic will sign a BAA.

What is a BAA?

https://www.techtarget.com/healthtechsecurity/feature/What-I... agreement that lets a business associate handle HIPAA-protected data.

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#654

Earlier quoted context omitted.

While we expect this capability to improve rapidly in the coming months, Claude's current ability to use computers is imperfect. Some actions that people perform effortlessly—scrolling, dragging, zooming—currently present challenges for Claude and we encourage developers to begin exploration with low-risk tasks.

Can someone please try this on a MAC/OS and just 100% verify if this puppy can scroll or not? thnks

It does in the video. Just not the spreadsheet at the start.

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#656
post #369

I really don't get their model. They have very advanced models, but the service overall seems to be a jumble of priorities. Some examples: Anthropic doesn't offer an unlimited chatbot service, only plans that give you "more" usage, whatever that means. If you have an API key, you are "unlimited," so they have the capability. Why doesn't the chatbot allow one to use their API key in the Claude app to get unlimited usa…

> Anthropic doesn't offer an unlimited chatbot service,

because its expensive, if they give unlimited service someone will misuse it

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#657

Earlier quoted context omitted.

> There is nothing magical or special about human brain. There is a lot about the human brain that even the world's top neuroscientists don't know. There's plenty of magic about it if we define magic as undiscovered knowledge. There's also no consensus among top AI researchers that current techniques like LLMs will get us anywhere close to AGI. Nothing I've seen on current models (not even o1-preview) suggests to me…

Defining AGI as “can reason about 5MLOC” is ridiculous. When do the goal posts stop moving? When a computer can solve time travel? Babies have behavior all the time that is no more differentiable from what an LLM does on a normal basis (including terrible logic and hallucinations). The majority of people on the planet can barely reason about how any given politician will affect them, even when there’s a billion resou…

> When do the goal posts stop moving?

When someone comes up with a rigorous definition of intelligence.

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#658
post #22
post #9

I still feel like the difference between Sonnet and Opus is a bit unclear. Somewhere on Anthropic's website it says that Opus is the most advanced, but on other parts it says Sonnet is the most advanced and also the fastest. The UI doesn't make the distinction clear either. Then on Perplexity, Perplexity says that Opus is the most advanced, compared to Sonnet. And finally, in the table in the blogpost, Opus isn't eve…

Opus is a larger and more expensive model. Presumably 3.5 Opus will be the best but it hasn't been released. 3.5 Sonnet is better than 3.0 Opus kind of like how a newer i5 midrange processor is faster and cheaper than an old high-end i7.

Makes me wonder if perhaps they do have 3.5 Opus trained, but that they're not releasing it because 3.5 Sonnet is already enough to beat the competition, and some combination of "don't want to contribute to an arms race" and "it has some scary capabilities they weren't sure were ready to publish yet".

Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku

#659
If "computer use" feature is able to find it's way in Azure, AAD/Entra, SharePoint settings, etc. - it has a chance of becoming a better user interface for Microsoft products. :)

Can you imagine how simple the world would be if you'd just need to tell Claude: "user X needs to have access to feature Y, please give them the correct permissions", with no need to spend days in AAD documentation and the settings screens maze. I fear AAD is AI-proof, though :)

Post reply on HN