Completely irrelevant, and it might just be me, but I really like Anthropic's understated branding. OpenAI's branding isn't exactly screaming in your face either, but for something that's generated as much public fear/scaremongering/outrage as LLMs have over the last couple of years, Anthropic's presentation has a much "cosier" veneer to my eyes. This isn't the Skynet Terminator wipe-us-all-out AI, it's the adorable…
As a Kurt Vonnegut fan, their asterisk logo on claude.ai always amuses me. It must be intentional: https://en.m.wikipedia.org/wiki/File:Claude_Ai.svg https://www.redmolotov.com/vonnegut-ahole-tshirt
Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
681–690 of 758 posts
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#682Earlier quoted context omitted.
> There is nothing magical or special about human brain. There is a lot about the human brain that even the world's top neuroscientists don't know. There's plenty of magic about it if we define magic as undiscovered knowledge. There's also no consensus among top AI researchers that current techniques like LLMs will get us anywhere close to AGI. Nothing I've seen on current models (not even o1-preview) suggests to me…
Defining AGI as “can reason about 5MLOC” is ridiculous. When do the goal posts stop moving? When a computer can solve time travel? Babies have behavior all the time that is no more differentiable from what an LLM does on a normal basis (including terrible logic and hallucinations). The majority of people on the planet can barely reason about how any given politician will affect them, even when there’s a billion resou…
I haven't moved any goal posts - it is your definition which is way too narrow.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#683Earlier quoted context omitted.
This has existed for a long time, it's called "RPA" or Robotic Process Automation. The biggest incumbent in this space is UiPath, but there are a host of startups and large companies alike that are tackling it. Most of the things that RPA is used for can be easily scripted, e.g. download a form from one website, open up Adobe. There are a lot of startups that are trying to build agentic versions of RPA, I'm glad to s…
I was going to comment about this. Worked at a place that had a “Robotics Department”, wow I thought. Only to find out it was automating arcane software. UI is now much more accessible as API. I hope we don’t start seeing captcha like behaviour in desktop or web software.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#684Earlier quoted context omitted.
As a Kurt Vonnegut fan, their asterisk logo on claude.ai always amuses me. It must be intentional: https://en.m.wikipedia.org/wiki/File:Claude_Ai.svg https://www.redmolotov.com/vonnegut-ahole-tshirt
It's also a joke in the TV show "Community": https://www.youtube.com/watch?v=HP1Atb8nAGY
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#685Earlier quoted context omitted.
This has existed for a long time, it's called "RPA" or Robotic Process Automation. The biggest incumbent in this space is UiPath, but there are a host of startups and large companies alike that are tackling it. Most of the things that RPA is used for can be easily scripted, e.g. download a form from one website, open up Adobe. There are a lot of startups that are trying to build agentic versions of RPA, I'm glad to s…
Honestly, this is going to be huge for healthcare. There's an incredible amount of waste due to incumbent tech making interoperability difficult.
Similarly I expect that once processing/searching laws/legal records becomes easy through LLMs, we'll compensate by having orders of magnitude more laws, perhaps themselves generated in part by LLMs.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#686Earlier quoted context omitted.
> There's an incredible amount of waste due to incumbent tech making interoperability difficult. So the solution to that is to add another layer of complex AI tech on top of it?
Well nothing else we've tried has worked.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#687Earlier quoted context omitted.
Exactly. I have been wondering for a while how GenAI might upend RPA providers guess this might be the answer.
I've been wondering the same and started exploring building a startup around this idea. My analysis led me to the conclusion that if AI gets even just 2 orders of magnitude better over the next two years, this will be "easy" and considered table stakes. Like connecting to the internet, syncing with cloud or using printer drivers I don't think there will be a very big place for standalone next gen RPA pure plays. it m…
You realize this means '100 times better', right?
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#688Earlier quoted context omitted.
That's exactly it. I've been peddling my vision of "AI automation" for the last several months to acquaintances of mine in various professional fields. In some cases, even building up prototypes and real-user testing. Invariably, none have really stuck. This is not a technical problem that requires a technical solution. The problem is that it requires human behavior change. In the context of AI automation, the promis…
There’s nothing to gain for anyone there. Workers will lose their jobs, and managers will lose their reports.
Nobody likes to change a system where they already have their own little comfortable spot and figured it out and just want to seep in the lukewarm there until retirement. Fully understandable. But at least in the private sector this will not save them.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#689I still feel like the difference between Sonnet and Opus is a bit unclear. Somewhere on Anthropic's website it says that Opus is the most advanced, but on other parts it says Sonnet is the most advanced and also the fastest. The UI doesn't make the distinction clear either. Then on Perplexity, Perplexity says that Opus is the most advanced, compared to Sonnet. And finally, in the table in the blogpost, Opus isn't eve…
Opus hasn't yet gotten an update from 3 to 3.5, and if you line up the benchmarks, the Sonnet "3.5 New" model seems to beat it everywhere. I think they originally announced that Opus would get a 3.5 update, but with every product update they are doing I'm doubting it more and more. It seems like their strategy is to beat the competition on a smaller model that they can train/tune more nimbly and pair it with outside-…
This theory is consistent with the other two top players, Open AI and Google, they both were expected to release a heavy model, but instead have just released multiple medium and small tier models. It's been so long since google released gemini ultimate 1.0 (the naming clearly implying that they were planning on upgrading it to 1.5 like they did with Pro)
Not seeing anyone release a heavyweight model, but at the same time releasing many small and medium sized models makes me think that improving models will be much more complicated than scaling it with more compute, and that there likely are diminishing returns with that regard.
Re: Computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku
#690Earlier quoted context omitted.
By reputation -- I can't vouch for this personally, and I don't know if it'll still be true with this update -- Opus is still often better for things like creative writing and conversations about emotional or political topics.
Yes, (old) 3.5 Sonnet is distinctly worse at emotional intelligence, flexibility, expressiveness and poetry.