Live data from Hacker News

Claude for Chrome

anthropic.com

211–220 of 433 posts

Re: Claude for Chrome

#211

According to their own blog post, even after mitigations, the model still has an 11% attack success rate. There's still no way I would feel comfortable giving this access to my main browser. I'm glad they're sticking to a very limited rollout for now. (Sidenote, why is this page so broken? Almost everything is hidden.)

11% success rate for what is effectively a spear-phishing attempt isn't that terrible and tbh it'll be easier to train Claude not to get tricked than it is to train eg my parents.

>Claude not to get tricked than it is to train eg my parents.

One would think but apparently from this blog post it is still succeptible to the same old prompt injections that have always been around. So I'm thinking it is not very easy to train Claude like this at all. Meanwhile with parents you could probably eliminate an entire security vector outright if you merely told them "bank at the local branch," or "call the number on the card for the bank don't try and look it up."

Re: Claude for Chrome

#212

Earlier quoted context omitted.

> It's clear to me that the tech just isn't there yet. Totally agree. This was the thesis behind MCP-B (now WebMCP https://github.com/MiguelsPizza/WebMCP ) HN Post: https://news.ycombinator.com/item?id=44515403 DOM and visual parsing are dead ends for browser automation. Not saying models are bad; they are great. The web is just not designed for them at all. It's designed for humans, and humans, dare I say, are prett…

I suspect this kind of framework will be adopted by websites with income streams that are not dependent on human attention (i.e. advertising revenue, mostly). They have no reason to resist LLM browser agents. But if they’re in the business of selling ads to human eyeballs, expect resistance. Maybe the AI companies will find a way to resell the user’s attention to the website, e.g. “you let us browse your site with an…

Even the websites whose primary source of revenue is not ad impressions might be resistant to let the agents be the primary interface through which users interact with their service.

Instacart currently seems to be very happy to let ChatGPT Operator use its website to place an order (https://www.instacart.com/company/updates/ordering-groceries...) [1]. But what happens when the primary interface for shopping with Instacart is no longer their website or their mobile app? OpenAI could demand a huge take rate for orders placed via ChatGPT agents, and if they don't agree to it, ChatGPT can strike a deal with a rival company and push traffic to that service instead. I think Amazon is never going to agree to let other agents use its website for shopping for the same reason (they will restrict it to just Alexa).

[1] - the funny part is the Instacart CEO quit shortly after this and joined OpenAI as CEO of Applications :)

Re: Claude for Chrome

#213

I don’t know if this will make anything better. Internet is now filled with ai generated text, picture or videos. Like we havent had enough already, it is becaming more and more. We make ai agents to talk to each other. Someone will make ai to generate a form, many other will use ai to fill that form. Even worst, some people will fill millions of forms in matter of second. What is left is the empty feeling of having…

I am starting to see this age of internet-for-robots-by-robots as our second chance to detach from those devices and start living irl again.

The subtext is the one technology capable of potentially rallying, unifying, and mobilizing the working class across the globe is lost in this design. Probably intentionally. A shame we couldn't rise up and do something about wealth distribution before the powers that be that maintain the world's status quo locked it down.

Re: Claude for Chrome

#214

It's wild to see an AI company put out a press release that is basically "hey, you kids wanna see a loaded gun?" Normally all their public coms are so full of optimism and salesmanship around the potential. They are fully aware of how dangerous this is.

This is what they need for the next generation of models. The key line is: > We view browser-using AI as inevitable: so much work happens in browsers that giving Claude the ability to see what you're looking at, click buttons, and fill forms will make it substantially more useful. A lot of this can be done by building a bunch of custom environments at training time, but only a limited number of usecases can be handle…

I don’t get the argument. Why is the loaded foot gun better in the hands of “select” customers better than in the hands of self selecting group of beta testers?

Re: Claude for Chrome

#215

I don’t know if this will make anything better. Internet is now filled with ai generated text, picture or videos. Like we havent had enough already, it is becaming more and more. We make ai agents to talk to each other. Someone will make ai to generate a form, many other will use ai to fill that form. Even worst, some people will fill millions of forms in matter of second. What is left is the empty feeling of having…

It’s wild to me that people see this as bad. The point of the form is not in the filling. You shouldn't want to fill out a form. If you could accomplish your task without the busywork, why wouldn’t you? If you could interact with the world on your terms, rather than in the enshitified way monopoly platforms force on you, why wouldn't you? And yeah, if you could consume content in the way you want, rather than the way…

you are getting this from the wrong perspective. I agree what you say here, but things you are listing here implies one thing;

"you didnt want to do this before, now with the help of ai, you dont have to. you just live your life as the way you want"

and your assumption is wrong. I still want to watch videos when it is generated by human. I still want to use internet, but when I know it is a human being at the other side. What I don't want is AI to destroy or make dirty the things I care, I enjoy doing. Yes, I want to live in my terms, and AI is not part of it, humans do.

I hope it is clear.

Re: Claude for Chrome

#217
post #26

Personally, the only way I’m going to give an LLM access to a browser is if I’m running inference locally. I’m sure there’s exploits that could be embedded into a model that make running locally risky as well, but giving remote access to Anthropic, OpenAI, etc just seems foolish. Anyone having success with local LLMs and browser use?

The primary risk with these browser agents is prompt injection attacks. Running it locally doesn't help you in that regard.

Re: Claude for Chrome

#218

I don’t know if this will make anything better. Internet is now filled with ai generated text, picture or videos. Like we havent had enough already, it is becaming more and more. We make ai agents to talk to each other. Someone will make ai to generate a form, many other will use ai to fill that form. Even worst, some people will fill millions of forms in matter of second. What is left is the empty feeling of having…

I am starting to see this age of internet-for-robots-by-robots as our second chance to detach from those devices and start living irl again.

I really wish, but I doubt that. I will definitely move to that direction though. I am a professional software engineer, and seriously considering doing another job.

not because AI can take over my job or something, hell no it can't, at least for now. but day by day I am missing the point of being an engineer. problem solving, building and seeing that it works. the joy of engineering is almost gone. Personally, I am not satisfied with my job as I used to do, and that is really bothering.

Re: Claude for Chrome

#219

TikTokification of the browser by AI is the killer feature, not writing an email. When on a page it automatically suggests the next site(s) to visit based on my history and the page I am on. And when I say killer, this kills google search by pivoting away from the urlbar and provides a new space to put ads. Spent years in the browser space, on Chrome, DDG, Blackberry and more developing browsers, prototype browser an…

TikTokification is an odd example to pick here, given that TikTok is a platform which didn't kill its Google competitor YouTube.

What do you mean? Youtube ticktocked itself complete with shoehorning vertical videos on the desktop experience.

Re: Claude for Chrome

#220
post #210

Earlier quoted context omitted.

I have built a custom "deep research" internally that uses puppeteer to find business information, tech stack and other information about a company for our sales team. My experience was that giving the LLM a very limited set of tools and no screenshots worked pretty damn well. Tbf for my use case I don't need more interactivity than navigate_to_url and click_link. Each tool returning a text version of the page and th…

Seems navigate_to_url and click_link would be solved with just a script running puppeteer vs having an llm craft a puppeteer script to hopefully do this simple action reliably? What is the great advantage with the llm tooling in this case?

Oh the tools are hand coded (or rather built with Claude Code) but the agent can call them to control the browser.

Imagine a prompt like this:

You are a research agent your goal is to figure out this companies tech stack: - Company Name

Your available tools are: - navigate_to_url: use this to load a page e.g. use google or bing to search for the company site It will return the page content as well as a list of available links - click_link: Use this to click on a specific link on the currently open page. It will also return the current page content and any available links

A good strategy is usually to go on the companies careers page and search for technical roles.

This is a short form of what is actually written there but we use this to score leads as we are built on postgres and AWS and if a company is using those, these are very interesting relevancy signals for us.

Post reply on HN