Live data from Hacker News

Claude can now search the web

anthropic.com

221–230 of 758 posts

Re: Claude can now search the web

#221
post #161

Earlier quoted context omitted.

How long have you been using Google search for? It used to be SO much less likely to return junk.

Since around 2012. What year would be the golden age of google search? I wonder if anyone has archived search result pages for relatively timeless queries so that we could compare. Wayback Machine seems to archive some of them. https://web.archive.org/web/20200801000000*/https://www.goog... https://web.archive.org/web/20200801000000*/https://www.goog...

Around 2011 to 2012 after the first of many updates with names like hurricanes came and washed away the good.

Re: Claude can now search the web

#222
post #102

Searching the web is a great feature in theory, but every implementation I've used so far looks at the top X hits and then interprets it to be the correct answer. When you're talking to an LLM about popular topics or common errors, the top results are often just blogspam or unresolved forum posts, so the you never get an answer to your problem. More of an indicator that web search is more unusable than ever, but inte…

Exa (YC S21) is trying to solve this problem by re-indexing the web in an LLM-friendly way.

2021? how are they doing?

Re: Claude can now search the web

#223

Earlier quoted context omitted.

Google search is crap. It seems to be a sentiment among many HNers, but is it really that bad? I mostly use it for programming, so documentation/forums and it works out greatly. For some queries it even returns personal blogs (which people seem to bash google for not happening). Of course there are some queries that return purely AI blogspam, but reformatting the query with a bit more thought usually solves it. I won…

Is google search bad? Click here to find ten reasons why it is bad and 10 reasons why you should still use it. Yes, it is that bad. Website of Nike? Website of Starbucks? Likely position number one. Every product, category etc., e.g. what rice cooker should I buy? Is diseased by link and affiliate spam. There is a reason why people put +reddit on search terms.

I think part of the reason for this is that web site developers have got out of the habit of optimizing for search engines. I'm often surprised by how self-contained the requirements for a website are now, even among otherwise technically sophisticated clients. There'll be a beautiful site in React that absolutely sucks for SEO, but no-one will mind because a) it's unclear how big an audience there should be for the site, and b) the "all your hits come from search engines" was broken ten or more years ago by social network linking, so the question of how you get an audience seems much more arbitrary, and less connected to google.com.

Re: Claude can now search the web

#224
i stopped using Claude about 2 months ago. went to Grok (the code was better, everything was better - politics aside). i wonder if this update will improve it.

the main issue i find with Claude is, he fights you. He refuses so many requests and i need 3 or 4 replies to get what i want vs deepseek/grok. i've kept the monthly subscription to help anthropic, but it's trounced by the free options imo.

Re: Claude can now search the web

#226

Earlier quoted context omitted.

Google will still scrape it for training data either way, this only impacts search results.

> Today we’re announcing Google-Extended, a new control that web publishers can use to manage whether their sites help *improve Bard and Vertex AI generative APIs*, including future generations of models that power those products.

https://www.theverge.com/news/630079/openai-google-copyright...

they're literally asking to break laws to train AI for national security. A sentence in a press release from 2 years ago is worthless... look at what they're actually doing

Re: Claude can now search the web

#227
post #102

Searching the web is a great feature in theory, but every implementation I've used so far looks at the top X hits and then interprets it to be the correct answer. When you're talking to an LLM about popular topics or common errors, the top results are often just blogspam or unresolved forum posts, so the you never get an answer to your problem. More of an indicator that web search is more unusable than ever, but inte…

Ugh, what a nightmare, now search engines are going to start optimizing for bots.

Re: Claude can now search the web

#228
post #102

Searching the web is a great feature in theory, but every implementation I've used so far looks at the top X hits and then interprets it to be the correct answer. When you're talking to an LLM about popular topics or common errors, the top results are often just blogspam or unresolved forum posts, so the you never get an answer to your problem. More of an indicator that web search is more unusable than ever, but inte…

Do you think that if it's a non-Google company, that maybe doesn't rank search by ad payment $$$, that this new company could in theory do a better job?

Re: Claude can now search the web

#229
post #102

Searching the web is a great feature in theory, but every implementation I've used so far looks at the top X hits and then interprets it to be the correct answer. When you're talking to an LLM about popular topics or common errors, the top results are often just blogspam or unresolved forum posts, so the you never get an answer to your problem. More of an indicator that web search is more unusable than ever, but inte…

Is there any viable alternative to pass knowledge to the LLMs that goes beyond their training cut off date?

Re: Claude can now search the web

#230
post #200

Earlier quoted context omitted.

If this feature isn’t already part of the Claude API it likely will be at some point, in which case many Claude requests will be automated with no way to distinguish between user-driven or otherwise.

Simply put, at the end of the day you lose, AI blocking will not work. I mean, currently the AI request comes from the datacenter running the AI, but eventually one of two things will happen. AI models will get small/fast enough to run on user hardware and use the users resources: End result? You lose. The user will set their own headers and sites will play the impossible game of identifying AI. AI sites will figure…

Welcome to the world of CAPTCHAs
Post reply on HN