Live data from Hacker News

Google Removed 749M Anna's Archive URLs from Its Search Results

torrentfreak.com

21–30 of 168 posts

Re: Google Removed 749M Anna's Archive URLs from Its Search Results

#21

Feels weird to say but I have found using Yandex of all places an excellent search engine for content that get taken down by DMCA requests. Eg if you want to watch a movie that's not on Netflix using a web stream the search results are far better. Feels like Google circa 2005.

I've been playing around with a variety of search engines such as Kagi, Startpage, Ecosia, DDG.

All of them are better than google in finding relevant results. Lol

Google is way too "personalized".

Re: Google Removed 749M Anna's Archive URLs from Its Search Results

#22
post #21

Feels weird to say but I have found using Yandex of all places an excellent search engine for content that get taken down by DMCA requests. Eg if you want to watch a movie that's not on Netflix using a web stream the search results are far better. Feels like Google circa 2005.

I've been playing around with a variety of search engines such as Kagi, Startpage, Ecosia, DDG. All of them are better than google in finding relevant results. Lol Google is way too "personalized".

You can turn off personalization. (Operating under the assumption that most people search for facts, I personally don't see why one would ever want personalized results.)

Re: Google Removed 749M Anna's Archive URLs from Its Search Results

#24
post #21

Earlier quoted context omitted.

I've been playing around with a variety of search engines such as Kagi, Startpage, Ecosia, DDG. All of them are better than google in finding relevant results. Lol Google is way too "personalized".

You can turn off personalization. (Operating under the assumption that most people search for facts, I personally don't see why one would ever want personalized results.)

> I personally don't see why one would ever want personalized results.

The same short combination of words can mean very different things to different people. My favorite example of this is "C string" because when I was a kid learning C I was introduced to a whole new class of lingerie because Google didn't really personalize results back then. Now when I search "C string" Google knows exactly what I mean.

Re: Google Removed 749M Anna's Archive URLs from Its Search Results

#25
post #21

Earlier quoted context omitted.

I've been playing around with a variety of search engines such as Kagi, Startpage, Ecosia, DDG. All of them are better than google in finding relevant results. Lol Google is way too "personalized".

You can turn off personalization. (Operating under the assumption that most people search for facts, I personally don't see why one would ever want personalized results.)

I won't bother defending Google-style personalization as it exists for their search results, but since collisions in terminology across fields are common, it's not that hard to see how actual, thoughtful personalization could be useful. Someone searching for "Kafka" is going to want very different results based on whether they're thinking of software or literature. Opinions may also differ over the usefulness of sources, even for people ultimately interested primarily in facts; I find Kagi-style personalization (make your own domain list) very useful, but across Kagi's userbase Reddit is simultaneously one of the most lowered, most raised, and most pinned domains: https://kagi.com/stats?stat=leaderboard

Re: Google Removed 749M Anna's Archive URLs from Its Search Results

#27
post #14
post #8

Earlier quoted context omitted.

I’m pretty sure Google indexing pages from Anna’s archive would only get metadata, because AA doesn’t have the full text of the books on those pages. I think to get the full text you have to download the torrents, and I don’t think Google was doing that.

No, thats more meta's trick. and they were "only doing it for the articles" not the pictures. I think. I dunno..

They were doing it for the videos too, but only for "personal use"...

https://www.wired.com/story/meta-claims-downloaded-porn-at-c...

Re: Google Removed 749M Anna's Archive URLs from Its Search Results

#28
post #21

Earlier quoted context omitted.

I've been playing around with a variety of search engines such as Kagi, Startpage, Ecosia, DDG. All of them are better than google in finding relevant results. Lol Google is way too "personalized".

You can turn off personalization. (Operating under the assumption that most people search for facts, I personally don't see why one would ever want personalized results.)

Location based personalization is pretty useful - if I search for 'Bob's Discount Linguine' I want the one in my neighborhood.

Lots of niche things (like programming) also reuse common english words to mean specific things - if I search e.g. 'locking' it's nice to get results related to asynchronous programming instead of locksmiths because google knows I regularly search for programming related terminology.

Of course it's questionable whether google does a good job at any of this, but I absolutely see the value.

Re: Google Removed 749M Anna's Archive URLs from Its Search Results

#29

Earlier quoted context omitted.

You can turn off personalization. (Operating under the assumption that most people search for facts, I personally don't see why one would ever want personalized results.)

I won't bother defending Google-style personalization as it exists for their search results, but since collisions in terminology across fields are common, it's not that hard to see how actual, thoughtful personalization could be useful. Someone searching for "Kafka" is going to want very different results based on whether they're thinking of software or literature. Opinions may also differ over the usefulness of sour…

> Kafka" is going to want very different results based on whether they're thinking of software or literature.

Speak for yourself. I've worked in several "Kafka-esque" software organizations.

Re: Google Removed 749M Anna's Archive URLs from Its Search Results

#30
I was surprised that those pages showed up in book title searches at all. Makes sense to get rid of them, you don't want a search for a book to be topped by a link to pirate the book. The top-level domains still come up, and people who know they want to pirate a book can still find the site.
Post reply on HN