Google really doesn't have a leg to stand on here. They scrape the Internet. They replace content against the wishes of users multiple different times, such as with AMP. Their entire business model recently has been to provide you answers they learned from scraping your website and now they want to sue other people who are doing the same. Data wants to be free. They knew that once. EDIT: Also to be clear I am not say…
As the post says, Google only scrapes the websites that want to be scraped. Sure, it's opt-out (via robots.txt) rather than opt-in, but they do give you a choice. You can even decide between no scraping at all and opting out on a per-scraper basis, and Google will absolutely honor your preferences in that regard. SERP API just assumes everybody wants to be scraped, and doesn't give you a choice. (whether websites sho…
Why we're taking legal action against SerpApi's unlawful scraping
81–90 of 113 posts
Re: Why we're taking legal action against SerpApi's unlawful scraping
#82https://blog.cloudflare.com/perplexity-is-using-stealth-unde...
Re: Why we're taking legal action against SerpApi's unlawful scraping
#83Google really doesn't have a leg to stand on here. They scrape the Internet. They replace content against the wishes of users multiple different times, such as with AMP. Their entire business model recently has been to provide you answers they learned from scraping your website and now they want to sue other people who are doing the same. Data wants to be free. They knew that once. EDIT: Also to be clear I am not say…
Unfortunately they do have a couple of points that may prove salient (though I fully agree about them being scrapers also). You can search Google _for free_ (with all the caveats of that statement), part of their grievance is that serpapi use the scraped data as a paid for service Lots of Google bot blocking is also circumvented, which they seem to have made a lot of efforts towards in the past year - robots.txt dire…
I thought the ads counted as payment? That seems to be the logic used to take technical measures against adblockers on YouTube while pushing users towards a paid ad-free subscription, at least.
If viewing ads is payment, then Google isn't a free service. If viewing ads isn't payment, then Google should have no problem with people using adblockers.
Re: Why we're taking legal action against SerpApi's unlawful scraping
#84Google really doesn't have a leg to stand on here. They scrape the Internet. They replace content against the wishes of users multiple different times, such as with AMP. Their entire business model recently has been to provide you answers they learned from scraping your website and now they want to sue other people who are doing the same. Data wants to be free. They knew that once. EDIT: Also to be clear I am not say…
Eh, and in 20 if SerpApi or whatever the fuck becomes the next google, they’ll have a blog post titled “Why we’re taking legal action against BlemFlamApi data collection”. The biggest joke was all the “hackers” 25 years ago shouting “Don’t be evil like Oracle, Microsoft, Apple or Adobe and charge for your software, be good like Google and just put like a banner ad or something and give it away for free”
Re: Why we're taking legal action against SerpApi's unlawful scraping
#85Earlier quoted context omitted.
They're in a unique position where many people allow googlebot but try to block most other bots
Allow for the purpose of indexing, not training models. Like if you give a friend a key to your house so they can check on your plants when you're out of town but they throw a rager and trash the place.
That was not a phrase I expected to read on Hacker News! Haven't heard it since I was about 13. I always assumed it was a Scottish phrase.
Re: Why we're taking legal action against SerpApi's unlawful scraping
#86"Google follows industry-standard crawling protocols, and honors websites’ directives over crawling of their content." Is that true with how they trained Gemini? Doesn't everyone with a foundational model scrape the web relentlessly without regard for robots.txt?
Re: Why we're taking legal action against SerpApi's unlawful scraping
#87Earlier quoted context omitted.
[flagged]
Scraping for search engines is different. They need to scrape the data to build the index, otherwise the search wouldn’t work.
Being used as AI training data provides negative value for a website owner, as it takes traffic away.
It's the difference between a movie review, and a ripped torrent.
Re: Why we're taking legal action against SerpApi's unlawful scraping
#88Google really doesn't have a leg to stand on here. They scrape the Internet. They replace content against the wishes of users multiple different times, such as with AMP. Their entire business model recently has been to provide you answers they learned from scraping your website and now they want to sue other people who are doing the same. Data wants to be free. They knew that once. EDIT: Also to be clear I am not say…
Re: Why we're taking legal action against SerpApi's unlawful scraping
#89They also started caring about this, probably because they don't want their competitors to get the same data as they have.
Re: Why we're taking legal action against SerpApi's unlawful scraping
#90Google really doesn't have a leg to stand on here. They scrape the Internet. They replace content against the wishes of users multiple different times, such as with AMP. Their entire business model recently has been to provide you answers they learned from scraping your website and now they want to sue other people who are doing the same. Data wants to be free. They knew that once. EDIT: Also to be clear I am not say…
Unfortunately they do have a couple of points that may prove salient (though I fully agree about them being scrapers also). You can search Google _for free_ (with all the caveats of that statement), part of their grievance is that serpapi use the scraped data as a paid for service Lots of Google bot blocking is also circumvented, which they seem to have made a lot of efforts towards in the past year - robots.txt dire…
Well not through their API which you do need to pay for and is a paid service.