Live data from Hacker News

How Google’s Web Crawler Bypasses Paywalls

elaineou.com

1–10 of 243 posts

Re: How Google’s Web Crawler Bypasses Paywalls

#4
Doesn't this kind of also hurt SEO? I'm would guess Google has some automated system to detect and apply a negative signal to sites that provide different content to a Googlebot user agent than a non-Googlebot user agent. I guess these sites are counting that the other signals outweigh that negative hit.

Otherwise, why would expertsexchange be obligated to provide the answers at the very bottom? Did something change?

Re: How Google’s Web Crawler Bypasses Paywalls

#9
post #7

If they're now blocking clicks from Google, doesn't that mean that they're cloaking and violating the Google's Webmaster Guidelines [1]? [1]: https://support.google.com/webmasters/answer/66355?hl=en

Google is not okay with cloaking, but they will whitelist publishers if the publisher specifically includes a parameter that declares if the site requires registration or subscription. This is done in the sitemap.

https://support.google.com/news/publisher/answer/74288?hl=en

Re: How Google’s Web Crawler Bypasses Paywalls

#10
post #4

Doesn't this kind of also hurt SEO? I'm would guess Google has some automated system to detect and apply a negative signal to sites that provide different content to a Googlebot user agent than a non-Googlebot user agent. I guess these sites are counting that the other signals outweigh that negative hit. Otherwise, why would expertsexchange be obligated to provide the answers at the very bottom? Did something change?

I'm 99% sure I've encountered a Googlebot crawling pages with the UA of a regular browser, presumably for exactly this purpose.
Post reply on HN