I agree. It's ironic they protect their page so aggressively against scraping, being the biggest scraper in the universe at the same time.
Often they ignore robots.txt also, they still continue scraping, sometimes they even publish things in their SERPS.
From Google's FAQ:
However, robots.txt Disallow does not guarantee that a page will not appear in results: Google may still decide, based on external information such as incoming links, that it is relevant. If you wish to explicitly block a page from being indexed, you should instead use the noindex robots meta tag or X-Robots-Tag HTTP header. In this case, you should not disallow the page in robots.txt, because the page must be crawled in order for the tag to be seen and obeyed.