Live data from Hacker News

Suddenly, Hacker News is not the first result for 'Hacker News'

google.com

211–216 of 216 posts

Re: Suddenly, Hacker News is not the first result for 'Hacker News'

#211

Earlier quoted context omitted.

seriously. i can't get a hospital to show up in google maps... no human for me to talk to. HN is number 4 instead of 1 google page one, they're right on it.

Which hospital? Seems like this thread has at least one pair of google-eyeballs looking at it.

i'd rather not drop the name here. but it's a verified listing and i've used the "report a problem" 3 times now. If they're not checking that... :(

Re: Suddenly, Hacker News is not the first result for 'Hacker News'

#212

Earlier quoted context omitted.

seriously. i can't get a hospital to show up in google maps... no human for me to talk to. HN is number 4 instead of 1 google page one, they're right on it.

You can fix this yourself. In the lower-right corner of Google Maps there is a tiny link that says, "Edit in Google Map Maker". Click this link and you can edit Google Maps. Your edits get sent to Google and they'll approve/deny it in typically a few days.

it's a verified listing (do you know hard it is to convince the IT dept to take their automated phone system offline so i could verify a google maps listing? It was nuts and no google didn't offer the postcard method) so i don't see why i have to enter the same info again, but i did.

the listing only shows up if you type the exact name of the hospital into the search bar, which is useless.

Re: Suddenly, Hacker News is not the first result for 'Hacker News'

#213
Bottom line is that if you don't trust Google & use their tools (GWT) then you can end up in sticky situations like this one.

My experience with the Crawl Rate feature via GWT is that they do honour it pretty strictly, but for large sites Gbot can cause a lot of extra load even if pages are static.

A good CDN and stateless cache server will help but for sites as large as HN every request adds up!

Re: Suddenly, Hacker News is not the first result for 'Hacker News'

#214

I think I know what the problem is; we're detecting HN as a dead page. It's unclear whether this happened on the HN side or on Google's side, but I'm pinging the right people to ask whether we can get this fixed pretty quickly. Added: Looks like HN has been blocking Googlebot, so our automated systems started to think that HN was dead. I dropped an email to PG to ask what he'd like us to do.

what about the massive amount of duplicate content indexed via this domain: hackerne.ws ?

Re: Suddenly, Hacker News is not the first result for 'Hacker News'

#215

Earlier quoted context omitted.

I would think that the number of people who (a) know how to create a valid robot.txt file, (b) have some idea of how to use the "crawl-delay" directive and (c) write a "shoot-themselves-in-the-foot" worthy error is vanishingly small.

I alluded to some of the ways that I've seen people shoot themselves in the foot in a blog post a few years ago: http://www.mattcutts.com/blog/the-web-is-a-fuzz-test-patch-y... "You would not believe the sort of weird, random, ill-formed stuff that some people put up on the web: everything from tables nested to infinity and beyond, to web documents with a filetype of exe, to executables returned as text documents. In…

"...everything from tables nested to infinity..."

The irony of that statement on hacker news is pretty amazing. Have you looked at how the threads are rendered on this page. It is tables all the way down.

Re: Suddenly, Hacker News is not the first result for 'Hacker News'

#216

This is an example of why Goog's search algorithms (and others') should be open: http://news.ycombinator.com/item?id=3268371 A subtle attack may be by making bots stop indexing it or using SEO practices to lower it enough so it would become unsearchable, and therefore, non-existent. Or just crack into Google...

UPDATE: http://www.itworld.com/software/228393/free-software-activis...
Post reply on HN