The Ultimate Guide To Duplicate Content
david-whitehouse.org
The Ultimate Guide To Duplicate Content
1–10 of 11 posts
Re: The Ultimate Guide To Duplicate Content
#2I've heard this before, but I have a hard time believing it. Google really can't figure out on its own that for some sites the www is optional?
Re: The Ultimate Guide To Duplicate Content
#3"www vs non-www" I've heard this before, but I have a hard time believing it. Google really can't figure out on its own that for some sites the www is optional?
Besides, why bother risking it when it's a 5 minute job?
Re: The Ultimate Guide To Duplicate Content
#4mixedcase urls (on a windows server), ending slash vs no ending slash, double slashes in url, too much border (nav, footer, right hand side) html on pages with otherwise minor content, indexed ip address urls, additional url (i.e.: tracking parameters), .... and these are just on top of my head
Re: The Ultimate Guide To Duplicate Content
#5missed a few: mixedcase urls (on a windows server), ending slash vs no ending slash, double slashes in url, too much border (nav, footer, right hand side) html on pages with otherwise minor content, indexed ip address urls, additional url (i.e.: tracking parameters), .... and these are just on top of my head
Definitely should have added the slash vs no slash one though, will add later, feel free to add them yourself if you like ;)
Re: The Ultimate Guide To Duplicate Content
#6Other content that should be checked is the printable versions should not be indexed. Wordpress also generates some crawlable urls that cause the same content to be returned, e.g. archives, categories.
Presence of session ids in urls used to be a big problem, but search engines seems to have matured and a bit cleverer. However, if you aren't using a well known CMS, then it is better to make sure you link to a canonical version and don't put the one with session id in the search engine index.
Re: The Ultimate Guide To Duplicate Content
#7But a good start non-the-less for those who keep tripping up on this simple to fix issue.
Re: The Ultimate Guide To Duplicate Content
#8"www vs non-www" I've heard this before, but I have a hard time believing it. Google really can't figure out on its own that for some sites the www is optional?
I don't think they can, unless you submit both to Google Webmasters Tools and then select a default domain to show. Besides, why bother risking it when it's a 5 minute job?
Re: The Ultimate Guide To Duplicate Content
#9Earlier quoted context omitted.
I don't think they can, unless you submit both to Google Webmasters Tools and then select a default domain to show. Besides, why bother risking it when it's a 5 minute job?
Because the internet at large is not going to do that 5 minutes of work - just like with web standards, it's up to browsers to support the decade-of-html that was produced before anyone gave a crap about the w3c.
Think of Google as a retarded 30 year old user of site. If 67% of your users relied on the opinion of that single user, you'd damn sure do 5 mins worth of work to help him along.
And if you didn't it would be your own fault for failing to get his reference (read: ranking).
Re: The Ultimate Guide To Duplicate Content
#10Earlier quoted context omitted.
Because the internet at large is not going to do that 5 minutes of work - just like with web standards, it's up to browsers to support the decade-of-html that was produced before anyone gave a crap about the w3c.
Googlebot is a not a browser. It's a distinct "user" of your site with special needs. Think of Google as a retarded 30 year old user of site. If 67% of your users relied on the opinion of that single user, you'd damn sure do 5 mins worth of work to help him along. And if you didn't it would be your own fault for failing to get his reference (read: ranking).