Interesting comparison from 1998: http://slashdot.org/poll/421/My-Age-is In 1998, 22% of respondents were ages 16-20, 53% were in their 20s, 15% in their 30s and 5% over 40. As of this writing (and assuming that I did my arithmetic correctly), ~7% of respondents on this poll are ages 16-20, ~60% are in their 20s, ~25% in their 30s, and ~7% are over 40. Obviously there are a lot of caveats here, but the slight increas…
As I commented previously when we had a poll on the ages of HNers, the data can't be relied on to make such an inference. That's because the date are not from a random sample of the relevant population. One professor of statistics, who is a co-author of a highly regarded AP statistics textbook, has tried to popularize the phrase that "voluntary response data are worthless" to go along with the phrase "correlation doe…
In this case, there's a strong argument to be made that the samples could skew young (if for no other reason than because younger people frequent sites like HN more than older people). It's also quite possible that the demographics for HN and for slashdot are different in a way that would make the HN audience skew younger. But the voluntary response process itself? I think it's a stretch to claim that it'll cause a measurable bias in age distributions, across two different surveys, spaced a decade apart. You'd have to argue that younger people in 2011 are more likely to answer a poll than younger people in 1998 were to answer a poll. Possible? Sure. The most likely source of error here? Nah.
In other words, if you're going to pick (and you should!), pick on the BIG, OBVIOUS differences between the data sets...not the little ones.