Earlier quoted context omitted.
It's the great irony of digital media. Copying data is accurate to the bit and is preserved "as-is", but in practice, it requires someone to maintain servers, to care about it. To separate out what is worth preserving. We hot-linked to all those image hosts because we couldn't imagine them disappearing. Archive.org had incredible foresight and if it didn't already exist, I'd call such a project a pipe dream.
Even Archive.org is rather limited in what it keeps. I know of a very large site that recently disappeared. archive only has part of the web html part of the site. Everything else is either gone or non accessible.
Where did the old web go? We followed 657,607 links to find out
31–40 of 231 posts
Re: Where did the old web go? We followed 657,607 links to find out
#32I found an old database backup of 0.mk on a disk I had kept. 0.mk started in 2009 as a passion project built by three of us. We worked on it for a few hours each week around our regular jobs. We eventually closed it in 2014 because the revenue (hint: no revenue) could not cover hosting, development, and the constant work of fighting spam and reviewing abuse. The recovered historical corpus contains 657,607 links. For…
How did you managed to obtain that domain? Usually single digit or letter domains are “reserved”.
Re: Where did the old web go? We followed 657,607 links to find out
#33Am I getting old? 09-14 is not even close to the old web for me. The old web, to me, was back when people still published physical 'phone' books for websites.
Re: Where did the old web go? We followed 657,607 links to find out
#34Re: Where did the old web go? We followed 657,607 links to find out
#35Re: Where did the old web go? We followed 657,607 links to find out
#36Re: Where did the old web go? We followed 657,607 links to find out
#372009-2014? That's not the old web. Or I'm old. Take your pick.
Re: Where did the old web go? We followed 657,607 links to find out
#38Re: Where did the old web go? We followed 657,607 links to find out
#39Webpages dying is probably one of the biggest design flaws of the original web. I am not saying old content needs to be preserved forever, but so much content has factually been lost over time. Old logs from text-based MUDs for instance, even for MUDs that still exist today.
It's the great irony of digital media. Copying data is accurate to the bit and is preserved "as-is", but in practice, it requires someone to maintain servers, to care about it. To separate out what is worth preserving. We hot-linked to all those image hosts because we couldn't imagine them disappearing. Archive.org had incredible foresight and if it didn't already exist, I'd call such a project a pipe dream.
This is a bit idealized. In practice copying data is not quite accurate (especially in bulk) and bit-rot is a very real phenomenon, both in flight and in storage.
You sometimes encounter it when dealing with files from the early '00s, it's very common to discover a few of them are corrupt, even if they've only ever been copied between harddrives.