I don't care. In fact, I prefer it this way. Death in species is an evolutionary advantage. So too with culture. We mustn't let it ossify. For the whitewashing thing, it will happen anyway. Only vigilance can protect against rewriting. Websites can be altered. There is no provenance. I'm not convinced infinite recall is useful.
I don't think it's for us to decide. We've been mourning the loss of the Library of Alexandria for circa 1,500 years and I can't notice that we're getting radically wiser during the last 25 ones. Letting all go is definitely the cheaper and convenient attitude... for us. But we might be leaving nothing to build upon for the future generation. They should have the same opportunity to ignore what they want that we have…
Internet Data Is Rotting
51–60 of 116 posts
Re: Internet Data Is Rotting
#52> Then there is also a problem of software preservation: How can people today or in the future interpret those WordPerfect or WordStar files from the 1980s, when the original software companies have stopped supporting them or gone out of business? This issue in particular we have great solutions for (open formats / text), but they are of course less profitable than only-my-app-can-read-this formats.
FWIW those particular formats are widely understood even if they are proprietary (well, at least in WordStar's case). And as long as the software runs (be it natively or via an emulator or VM), you can always open and convert/print the files (e.g. you could use vDOS to run WordStar or whatever and use its printer emulator functionality with Windows' PDF printer to create a PDF from the WordStar files).
Re: Internet Data Is Rotting
#53I read somewhere that the lifespan of the average hyperlink is only about two years. I count myself lucky I was introduced to the HTTRACK archiver program many years ago and thus have complete offline copies of many of my favorite websites of the early 00's.
Re: Internet Data Is Rotting
#54Earlier quoted context omitted.
Can you give some examples of these 'favorite websites'? I'm interested in knowing what kind of website would be so interesting that I would want an entire offline copy of it. (Besides maybe Wikipedia)
Mostly defunct webcomics but also some of the small personal sites that documented and collected resources for particular events or strange people. A lot of those arose out of the SomethingAwful forums. For example, there was a guy named Brian who wrote batshit insane fanfiction about himself. One of these sites archived the fiction, interviews with Brian, videos, recordings of collaborative reading skype parties, et…
Re: Internet Data Is Rotting
#55I’m okay with internet rot and you should be too. I’m not sure where we got the idea that “our data must be preserved forever”. This can be especially harmful for teens and young adults whose indiscretions now follow them forever. Think of the privilege you had when you were younger. You could do something stupid and nobody could whip out a high def camera to record it and make it part of your history forever. Let it…
I think that ship has sailed. All our most personal data is being archived forever competently by multiple parties.
So let's think about what we can do for the generation that is about to be born.
Re: Internet Data Is Rotting
#56This needs to be solved on the protocol level. Of course, the players who have control over our protocols are exactly the people who don't want this to be solved at all. The next best thing would be to redefine what "bookmarking" is. When I bookmark a page, I want it to be permanently stored on my local machine and full-text indexed. In fact, it's rather ridiculous that after 25 years browsers don't have anything of…
Re: Internet Data Is Rotting
#57Earlier quoted context omitted.
Mostly defunct webcomics but also some of the small personal sites that documented and collected resources for particular events or strange people. A lot of those arose out of the SomethingAwful forums. For example, there was a guy named Brian who wrote batshit insane fanfiction about himself. One of these sites archived the fiction, interviews with Brian, videos, recordings of collaborative reading skype parties, et…
Are you able to navigate through the sites using the original links? I notice that on the Wayback Machine, internal site links only work if that particular page was also archived.
Re: Internet Data Is Rotting
#58Earlier quoted context omitted.
I'm OK with it because otherwise you are whitewashing history. For example I have recordings of the Colbert report going back to ~2005. Some of his skits released during that time would be classified as "hate speech" in 2019. Of course he, and mainstream broadcasting companies would love it if you didn't think about that. There are plenty of news clips and interviews where mainstream politicians (on Left AND Right) c…
> Some of his skits released during that time would be classified as "hate speech" in 2019. Yeah, I'm going to need some proof for that. Eddie Murphy's '80s HBO standup routines are available in full on YouTube, and some of that material is absurdly homophobic. He isn't a pariah by any means. People do see nuance and understand that values and mores change over time, and what might have been socially acceptable at on…
> GLAAD: Kevin Hart ‘Shouldn’t Have Stepped Down’ as Oscars Host, but Used Gig to Bring Unity
https://www.indiewire.com/2018/12/glaad-kevin-hart-oscars-ho...
Re: Internet Data Is Rotting
#59Earlier quoted context omitted.
wget -r ?
No. There are some cases where it is useful to download many pages in a batch, but what I am talking about is, effectively, partial local replication. Bookmarking I describe should create a tiny (but useful) version of the web on your computer. It needs to be seamless. It needs to be searchable. It would be incredibly useful if it would capture relationships between pages (links) in addition to pages themselves to na…
You can even take it a step further and create a viewable timeline of all the pages you've visited in case you have something at the tip of your tongue, but need to retrace your steps and logic to get there. Browsing history is kinda lackluster for this. My last ten entries on firefox are three hackernews articles with "Add Comment" dispersed half a dozen times in that list. If I had something along the lines of zoteros timeline for papers I could probably find stuff way easier.
Re: Internet Data Is Rotting
#60I’m okay with internet rot and you should be too. I’m not sure where we got the idea that “our data must be preserved forever”. This can be especially harmful for teens and young adults whose indiscretions now follow them forever. Think of the privilege you had when you were younger. You could do something stupid and nobody could whip out a high def camera to record it and make it part of your history forever. Let it…
Some reddit users are egregious about it, installing scripts that overwrite their comments after x amount of time, seemingly oblivious to the fact that every edit on reddit can be found through archival tools. The solution is to take better care to not conflate your anonymous online persona to your real life persona—just don't post identifying information publicly online and you will be head and shoulders above many users on the internet in terms of privacy. There's no need to purge the internet of it's collective knowledge and history.
We are very lucky to have the wayback machine preserving this stuff from dissapearing into the void, but it doesn't cache everything on the internet, especially if that forum I visited had shut down and became impossible to find in a search result.
Side note: Is there an extension or bookmarklet available to automatically pull a web archive?