This is why I’ve gotten into the habit of maintaining my own WWW archive of sites I find interesting. Probably have around 1 TiB now, and One Of These Days I’d like to set my network up so it can serve arbitrary sites directly from local archive to revive any site I want. I have a `wget-mirror` shell function invoking wget with all the trimmings that takes care of 99% of sites. I’ll edit the full command into this co…
Assume everyone is familiar with this project, dating back to 1996: https://en.wikipedia.org/wiki/WWWOFFLE https://ftp.netbsd.org/pub/pkgsrc/distfiles/wwwoffle-2.9j.tg... The way the www is going, it seems like downloading a copy of libgen, i.e., nonfiction books, and scimag, i.e., academic journals, via torrent, would be more valuable than archiving websites, in general. These primary sources are part of the materia…
https://news.ycombinator.com/item?id=40258584
https://arxiv.org/pdf/2005.14165.pdf
https://www.wired.com/story/battle-over-books3
https://www.washingtonpost.com/technology/interactive/2023/a...
https://www.theguardian.com/technology/2023/apr/20/fresh-con...
https://storage.courtlistener.com/recap/gov.uscourts.cand.41...
See 40-45.
https://storage.courtlistener.com/recap/gov.uscourts.nysd.60...
See 87-116.