Live data from Hacker News

Internet Archive Scholar

scholar.archive.org

11–20 of 59 posts

Re: Internet Archive Scholar

#11
archive.org is an alternative good internet as a giant library, as dreamed in early 90's: web archive, film archive, software archive, media archive... and now research papers archive.

Re: Internet Archive Scholar

#12
post #4
post #3

I've already made my donation to IA this year but I might need to make another. Somehow it's the IA's job to fix problems that we all know are problems, sadly.

Wikipedia reminded me multiple times to donate to the Internet Archive this year.

Is this so Wikipedia can be archived by the IA?

Re: Internet Archive Scholar

#13
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

I use the markdownload extension[1] on firefox and move the .md file into my notes folder (notable[2]). Works very well.

1. https://addons.mozilla.org/en-US/firefox/addon/markdownload/

2. https://notable.app/

Re: Internet Archive Scholar

#14
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

I find the SingleFile extension superb for this.

Re: Internet Archive Scholar

#15
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

While the author(s) are still alive, they are often a productive contact. (In one case I was able to give back: bundling up the several scans an author had of a half-century old paper from their student days into a single, hopefully cromulent, PDF) Edit: recall also that accepting that links are one-way and might be dead was the key simplification that allowed HTTP to take off after prior attempts at hypermedia had f…

Good (Edit) point! It's good that the web accepts dead links by design, we can't expect perfection from our distributed information, but it seems that the bitrot of information is too high compared to the information storage technologies available.

2 spinning rust drive can store the library of congress. ~ 2,000 drives would store the web (1). How many millions of these drives get manufactured per year? Our technology systems are failing us - all those words are being lost, like tears in rain.

(1)Back of the envelope estimation: https://www.worldwidewebsize.com/ ~ estimates 50 billion websites, with some estimates ~ 6 pages of information per website. Let's say 1mb per page on average. So ~2,000 drives would store the entire web.

Re: Internet Archive Scholar

#16
post #12
post #4

Earlier quoted context omitted.

Wikipedia reminded me multiple times to donate to the Internet Archive this year.

Is this so Wikipedia can be archived by the IA?

I always take the wikipedia donation drive as a reminder to donate to archive.org instead.

Re: Internet Archive Scholar

#17
post #4
post #3

I've already made my donation to IA this year but I might need to make another. Somehow it's the IA's job to fix problems that we all know are problems, sadly.

Wikipedia reminded me multiple times to donate to the Internet Archive this year.

Same. Oh hey these scammers are asking for money again? Wait, I haven’t given to IA in a while.

Re: Internet Archive Scholar

#18
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

Some people use this tool (of mine) for saving web content from either bookmarks or just everything you browse: https://github.com/crisdosyago/Diskernet

There's also plent¥ of other similar tools:

- https://github.com/ArchiveBox/ArchiveBox

- https://github.com/gildas-lormeau/SingleFile

Re: Internet Archive Scholar

#19
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

Some people use this tool (of mine) for saving web content from either bookmarks or just everything you browse: https://github.com/crisdosyago/Diskernet There's also plent¥ of other similar tools: - https://github.com/ArchiveBox/ArchiveBox - https://github.com/gildas-lormeau/SingleFile

Zotero also saves snapshots of pages if you already cite academic pages

Re: Internet Archive Scholar

#20
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

Some people use this tool (of mine) for saving web content from either bookmarks or just everything you browse: https://github.com/crisdosyago/Diskernet There's also plent¥ of other similar tools: - https://github.com/ArchiveBox/ArchiveBox - https://github.com/gildas-lormeau/SingleFile

How have I never seen your tool before.

>22120 archives content exactly as it is received and presented by a browser, and it also replays that content exactly as if the resource were being taken from online.

I've been looking for this for a really long time.

Post reply on HN