Live data from Hacker News

Internet Archive Scholar

scholar.archive.org

21–30 of 59 posts

Re: Internet Archive Scholar

#21
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

Some people use this tool (of mine) for saving web content from either bookmarks or just everything you browse: https://github.com/crisdosyago/Diskernet There's also plent¥ of other similar tools: - https://github.com/ArchiveBox/ArchiveBox - https://github.com/gildas-lormeau/SingleFile

I'd never heard of SingleFile before but it looks excellent. It would be great if Firefox could incorporate it into its save function too. Firefox save page works but as shown on the SingleFile demo video, it's not really what a user would expect, it's often not complete and splitting it across multiple files/directories isn't ideal either.

Re: Internet Archive Scholar

#23
post #3

I've already made my donation to IA this year but I might need to make another. Somehow it's the IA's job to fix problems that we all know are problems, sadly.

If you're able/comfortable, please consider setting up a recurring donation. For long-term planning reasons, it's helpful for organizations to have a consistent recurring revenue stream that they can use to project assets further into the future. One-off donations are good, too! But if you're going to consistently send them money anyway, you may as well do it in a predictable manner to help their accounting.

Re: Internet Archive Scholar

#24
post #4
post #3

I've already made my donation to IA this year but I might need to make another. Somehow it's the IA's job to fix problems that we all know are problems, sadly.

Wikipedia reminded me multiple times to donate to the Internet Archive this year.

Another chance to upvote the donation link to top thread on an IA story since a direct submission got swallowed by the dupe detector! They are doing such much amazing things.

https://archive.org/donate

Your Donation Will Be Matched 2-to-1! [...] Right now, we have a 2-to-1 Matching Gift Campaign, tripling the impact of every donation. (from the home page)

Re: Internet Archive Scholar

#25
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

I've been wanting to run my own search engine sorta thingy that indexes websites I feed it. I sometimes find little nooks of the net that post resources I may need in the future. Like my own mini-Google that indexes a list of sites.

How can I go about creating this? Are their off-the-shelf solutions, will I need to say combine scrapy with elastic search? The links in this thread look promising.

Re: Internet Archive Scholar

#26
This seems like the type of thing that will become the search engine of first resort in the future as AI-generated propaganda and nonsense pollutes the spectrum of websites and search results.

Re: Internet Archive Scholar

#27
post #5

Some days, nearly half the links I click are dead so I've found myself relying on the waybackmachine more and more over the past few months. It's really shocking just how fast digital obsolescence reared its ugly head. Of course angelcites etc. were a clear early blow, but nowadays... I've started saving the html (including the css seems like too much overhead, and often it's incomplete or relies on downloads still -…

I just save the pdf of a site that's really important to me.

When I did a major college project in 2003, I made sure to make pdfs of any academic article that I referenced. It actually saved me, because some articles disappeared by the time I went to revise my references.

Re: Internet Archive Scholar

#28

Earlier quoted context omitted.

> I've started saving I do similar. I've had https://github.com/ArchiveBox/ArchiveBox bookmarked for a while as something to try better organise all that, but like a great many things I haven't go around to it yet.

I use Raindrop for this. It’s a pretty great bookmark manager made by an indie dev, but it also can create archives of bookmarked pages. https://help.raindrop.io/backups#permanent-library

I use Raindrop but didn't know abou that feature. Thanks.

Re: Internet Archive Scholar

#30

Earlier quoted context omitted.

> I've started saving I do similar. I've had https://github.com/ArchiveBox/ArchiveBox bookmarked for a while as something to try better organise all that, but like a great many things I haven't go around to it yet.

I use Raindrop for this. It’s a pretty great bookmark manager made by an indie dev, but it also can create archives of bookmarked pages. https://help.raindrop.io/backups#permanent-library

Thanks, I'll add that to the list of things to try out.
Post reply on HN