Earlier quoted context omitted.
LigGen / SciHub are model citizens in this regard. Use the best tech for the job . Torrents + IPFS + simple http mirrors all over . While their data is big its not as big as Archive.org https://libgen.is/repository_torrent/ so I guess one still needs the funding to go to these decentralized nodes and they are there mostly for resiliancy
If so, why do they keep jumping from one dubious domain and TLD from another?
Internet Archive breached again through stolen access tokens
351–360 of 376 posts
Re: Internet Archive breached again through stolen access tokens
#352Earlier quoted context omitted.
Internet archive allows full text search of books, newspapers, etc.. Or anyway it did, before being breached.
It does transcribe books (through imperfect OCR) so I guess that's possible. Never relied on it as I search by title and author. But anyways not the case for the wayback product which is the unique core to IA.
Well, OK, maybe other webpage archives don't work as well, I haven't tried them, but there are others. And they're newer, so don't have such extensive historical pages.
Large numbers of Wikipedia references (which relied on IA to prevent link rot) must be completely broken now.
Re: Internet Archive breached again through stolen access tokens
#353Earlier quoted context omitted.
A million is out of the parameters of the case. Realistically you won't get enough volunteer-storage to cover one IA. And even if you did, it wouldn't satisfy the mission requirements, which is to store reliably for decades all of the data.
This isn't meant to be storage for IA, it's meant to be a distributed backup.
Re: Internet Archive breached again through stolen access tokens
#354Earlier quoted context omitted.
Ooo, excellent. Yes, hiding items is imperfect, but I understood that it was legally required or something. (IANAL and IDFK, TBH) I wonder how perma.cc gets around that.
I'm afraid that it just hasn't been tested in court yet. I haven't read this paper yet, but... https://www.tesble.com/10.1080/0270319x.2021.1886785 from the abstract: > The article concludes that Perma.cc's archival use is neither firmly grounded in existing fair use nor library exemptions; that Perma.cc, its "registrar" library, institutional affiliates, and its contributors have some (at least theoretical) exposure…
I'll hold my breath.
Re: Internet Archive breached again through stolen access tokens
#355Earlier quoted context omitted.
This seems to get brought at least once in the comments for every one of these articles that pops up. The IA has tried distributing their stores, but nowhere near enough people actually put their storage where their mouths are.
Perhaps a naïve question, but hasn't this problem been solved by the FreeNet Project (now HyphaNet) [0]? (the re-write — current FreeNet — was previously called Locutus, IIRC [1]). Side note: As an outsider, and someone who hasn't tried either version of FreeNet in more than almost 2 decades, was this kind of a schism like the Python 2 vs. Python 3 kerfuffle? Is there more to it? [0]: https://www.hyphanet.org/ [1]: h…
Neither version of Freenet is designed for long-term archiving of large amounts of data so it probably isn't ideally suited to replacing archive.org, but we are planning to build decentralized alternatives to services like wikipedia on top of Freenet.
[1] https://freenet.org/faq/#why-was-freenet-rearchitected-and-r...
Re: Internet Archive breached again through stolen access tokens
#356We need archives built on decentralized storage. Don't get me wrong, I really like and support the work Internet Archive is doing, but preserving history is too important to entrust it solely to singular entities, which means singular points of failure.
Re: Internet Archive breached again through stolen access tokens
#357Nobody has ever stopped a competitive alternative from existing. Feel free to give it a shot. You have a head start with all the work that they've done and shared.
Re: Internet Archive breached again through stolen access tokens
#358Earlier quoted context omitted.
This isn't meant to be storage for IA, it's meant to be a distributed backup.
Ah my bad, so it's not a replacement of IA. In that case it makes sense
Re: Internet Archive breached again through stolen access tokens
#359Earlier quoted context omitted.
I'm not going to look up legal precedent, hire a lawyer if you want that. You are wrong, copyright specifically prohibits copying, not distribution. They can get a cease and desist that requests you destroy property and they ca get a court order backing that which will put you into contempt of court if you fail to do so. Proving damages is easier with distribution, but that is a civil matter not a criminal matter.
They’re not making copies either.
You do realise the "downloading" is implicitly a copy.
If you want to actually have a civil discussion then you need to make some reasonable argument than "They're not making copies either."
Sounds like whatever role you played at IA when you were there didn't give you any actual insight into what happens in operation and you simply tried to prove your point with an appeal to authority instead of backing it with facts and reason.
Re: Internet Archive breached again through stolen access tokens
#360Earlier quoted context omitted.
Ah my bad, so it's not a replacement of IA. In that case it makes sense
Yes, the idea is that this is a replacement for the torrents they make public. In case the IA goes away, we'll have this distributed dataset to fall back on.