Live data from Hacker News

Anna's Archive: An Update from the Team

annas-archive.org

441–450 of 560 posts

Re: Anna's Archive: An Update from the Team

#442

Also, they provide a torrents list that anyone can seed and be part of the long-term preservation. https://annas-archive.org/torrents

Interesting to see that sci-hub is about 90TB and libgen-non-fiction is 77.5TB. To me, these are the two archives that really need protecting because this is the bulk of scientific knowledge - papers and textbooks. I keep about 16TB of personal storage space in a home server (spread over 4 spinning disks). The idea of expanding to ~200 TB however seems... intimidating. You're looking at ~qty 12 16TB disks (not counti…

It's 167.5, not ~200, and you can get disks much larger than 16 TB these days - a quick check shows 30 TB being sold in normal consumer stores although ~20 TB disks may still be more affordable per byte.

Re: Anna's Archive: An Update from the Team

#443

Earlier quoted context omitted.

Precisely. To be clear, I don't agree with a comment upthread saying the "shoutout" is what might potentially do harm to the IA in court. I think the actual act of having scraped all those books from the IA's lending system could potentially do harm to the IA in court. The publishers can now point to all the copies of the books in the wild that IA had in their lending system and argue that IA's system is not legally…

I believe this was already brought up in the court proceedings, and Brewster Kahle already addressed it in April 2024: «Trying to blow protections we have put on files, for instance, does not help us– and usually hurts». https://old.reddit.com/r/DataHoarder/comments/1bswhdj/commen...

IA lending books with "weak" DRM also hurts efforts in reducing DRM and reforming copyright though and that is much more important in the long term. It was always a deal with the devil that IA should have never made and them now being at odds with others that preserve those books and actually make them available only makes that more clear.

It's like a food kitchen under a tyrannical regime complaining that people passing their food to rebels might get them shut down.

Re: Anna's Archive: An Update from the Team

#444

Earlier quoted context omitted.

So what about the authors and creators of the works? They did it for free?

Information and well-crafted sentences are available on the Language Tree, easily plucked by anyone at zero cost. It's greedy for those so-called novelists and subject matter experts to expect a living wage. "Information wants to be free," which means that any cost of producing that information can be abstracted away due to ideological inconvenience.

Then show me the easily available "information on the langauge tree" to solve the unsolved problems in science. Btw. books are not mere information, they are also products of effort and sacrifice and intentions. They are also embedded in an economic system of paper, books, ink, transport and what not producers.

So you are either poor or too lazy to buy a book from the store. But this doesn't justify mind theft or it's distribution.

Re: Anna's Archive: An Update from the Team

#445

Earlier quoted context omitted.

So what about the authors and creators of the works? They did it for free?

they already work almost for free, since all the money goes to the publisher and retailer. out of $20 book, the authors earn about $1 - $1.5, for e-books its about $1.7 - $2 The value from book sales goes to retailer and publisher: two large corporations, and in case of amazon - a single big corporation so please cry me a river about amazon's lost profits earned at the back of the book authors

This is a problem of publishers and retailers, and not a justification for distribution of mind theft.

Re: Anna's Archive: An Update from the Team

#447

Earlier quoted context omitted.

If I go to the public library, check out a movie on disc, back it up, and share the back up file online, is my public library legally liable

Depends a bit probably if your local library has major lawsuits for operating in a very sketchy side of the legal gray area

Maybe your library shouldn't have made choices that put it at odds with the data preservation community then.

Re: Anna's Archive: An Update from the Team

#448
post #90

Earlier quoted context omitted.

Except that that’s CloudFlare, which is also blocking Anna’s Archive.

Luckily it isn't the only public DNS. 8.8.8.8, 9.9.9.9, and many others exist.

Here it's Cloudflare the CDN, sitting in front of Anna's Archive, that's doing the blocking. The DNS resolver used doesn't come into play.

(Case in point, I am using Google's DNS, yet still encounter the block when accessing from a Belgian IP.)

Re: Anna's Archive: An Update from the Team

#449
Meta illegally scraped 80TB of data from Anna's archive, Libgen, Zlib etc. I'm sure other tech giants did too. Without paying them a cent, costing these projects $$$ in bandwidth/hosting etc.

when I hear people complain about these projects it just sounds like hypocrisy.

Post reply on HN