Earlier quoted context omitted.
You seem to be assuming 1. I'm scraping the resized galleries. 2. I don't have the Hath perk that makes the galleries full sized. 3. I don't have a phash-based fuzzy image deduplication system on top of all this (see https://github.com/fake-name/IntraArchiveDeduplicator ). It's main purpose is to deduplicate manga ( https://github.com/fake-name/MangaCMS ).
Jesus, your projects are massive. Does your job involve working on these or are these just side things?
Re: Web Scraping in 2016
#401It's all entirely hobby things.