Viewing profile — dataslap
dataslap
HN member- Joined
- Tue, Nov 14, 2017, 9:27 PM UTC
- HN karma
- 21
- Public activity
- 12 items
- HN profile
- View on Hacker News ↗
About dataslap
No profile information was provided.
Recent public activity
-
comment
Comment #16314910
21 Gigs in total for everything, ignoring errors (files too big / broken links)
-
comment
Comment #16314135
I'm really happy their share it though. The collection is just amazing - it's a good idea for "what books to buy for a personal library when you start making decent money". Definit…
-
comment
Comment #16314122
I'm at 300 right now, 9.5 gigs. I think some are skipped cause the download times out (can redownload manually I suppose)
-
comment
Comment #16311324
P.S. it will take hours or days =))) <3 <3 <3
-
comment
Comment #16311320
Please use responsibly. Recommended to delay by a few days - it seems like there is an initial surge going on right now. Increase download delay to 20 sec - 1 minute. Be respectful…
-
comment
Comment #16311310
I think someone else is messing with it - got slower for me too.
-
comment
Comment #16311180
It's slow cause things are quite large there (just saw a 2 gig book). I'm using delays, don't worry - hence "in a few hours" <3 <3 <3
- comment
-
comment
Comment #16310785
Uploading a scrapy crawler that downloads PDF books to github in a few hours, gonna post the link here
-
comment
Comment #15699368
some "ethical" measures may do the trick to. scrapy has a setting to integrate delays + you can use fake headers. Some sites are pretty persistent with their cookies (include cooki…
-
comment
Comment #15699331
depends on the task. For example they have a decent file/image downloading middleware.
-
comment
Comment #15699299
scrapy has a pretty decent parser too