Live data from Hacker News

Sci-Hub statistics and database

sci-hub.ru

81–90 of 150 posts

Re: Sci-Hub statistics and database

#82

Alexandra Elbakyan is a titan and a saint. I couldn't have been able to finish my research without access to papers my institution wasn't subscribed to.

I snuck her into my own dissertation acknowledgements: https://imgur.com/bDgtBAE

Now i feel foolish for not acknowledging her, especially in elsevier papers.

Re: Sci-Hub statistics and database

#83

Earlier quoted context omitted.

> published on arxiv.org and open access Don't use the term "open access" like this. A paper published on arXiv is free to read, and was freely published. "Open access" is a scam by the big publishers, where they don't take money from the readers , but make the authors pay. Or, putting it another way, anyone can pay their way in those journals and publish (sometimes sub par) papers.

No, "open access" means that the paper is available to readers for free. Making the authors pay is typically termed "gold open access".

I wasn't aware that there were different distinct forms of "open access", so I had to read it up on Wikipedia. From what I understand, publications on arxiv are either gratis or libre open access.

Either way, we don't pay anyone any fee to publish on arxiv.

Re: Sci-Hub statistics and database

#84

Earlier quoted context omitted.

It is note worthy that most of physics (at least high energy physics) ist published on arxiv.org and open access. I don't know if sci hub bothers with publications that are available freely from an official source.

> published on arxiv.org and open access Don't use the term "open access" like this. A paper published on arXiv is free to read, and was freely published. "Open access" is a scam by the big publishers, where they don't take money from the readers , but make the authors pay. Or, putting it another way, anyone can pay their way in those journals and publish (sometimes sub par) papers.

As I wrote on another comment, I wasn't aware that there are multiple forms of open access. Since it appears that arxiv (again, at least high energy physics) employs mostly either gratis or libre open access, and since the Wikipedia article explicitly calls it an open access archive, I see no harm in calling it that either.

"arXiv (pronounced "archive"—the X represents the Greek letter chi [χ])[1] is an open-access repository of electronic preprints and postprints[...] "

Re: Sci-Hub statistics and database

#88

Is there any word about when sci hub is going to start adding new articles again? It's currently only useful as an archive of old research articles. New papers from the last year are not available. I never understood the rationale for stopping new content, though I believe it had some relation to some court case in India...but I don't understand why that was a reason to stop adding articles, and why it hasn't been re…

They had resumed adding articles after receiving legal advice that the Delhi High Court injunction only applied for a few months. https://mobile.twitter.com/ringo_ring/status/143435621720862...

[deleted]

Re: Sci-Hub statistics and database

#89
The publishers now encode their papers with individual identifiers, that generationP calls UUIDs on all pdf's. That means they can trace it back to the institution(and perhaps the actual prof?) They have then sent nastygrams to threaten them with fees or loss of access to the institution - potentially serious punishment. What is needed is a way to run the papers through an OCR recognition program to create renewed text and to further process to randomly vary a large number of adjacent letter spacings. (this is called 'micro kerning' and it allowed a second order of unique document recognition where the document is scanned by the journal to look for these fractional kerning gaps in the letter spacing which can lead back to the institution). I suppose a program flow could be made to OCR + random micro kerning changes - it would take time, but once set up it would be a rapid computer based document flow process. Photos/charts could be sanitised as well, but legends etc on them would also need OCR and micro kerning adjustments. With 25 words of text, each with 6 letters, micro kerning can easily create 10,000 unique ID's, easily enough to cope with all subscribed institutions - and that is on one chart/photo. I suspect the subscribing institutions have acted with firm words to their profs and grad students to block this. This can easily be told to us if a few people in assorted institutions let us know if they have been read firm words about this?

I am not sure how Sci-hub can get past this, unless they get a good Indian court ruling and can use Indian friends to scan printed copies of journals - if they exist in this online age?

Post reply on HN