Live data from Hacker News

Sci-Hub statistics and database

sci-hub.ru

91–100 of 150 posts

Re: Sci-Hub statistics and database

#91
post #89

The publishers now encode their papers with individual identifiers, that generationP calls UUIDs on all pdf's. That means they can trace it back to the institution(and perhaps the actual prof?) They have then sent nastygrams to threaten them with fees or loss of access to the institution - potentially serious punishment. What is needed is a way to run the papers through an OCR recognition program to create renewed te…

No “firm words” that I heard in my neck of the USA woods.

I was struck by very low number of German downloads. Did I miss something?

Re: Sci-Hub statistics and database

#93
post #70

Earlier quoted context omitted.

It didn't seem to actually work though. https://slate.com/technology/2022/01/ibm-watson-health-failu...

Did Watson fail because they where bad at their job or because the problem is much harder than people assumed?

Electronic health care records are not high quality data. They are qualitative, often discretized, and also distorted by fiscal shenanigans.

The best “EHR” data we have—quantitative and minimally biased—-are from large genetically diverse animal cohorts like the BXD mouse family.

Re: Sci-Hub statistics and database

#95
post #49

Earlier quoted context omitted.

What I read was that the Indian judicial system tends to be favorable to things like Sci Hub in its interpretation of copyright, and Sci Hub wanted to act in good faith with regard to that court, so as to have a fairly solid basis in international law for operating, should it rule in Sci Hub's favor. I might be off in this understanding, but that's what I understood.

Yeah, I have heard this reasoning, but it seems muddled. How is keeping the site online so old articles are available but no new articles are added acting in "good faith"? It's not like the old articles are any less copyrighted than the new articles, so this doesn't make sense to me. The court case has also been delayed for over a year now, so if it is delayed indefinitely, like it seems to be, then we will also not…

As I understand it, this was a sort of compromise the courts worked out in the interim. Seems likely that at some amount of delay, Sci-Hub would just break the injunction (I personally hope they don't, as the case seems to be wrapping up). I don't think any of this is ridiculous.

Relatedly, if any academics want to help a small bit to resolve this by signing the amicus brief (i.e. intervention application) we made for Sci-Hub, you can do so by contacting me through https://docs.google.com/forms/d/1_if6Lipu-YPBMLk6zYjBxDFRA_c... and I will connect you with the coordinating lawyer. You can read more about this at https://forum.effectivealtruism.org/posts/bEKwqNDGysnZRcmpw/...

Re: Sci-Hub statistics and database

#96

It's interesting how sci-hub's papers on medicine dwarf those in many other fields like comp-sci, math, and physics. I wonder if that reflects the number of papers in those fields, or if sci-hub just has a non-representative sample. If the latter, why?

I suspect that a lot of ill people, and people with loved ones who are ill, spend an enormous amount of time reading every paper they can find on the illness, or trying to narrow symptoms down to a particular illness. That's also a situation where you can go through 50 papers easily for every useful one (or even intelligible one, if you're a layman) you find, so a definite sci-hub situation unless you're independently wealthy. It's the requests that draw (or once drew, right now I guess) stuff into the database.

Re: Sci-Hub statistics and database

#97

It's interesting how sci-hub's papers on medicine dwarf those in many other fields like comp-sci, math, and physics. I wonder if that reflects the number of papers in those fields, or if sci-hub just has a non-representative sample. If the latter, why?

I don't even know the scale of medical research vs CS in terms of number of schools.

Re: Sci-Hub statistics and database

#98
post #89

The publishers now encode their papers with individual identifiers, that generationP calls UUIDs on all pdf's. That means they can trace it back to the institution(and perhaps the actual prof?) They have then sent nastygrams to threaten them with fees or loss of access to the institution - potentially serious punishment. What is needed is a way to run the papers through an OCR recognition program to create renewed te…

No “firm words” that I heard in my neck of the USA woods. I was struck by very low number of German downloads. Did I miss something?

No idea about Germany, but they may well be more law abiding?

Re: Sci-Hub statistics and database

#99
post #69

Earlier quoted context omitted.

Haven't got around to adding yet?

That is not how scihub used to function. Scihub used to have an engine, named Plato, which would fetch papers automatically if not already in their database. For the last year now, this essential service has not been operational. This is what the issue I am raising is about.

> this essential service

Give man a fish and he will praise you for a day. Give man a spinning and he will bitch at you because it is not a fish.

Re: Sci-Hub statistics and database

#100

Alexandra Elbakyan is a titan and a saint. I couldn't have been able to finish my research without access to papers my institution wasn't subscribed to.

Isn't it fantastic that we are alive and seeing the resurrection of the great library of Alexandria right before our eyes?

She has done more than any other organization or individual in the history of mankind when helping people in second and third world country pursue advance research since the advent of internet. Well she and the people who pirate and distribute MS Office. Faculties around the world recommend scihub as the main and only source of research and journals.

Post reply on HN