Live data from Hacker News

Sci-Hub statistics and database

sci-hub.ru

131–140 of 150 posts

Re: Sci-Hub statistics and database

#131

Sci-hub is sometimes the last resort to obtain a resource that is otherwise unobtainable. But what has become of the old way of obtaining unaccessible papers: Ask the authors for a copy? Sites like ResearchGate make this very easy. And often a simple email does the job, too. Advantages: * It is legal * The author gets feedback that someone out there reads their research * Making direct contact to your peers is a good…

> It is legal

Are you sure that the usual suspects don't make authors assign copyright or at least distribution rights? I wouldn't put it past them...

Re: Sci-Hub statistics and database

#132

Sci-hub is sometimes the last resort to obtain a resource that is otherwise unobtainable. But what has become of the old way of obtaining unaccessible papers: Ask the authors for a copy? Sites like ResearchGate make this very easy. And often a simple email does the job, too. Advantages: * It is legal * The author gets feedback that someone out there reads their research * Making direct contact to your peers is a good…

Except a researcher could easily need to skim 30+ papers in a day, this is not a solution.

Re: Sci-Hub statistics and database

#133

Sci-hub is sometimes the last resort to obtain a resource that is otherwise unobtainable. But what has become of the old way of obtaining unaccessible papers: Ask the authors for a copy? Sites like ResearchGate make this very easy. And often a simple email does the job, too. Advantages: * It is legal * The author gets feedback that someone out there reads their research * Making direct contact to your peers is a good…

> It is legal Are you sure that the usual suspects don't make authors assign copyright or at least distribution rights? I wouldn't put it past them...

I have signed quite a few of those copyright transfer forms, and there are always a clause that would allow sharing on a personal basis.

Re: Sci-Hub statistics and database

#134

Sci-hub is sometimes the last resort to obtain a resource that is otherwise unobtainable. But what has become of the old way of obtaining unaccessible papers: Ask the authors for a copy? Sites like ResearchGate make this very easy. And often a simple email does the job, too. Advantages: * It is legal * The author gets feedback that someone out there reads their research * Making direct contact to your peers is a good…

Except a researcher could easily need to skim 30+ papers in a day, this is not a solution.

Agreed, in that case that wouldn't work.

But honestly, how often does one have to skim that many papers in a day, to a level where the freely available abstract is not sufficient?

Perhaps every once in a while when one compiles a survey of a new field they enter. Once the project is set on the rails, one rarely has to read that much.

Re: Sci-Hub statistics and database

#135
post #129

Sci-hub is sometimes the last resort to obtain a resource that is otherwise unobtainable. But what has become of the old way of obtaining unaccessible papers: Ask the authors for a copy? Sites like ResearchGate make this very easy. And often a simple email does the job, too. Advantages: * It is legal * The author gets feedback that someone out there reads their research * Making direct contact to your peers is a good…

It's too time consuming and has an undefined likelihood of success. People will naturally flock to alternative methods, such as sci-hub, that are faster and until recently were near guaranteed to have the desired content.

Agreed, sci-hub is so much more convenient. But when the publishers finally shut it down for good, we'll have to find another solution.

A community of scientists sharing their papers would be a good thing already now.

I personally know active scientists who don't even try anymore to look up the paper, but rather go directly to sci-hub for any doi they need. I can understand why, but I also think that this doesn't lead to a sustainable publishing culture.

Re: Sci-Hub statistics and database

#136

Earlier quoted context omitted.

Except a researcher could easily need to skim 30+ papers in a day, this is not a solution.

Agreed, in that case that wouldn't work. But honestly, how often does one have to skim that many papers in a day, to a level where the freely available abstract is not sufficient? Perhaps every once in a while when one compiles a survey of a new field they enter. Once the project is set on the rails, one rarely has to read that much.

> how often does one have to skim that many papers in a day, to a level where the freely available abstract is not sufficient?

More often than you might think.

To take an example from my own work, I was doing assay design a while back, and needed to collect all existing primer sets in the literature. I probably went through a hundred papers over a several day period.

Re: Sci-Hub statistics and database

#137
post #105

Earlier quoted context omitted.

Makes me wonder how much they'd have to offer me to accept the task of implementing such tracking algorithms to modify other people's scientific papers for the purpose. Certainly I won't be the cheapest, but still. Either some developer is very vested in the idea of keeping science a secret or someone got a very nice bonus. Edit: also the sysadmin that keeps this database safe without 'accidental' data loss on UUID t…

It seems to me that this type of tracking, under the guise of guarding rights is against the GDPR = unlawful tracking that they are harassing Google and FB with? Is not justice's sword 2 edged?

(Note that this is all written from a European perspective, applies European Economic Area law and human rights as defined by the Council of Europe (which includes Russia, to everyone's surprise). I know that in the USA everyone is much more pro frontier justice, for example when it comes to pervasive and continual monitoring of employees while at work.)

I think it's very arguable that they have a legitimate interest here. Privacy has always been a weighing of interests, at least that's how I've always heard it explained by the Netherlands' face of digital law (Arnoud Engelfriet) also back in the days of WBP (the law from ~1995 that is 97% the same thing as GDPR), also in light of the European Convention of Human Rights (article 8 is a right to privacy).

A common example is filming the road: illegal, but if you park your car in front of your house and there have been car fires in your neighborhood lately, then it can be justified.

Filming employees inside a warehouse: invasion of privacy (illegal) but if there have recently been thefts from a certain part of the building then it's justified to hang up a camera there, introduce a lock that registers who went there at what time, or some such. (With adequate security measures so only authorized people can use it for the intended purpose.)

Personal example: monitoring everything I do on the company network is illegal, but because I work in a business where secrecy is important (security consultancy) it was considered justified to do spot checks, tell every employee upon entering into the employment contract that spot checks are a thing, and inform the subjects of spot checks after they were part of one. Transparent but still effective.

The two things to consider (iirc) are:

- Do my rights weigh heavier than the other party's right to privacy? (e.g. car fire is a fairly big impact on your right to the peaceful enjoyment of his possessions)

- Is there any other way in which I could achieve this goal with a lesser impact on the right to privacy?

In the case of Elsevier, from what I heard this whole scheme is a big mafia-like practice (wouldn't want to be published in a niche corner nobody reads now would you?) and so in my opinion it's entirely unethical to support (work for) them in the first place, at least in any role except one where you think you might be able to nudge things in the right direction. But I could see how a judge says: well, that's how today's law works, that you have moral objections is something you can take to your favorite religious leader and lament about, not a court of law.

If I'm being fair, there isn't even really an invasion of privacy because PDFs don't have executable code (usually) that can track you. Rather, they need to hide it somewhere so that, if it appears on the pirate bay, they can read out the ID and see who the perpetrator is. More like a criminal investigation using a fingerprint on a glass, and less like a cookie actively sent with every action you perform on a website.

TL;DR: GDPR applies, but it probably doesn't make this database illegal. It's not a loophole by which a person can say no to literally everything. (Would be cool if you could require the police to stop using your fingerprint in a legitimate investigation.)

Still, if I were that sysadmin... I probably wouldn't 'drop table elsevier', but I'd rather live off government benefits than support that scheme.

Re: Sci-Hub statistics and database

#138

Earlier quoted context omitted.

Isn't it fantastic that we are alive and seeing the resurrection of the great library of Alexandria right before our eyes? She has done more than any other organization or individual in the history of mankind when helping people in second and third world country pursue advance research since the advent of internet. Well she and the people who pirate and distribute MS Office. Faculties around the world recommend scihu…

The great library of Alexandra, you mean.

Hahaha, I actually typo-ed Alexandra and trust me I googled that immediately. Apparently it is "Alexandria"

https://en.wikipedia.org/wiki/Library_of_Alexandria

Re: Sci-Hub statistics and database

#139
post #72

Earlier quoted context omitted.

That is not how scihub used to function. Scihub used to have an engine, named Plato, which would fetch papers automatically if not already in their database. For the last year now, this essential service has not been operational. This is what the issue I am raising is about.

It's clear what you're talking about. Software bitrots over time. Plato might need fixes, might have a huge backlog, lots of stuff can happen.

I still have airgapped windows xp systems with one software package on them to do one job. Left alone Software doesn’t bit rot, it is the never ending stream of updates that cause it to cease working over time.

Re: Sci-Hub statistics and database

#140
post #72

Earlier quoted context omitted.

It's clear what you're talking about. Software bitrots over time. Plato might need fixes, might have a huge backlog, lots of stuff can happen.

I still have airgapped windows xp systems with one software package on them to do one job. Left alone Software doesn’t bit rot, it is the never ending stream of updates that cause it to cease working over time.

Think about it. Software like Plato downloads papers from around the web. If the environment around software (i.e the web) changes, software without updates bitrots.
Post reply on HN