Live data from Hacker News

Why isn't preprint review being adopted?

theroadgoeson.com

131–140 of 169 posts

Re: Why isn't preprint review being adopted?

#131
post #120

Earlier quoted context omitted.

You are mistaken in thinking that editors are supposed to vet submissions. That's not their job! Their job is only to weed out crackpot pseudoscience, which is a much easier task generally and also mostly solved by reputation, or papers outside of the journal's scope. And to edit the remaining papers into a constantly formatted journal publication. And that's done easy enough these days with LaTeX and online printing…

> Only now we can accept all 50, because why not? Because the downside to creating an ever growing haystack is that it becomes increasingly difficult to find a needle. Making it easier to create a deluge of bad research won’t help me find the worthwhile research that would actually help me in my job. If I had the choice between collecting “all the data” and just collecting “the really good and relevant data” I’m opti…

You just use better tools to manage it. Fine-tuned LLMs and Google Scholar like search engines help here.

To stretch an analogy it is like email. The job of the editor is the same as the spam detection service run by hosted email providers. They actually go in and actively hide scams and worthless ad email from you, and we thank them for it. Some email providers have recently started offering "focused inbox" modes where they prioritize emails for you too. I don't use that, but I could see why some people do. But importantly they don't block email based on those heuristics, like they might do for spam. You still get non-priority emails. But imagine a world where gmail straight up blocked/rejected email which it didn't consider priority. Would you want that?

The situation with journals is comparable. Editors have a spam/crank detection duty, but they shouldn't be rejecting manuscripts beyond that.

Re: Why isn't preprint review being adopted?

#132
post #120

Earlier quoted context omitted.

> Only now we can accept all 50, because why not? Because the downside to creating an ever growing haystack is that it becomes increasingly difficult to find a needle. Making it easier to create a deluge of bad research won’t help me find the worthwhile research that would actually help me in my job. If I had the choice between collecting “all the data” and just collecting “the really good and relevant data” I’m opti…

You just use better tools to manage it. Fine-tuned LLMs and Google Scholar like search engines help here. To stretch an analogy it is like email. The job of the editor is the same as the spam detection service run by hosted email providers. They actually go in and actively hide scams and worthless ad email from you, and we thank them for it. Some email providers have recently started offering "focused inbox" modes wh…

What you’re describing is essentially an arms race in quantity. Yes, we can use tools to help sort, but those same tools can also be used to deluge the inbox and obfuscate the bad. In fact, one of the best ways to sort is by using specific journals/journal metrics as a proxy for quality. That is much, much easier (and productive) than trying to sort based on some Google scholar advanced query. For example, it's much easier for a journal to retract an article that was shown an inability to replicate than to create a search to do the same.

The tone of your comment is very techno-optimist, which is very on brand for HN. In that view, every problem is solved by technology, even those that are created by technology. I would argue there are some problems that are better solved with less technology, not more.

Re: Why isn't preprint review being adopted?

#133
post #104

Earlier quoted context omitted.

A valid concern, but what is driving this red-queen race is that you are competing with people who cheat. If you’ve got one really solid paper published, but your competitor fraudulently published 10, it’s real tempting to fraud a few papers yourself. But they multiply fraudulent papers because they can get away with it, and they can get away with it because nobody is really reviewing or replicating those results. I…

The idea that fraud is rampant and driving the explosion in publication rates is deeply tempting to folks who don’t work directly in a scientific field. But it’s a misconception that’s largely driven by non-scientific media and bloggers. In practice literal fraud does exist, but it’s relatively rare in most scientific fields. The explosion in publication rates is largely caused by a combination of (1) more people ent…

> “paper slicing”

I'm glad we agree that paper slicing is also plaguing academia. When I was a grad student, it bugged me to no end that those guys at Stanford were out-publishing me--because they got enough results for 1.5 papers, and then made (3 choose 2) papers out of them.

And yeah, if you think I wasn't tempted to follow suit, think again. I eventually left academia because I didn't want to cheat, or compete with cheaters, when we were being graded on the curve of "publish or perish".

> fraud is rampant and driving the explosion in publication rates is deeply tempting to folks who don’t work directly in a scientific field.

So, you are an assistant prof who was passed over for tenure, or you didn't get that Harvard appointment--it went to Francesca Gino, whose vita looks sooooo much better than yours, because she's doing TED talks to promote her book called "Why It Pays To Break The Rules In Work And Life". She's making $1 million a year consulting, while you are trying to get funding for your research, which isn't as flashy but at least is real science...

... if you are graded on the curve, how will you look against someone who cheated? It's the prisoner's dilemma.

> How is anyone going to replicate a result if it’s not published?

In my proposal, it works the same that way can they review a paper if it's not published. You write up a paper, detailing the claims you are making and the experiments and methods you used to justify your claims, and you submit it for publication.

Then, 3 or 4 anonymous reviewers decide whether it's promising enough to go to the next step: hire another research group to replicate the results. If the results replicate, then and only then are they published.

Yes, it's more expensive, so to finance it I propose that when the grant is being written, the principle investigators should also estimate what it would cost to hire reviewers and to fund experiments at another lab to confirm the results.

When the grant is approved, that money is put in escrow until the research is done and submitted for publication, and disbursed to the reviewers and replicators to compensate them for pausing their own research and evaluating someone else's. If it replicates, it's published. If it doesn't replicate, you don't get to write up books on how to cheat your way to the top while you are cheating your way to the top.

> This is literally the reason that scientific publication exists ...

100% agree, its great to share new results and all---but if a "result" doesn't replicate, its not a result. What, exactly, are Gino's peers supposed to learn from her fraudulent papers? If I'm trying to decide what to research, how do I know which lines are actually promising or not? Do I just go with researchers at a name-brand university (like Gino at Harvard?)

These days papers can be written by machine as fast a a machine gun fires bullets. There's got to be some way of separating the signs from the noise.

> brownie-point credit for publications...

Publication is very important; one of my professors explained it this way: if you don't publish your results--i.e. you don't convince your peers that what you did is worthy of publishing--its like you didn't do anything at all. Then whole idea of research is to contribute to the edifice of science, and publishing is the vehicle by which that contribution is made. It's how you deliver the contribution to everybody else. And peer-review is how it is determined whether you actually made a contribution.

So the solution can't be to just stop caring about how many papers are published.

> Taken literally, your “no publication without replication” proposal would inhibit scientific replication entirely.

I hope my explanation above addresses this concern....

Re: Why isn't preprint review being adopted?

#134
post #67

Earlier quoted context omitted.

> Just make it accessible. We already have that system, it's called the internet. Nothing stops you or I from putting our ideas online for all to read, comment on, update, etc. The role of the publishers, flawed as it is, has little to do with the physical cost of producing or providing an article, and is filling (one can argue badly) a role in curation and archival that is clearly needed. Any proposal to change the…

That's just not true. Most publicly fundes research is hidden behind paywalls.

No it's exactly true. You can write up anything you want and put it on a site. The post I was replying to was suggesting an open access system (both read and write) for exchanging ideas. This exists.

What it doesn't do is effectively replace the non-open system for access to academic journals. I have a lot of sympathy for open (read) access to research, particularly publicly funded. It just isn't sensible to wave a wand and say "all papers are free to read now" without some plan for the other parts of the system and the ecosystem (academic research) that relies on it.

Re: Why isn't preprint review being adopted?

#135
post #97

Earlier quoted context omitted.

No, although the mechanism for hosting the content aren’t that important. Preprint servers are very useful but haven’t replaced journals for good reasons.

What are those reasons? The only thing I see that journals do which preprint servers couldn't easily take over is prestige .

You have the causality wrong. Prestige comes to journals by doing a good (or at least, better than peers) job of being a journal, which is providing a necessary function to the academic research process. If you want to improve on that system, you have to improve on those functions, or reduce the reliance on them by providing something better.

Put it another way, if you can design a system with a better ROC curve for classifying research, with a better TP rate for good papers, and have it cost less in real terms that current academic papers, then you are on to something. If all you've got is "papers should be free" or "it's too hard to access publishing from the outside" what you have are complaints, not solutions.

Re: Why isn't preprint review being adopted?

#136

Earlier quoted context omitted.

Quality control was handled fine enough by editors for literally all output on the planet pre 1970s. There’s nothing physically stopping that paradigm from returning now that we have the internet, other than the fact that there’s only a few thousand (?) such editors. If somehow there were fifty thousand such editors, then the whole peer review system would be completely unnecessary. Of course not enough people want t…

There are nearly 50,000 commercial journals and a long tail of non-commercial journals each with teams of editors. There are probably hundreds of thousands of people currently serving as editors. The issue isn't editorial bandwidth, it's that peer review is currently built into the promotion and tenure structure for academics, who produce the vast majority of scholarship and thus dictate the shape of the scholarly pu…

I’m including only actually competent, full time editors, with sufficiently high reputation that their decisions will be taken seriously. There’s definitely not 50000 of those.

A huge number of journals by numerical count, along with their ‘editors’, are literally laughed at in many fields.

As you’ve mentioned, trying to expand the actually reputable number by 10x, 20x, etc… is a huge problem.

Hence it has to be paid for, quite highly paid for, otherwise the coordination problem is probably impossibly difficult.

Re: Why isn't preprint review being adopted?

#137
post #104

Earlier quoted context omitted.

A valid concern, but what is driving this red-queen race is that you are competing with people who cheat. If you’ve got one really solid paper published, but your competitor fraudulently published 10, it’s real tempting to fraud a few papers yourself. But they multiply fraudulent papers because they can get away with it, and they can get away with it because nobody is really reviewing or replicating those results. I…

The idea that fraud is rampant and driving the explosion in publication rates is deeply tempting to folks who don’t work directly in a scientific field. But it’s a misconception that’s largely driven by non-scientific media and bloggers. In practice literal fraud does exist, but it’s relatively rare in most scientific fields. The explosion in publication rates is largely caused by a combination of (1) more people ent…

Nitpick: the term is salami slicing.

https://en.wikipedia.org/wiki/Salami_slicing_tactics#Salami_...

Re: Why isn't preprint review being adopted?

#139
post #120

Earlier quoted context omitted.

> Only now we can accept all 50, because why not? Because the downside to creating an ever growing haystack is that it becomes increasingly difficult to find a needle. Making it easier to create a deluge of bad research won’t help me find the worthwhile research that would actually help me in my job. If I had the choice between collecting “all the data” and just collecting “the really good and relevant data” I’m opti…

You just use better tools to manage it. Fine-tuned LLMs and Google Scholar like search engines help here. To stretch an analogy it is like email. The job of the editor is the same as the spam detection service run by hosted email providers. They actually go in and actively hide scams and worthless ad email from you, and we thank them for it. Some email providers have recently started offering "focused inbox" modes wh…

> Editors have a spam/crank detection duty, but they shouldn't be rejecting manuscripts beyond that.

If the system is working, publication in a reputable journal serves as a useful, albeit imperfect, indicator of scientific quality.

Top journals shouldn't be publishing deeply flawed work, or even decent work in clear need of a rewrite. It's not just about spam and cranks.

Re: Why isn't preprint review being adopted?

#140
post #10

One of my profs once remarked, "All of science is done on a volunteer basis." He was talking about peer review, which--as crucial as it is--is not something you get paid for. Writing a review--a good review--is 1) hard work, 2) can only be done by somebody who has spent years in postgraduate study, and 3) takes up a lot of time, which has many other demands on it. The solution? Its obvious. In a free market, how do y…

And also has limited value as we practice it today. A useful review would involve: (a) "This paper won't be accepted by Cochrane for meta analysis", "N=20 get out of here", ... (b) Researchers provided their data files and Jupyter notebooks, the reviewer got them to run (c) Reviewers attempt their own analysis for at least some of the data (think of the model of accounting where auditors look at a sample of the books…

> If you believe in meta-analysis, which you should,

I view meta-analysis as being like those mortgage-backed securities which crashed the world back in 2008. I mean, yeah, theoretically, a bond whose yield is a weighted average of mortgages should be less risky than any one of the mortgages in it.

But....the devil is in the details. When I start seeing meta-analysis which claim that, e.g., masks don't protect you from a respiratory illness which spreads by coughing and sneezing, or that vaccines are at best worthless and at worse cause autism, wellll.....

....I have to conclude that garbage-in, garbage-out.

Post reply on HN