Live data from Hacker News

ArXiv receives $10M for upgrades

news.cornell.edu

121–130 of 164 posts

Re: ArXiv receives $10M for upgrades

#121
post #25

Earlier quoted context omitted.

I feel this is already true of books, the question is will a publisher put it out... books cost money to produce and by comparision to hosting are massive investments.

I wish more self-published books would find a way to hire a copyeditor.

gpt to the rescue?

Re: ArXiv receives $10M for upgrades

#122
Arxiv is a blessing, I would not be where I am without Arxiv and all the folks openly sharing their research.

I don’t like that deepmind gates their research behind nature and other publications. I acknowledge it puts them in the cool kids club but it feels unauthentic. Putting research out in the open moves humanity forward.

Arxiv can say it made a little dent in the universe and I’m glad it’s getting the funding it needs to be sustainable.

I just hope they don’t burn it out trying to scale needlessly.

Re: ArXiv receives $10M for upgrades

#123
post #24

I am really glad arXiv is getting more funding. It is an essential resource. For me personally, it’ll be really interesting to see if the frontend changes. No doubt there are some important improvements that can be made (a website can always be made more accessible, moderation tools, support for name changes as mentioned, etc.) However, to first approximation, arXiV’s website already seems almost like a platonic idea…

Simple HTML, loads fast, has the information and features you need but otherwise gets out of your way Exactly. I'll be furious if it turns into another trendchasing JS-required SPA. By all means use the funds to pay for additional bandwidth or storage, but don't waste it on useless hostile UI changes.

I had this same thought but was reluctatnt to state it as it feels like unnecessary pessimism. But honestly, this website works flawlessly. The last thing it needs is software developers trying to keep themselves entertained or impress people. If it ain't broke...

Re: ArXiv receives $10M for upgrades

#124

Earlier quoted context omitted.

There's sales tax for customers, income (and payroll) tax for employees, and then taxes on dividends, and then capital gains on stocks. Low-margin SMEs contribute proportionately very little tax compared to Google when we account for tax contribution this way.

Employees payroll and income taxes come before margin?

I guess you're referring to my specification of low-margin. As I also specified SME, what that means is both low-margin and low-revenue. And the true intented meaning is that all cash flows one might want to tax (mostly with progressive rates) are small.

Re: ArXiv receives $10M for upgrades

#125
post #15

Oh that’s great to hear! I really like arXiv and am glad it’s getting more funding! I hope this will enable arXiv to fix some of the long-time issues it has, including the way identity and attribution are handled. For example: the current system assumes people’s names do not change [1], which in particular negatively affects trans authors. This seems like it could be fixed if they had the bandwidth to, and I hope the…

You must not change history. You can use a new name, but that should be considered a new and separate identity.

Agreed, what is needed is a way to link the old and new identities so others can know that it is the same person.

Re: ArXiv receives $10M for upgrades

#126

Earlier quoted context omitted.

Simple HTML, loads fast, has the information and features you need but otherwise gets out of your way Exactly. I'll be furious if it turns into another trendchasing JS-required SPA. By all means use the funds to pay for additional bandwidth or storage, but don't waste it on useless hostile UI changes.

I had this same thought but was reluctatnt to state it as it feels like unnecessary pessimism. But honestly, this website works flawlessly. The last thing it needs is software developers trying to keep themselves entertained or impress people. If it ain't broke...

Looking at how 95% of modern web sites appear today, the fear is grounded.

Re: ArXiv receives $10M for upgrades

#127

Earlier quoted context omitted.

Loads fast, except for the PDFs which load quite slowly

You might want to check out this site: https://www.arxiv-vanity.com/ I wish they would integrate something like that directly.

They have https://ar5iv.labs.arxiv.org/ in beta which has similar functionality.

Re: ArXiv receives $10M for upgrades

#128
I hope they just use the money for more bandwidth and maybe spend it on free api/download access so people can get big chunks of it easily and build all kinds of things (from alert on keyword to just stream of text to finetune models)

And I hope they don't touch the frontend at all :)

Re: ArXiv receives $10M for upgrades

#129
post #15

Oh that’s great to hear! I really like arXiv and am glad it’s getting more funding! I hope this will enable arXiv to fix some of the long-time issues it has, including the way identity and attribution are handled. For example: the current system assumes people’s names do not change [1], which in particular negatively affects trans authors. This seems like it could be fixed if they had the bandwidth to, and I hope the…

There is no good way to deal with name changes. You can change them in the metadata, but the old ones will still be around in the paper itself. You can let authors change them in the paper, but the old ones will still be around in others' bibliographies. You can't do much about the bibliographies, since (1) you'd need an automated way to find them all and disambiguate them (have fun with Chinese names) and (2) a lot…

For disambiguation there's ORCID, and for the citation there is nowadays the DOI to identify the paper. So long as you can update whatever the DOI points out with a note it seems like this is handled?

Re: ArXiv receives $10M for upgrades

#130
post #34

Earlier quoted context omitted.

Have to download the file, then load it on the client. Not sure you can make it much faster.

If the pdfs were served by some of the megacaps who posts tons of papers (e.g. Google or Facebook) then it would be an order of magnitude faster. And said megacaps would end up spending peanuts relative to the value they get from arXiv.

An order of magnitude faster sounds unimportant. Is all the time spent by researchers sitting around waiting for 0.4s for a paper to download really an issue?
Post reply on HN