Live data from Hacker News

Lyft releases self-driving research dataset

medium.com

1–10 of 123 posts

Re: Lyft releases self-driving research dataset

#3
The post indicates there is a competition and prizes but I'm not seeing any discussion of what sort of license the data is being made available under (or the competition for that matter). Hopefully it's there and I'm just not seeing it.

Re: Lyft releases self-driving research dataset

#4
post #3

The post indicates there is a competition and prizes but I'm not seeing any discussion of what sort of license the data is being made available under (or the competition for that matter). Hopefully it's there and I'm just not seeing it.

The Github they link to says it's under the CC BY-NC-SA 4.0.

Re: Lyft releases self-driving research dataset

#5
> Academic research accelerates innovation, but it requires costly data that is out of reach for most academic teams.

This is true of pretty much any AI research. Look at Puffer[0], which was just on HN a couple of days ago. They're running a free streaming service just to get enough data to train their algorithms, and in fact mention in their FAQ that they would love to use commercial data if they could get it.

Unfortunately, academic and commercial incentives don't really align here. Most commercial entities don't want to share their data because it's valuable to them, and if they let researchers in, they want the output of the research to remain proprietary to their commercial enterprise.

I wonder if there isn't some sort of governance solution to this. Like give companies big tax breaks for sharing their data with researchers, or something like that. Essentially subsidize academia indirectly.

[0] https://puffer.stanford.edu/player/

Re: Lyft releases self-driving research dataset

#7
post #5

> Academic research accelerates innovation, but it requires costly data that is out of reach for most academic teams. This is true of pretty much any AI research. Look at Puffer[0], which was just on HN a couple of days ago. They're running a free streaming service just to get enough data to train their algorithms, and in fact mention in their FAQ that they would love to use commercial data if they could get it. Unfo…

I've seen semiconductor industry companies collaborate on grant-funding fundamental condensed-matter physics research. If it is a question of interest to all parties, and the work is too blue-sky to be immediately profitable, sometimes they'll fund the work.

Re: Lyft releases self-driving research dataset

#8
“There will be $25,000 in prizes, and we’ll be flying the top researchers to the NeurIPS Conference in December, as well as allowing the winners to interview with our team.”

I guess it’s a decent opportunity if you’re trying to break into DL?

Re: Lyft releases self-driving research dataset

#9
post #6

$25k in prizes seems silly given this is a multi-billion dollar market to crack.

Such competitions do not usually result in a comprehensive "solution" by themselves - pushing the state-of-the-art is more common. Also the value is not going to be derived solely from the algorithm but more from its deployment to real world applications and the surrounding infrastructure to make it possible.

Re: Lyft releases self-driving research dataset

#10
post #5

> Academic research accelerates innovation, but it requires costly data that is out of reach for most academic teams. This is true of pretty much any AI research. Look at Puffer[0], which was just on HN a couple of days ago. They're running a free streaming service just to get enough data to train their algorithms, and in fact mention in their FAQ that they would love to use commercial data if they could get it. Unfo…

Slightly tangential but might this be another argument for people "owning" their data while companies "own" the processing procedures of it. If people "owned" their data it would presumably be much easier for them to give it out for research purposes
Post reply on HN