I acknowledge the issues in the dataset and that it has a lot of stars on github because it's from Udacity; but calling it 'a popular self-driving car dataset' is misleading as it implies this dataset is popularly used for self-driving cars when it is in fact only a small dataset Udacity uses to teach the basics of training neural networks for self-driving cars. I've been involved in the autonomous vehicle industry f…
Are these larger datasets routinely subject to the same kind of inspection this titanic.csv of self-driving car datasets?
It'd be interesting if the NHTSB had a held-back "test set" they used to evaluate self driving cars before letting them on the road.