Earlier quoted context omitted.
It's scary when you consider two other factors: First, the AI hype train. People think that calling something "Artificial Intelligence" implies that it is artificial, yes, but also, critically, that it is intelligent. Many enthusiastic people, and also many policymakers, don't fully realize the extent to which machine learning is constrained by both the quality and nature of its training data, and the capabilities of…
Given that nuclear power is significantly safer than other forms of power, are you asserting that the risks of self-driving cars are more about PR and perception than actual risk?
A popular self-driving car dataset is missing labels for hundreds of pedestrians
31–40 of 202 posts
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#32Earlier quoted context omitted.
It's scary when you consider two other factors: First, the AI hype train. People think that calling something "Artificial Intelligence" implies that it is artificial, yes, but also, critically, that it is intelligent. Many enthusiastic people, and also many policymakers, don't fully realize the extent to which machine learning is constrained by both the quality and nature of its training data, and the capabilities of…
Given that nuclear power is significantly safer than other forms of power, are you asserting that the risks of self-driving cars are more about PR and perception than actual risk?
Maybe a better analogy is driving a car, that is, nuclear power and self-driving cars are both vaguely like driving a car.
However, my intuition is that nuclear power has less variables than driving a car. I say intuition because I don't know much about neither.
I think the point is that self-driving cars trained on poor data are no better than poorly trained superhumans at driving.
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#33Well, if an autonomous vehicle outfit were running unmonitored Level 4 vehicles on public roads using only an open source data set I'd be worried. Even if it was labelled thoroughly and correctly, there isn't nearly enough data in any open source dataset to train an autonomous vehicle perception system that can operate safely without human supervision. This is not a safety critical issue.
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#34Earlier quoted context omitted.
A line of parked cars absolutely needs to be labeled as individual cars. Any one of them could pull out in front of you at any moment.
Indeed, whilst additionally presenting the chance of a door opening or an obscured pedestrian stepping out from between them.
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#35It is a self-correcting problem: these pedestrains won’t be present in the next dataset.
People should learn not to go outside if they're not labelled.
I worry what will happen when this idea breeds with the "why worry about privacy if you've got nothing to hide?" fallacy.
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#36Earlier quoted context omitted.
> And while a single frame might be missing a label, I bet that at-speed most everything important gets labeled correctly enough to be better than a distracted human driver, or the average human driver for that matter. Would you bet a family member? That a distracted driver is a hazard does not mean other drivers are safe, or even saf er .
>Would you bet a family member? "think of the children!" Lives at stake don't change anything here. The question is whether self-driving cars, even with the errors, are safer for people than regular drivers on average. If so, then absolutely yes everyone should bet their lives and their families'. Thousands of people are dying every day in cars. This is not something we need to wait for it to be perfect. It only need…
It only needs to be better.
A Pedestrian likely has a different definition of "better" than the car driver.
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#37This is really scary. I discovered this because we're working on converting and re-hosting popular datasets in many popular formats for easy use across models... I first noticed that there were a bunch of completely unlabeled images. Upon digging in, I was appalled that fully 1/3 of the images contained errors or omissions! Some are small (eg a part of a car on the edge of the frame or a ways in the distance not bein…
I understand your concern and share it myself. This is an important time and we should be really careful training these things. However, training as used in the real world isn't on a still frame only basis, it's used in sequence. And while a single frame might be missing a label, I bet that at-speed most everything important gets labeled correctly enough to be better than a distracted human driver, or the average hum…
If you look at public datasets, it's more often that things are incorrectly tagged/labeled rather than correctly.
Entropy is a real thing.
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#38Earlier quoted context omitted.
> And while a single frame might be missing a label, I bet that at-speed most everything important gets labeled correctly enough to be better than a distracted human driver, or the average human driver for that matter. Would you bet a family member? That a distracted driver is a hazard does not mean other drivers are safe, or even saf er .
>Would you bet a family member? "think of the children!" Lives at stake don't change anything here. The question is whether self-driving cars, even with the errors, are safer for people than regular drivers on average. If so, then absolutely yes everyone should bet their lives and their families'. Thousands of people are dying every day in cars. This is not something we need to wait for it to be perfect. It only need…
Maybe logically that makes sense but from an ethical perspective I argue it's much more complicated than that (e.g. the trolley problem)
In the current system if a human is at fault, they take the blame for the accident. If we decide to move to self driving cars that we know are far from perfect but statistically better than humans, who do we blame when an accident inevitably happens? Do we blame the manufacturer even though their system is operating within the limits they've advertised?
Or do we just say well, it's better than it used to be and it's no one's fault? When the systems become significantly better than humans, I can see this perhaps being a reasonable argument, but if it's just slightly better, I'm not sure people will be convinced.
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#39It is a self-correcting problem: these pedestrains won’t be present in the next dataset.
People should learn not to go outside if they're not labelled.
Re: A popular self-driving car dataset is missing labels for hundreds of pedestrians
#40Earlier quoted context omitted.
It's scary not because this specific dataset was used to train Teslas that are on the road today. Rather, because it makes us aware of an entire class of errors that most of us probably hadn't thought about before. I guess you are absolutely certain that training data used in production cars will be free of these issues, but it's not clear why.
It does not make use aware of a new class of errors. Labeling issues is nothing new, but plenty of systems trained on them continue to work just fine. This is FUD.