Earlier quoted context omitted.
True, but my point was more a scaling factor. Google can crank N amount of data with its small fleet, Tesla will be able to reach 10^5 N soon. If they're able to process it properly they may pass Google tests quickly.
I don't know if it's that easy. There are tons of images and text laying around, but research is focused on just a few datasets. Sometimes it's hard to make use of 100x more data.
In this case, what you need is mostly "what would most humans do?"
There would be things to refine about that (e.g. prevent speeding; analyse how humans reacted right before crashes etc. and improve on responses), but as a starting point it is immensely useful.