Earlier quoted context omitted.
> Similarly you can train a neural network to separate the speaker on a video call from the background, so you just don't need a depth camera. These still look so fake, and they tend to blur out objects that you're trying to hold up in the video, that I actually created my own virtual camera that blurs progressively more over depths based on RealSense measured depth and looks far more realistic. https://github.com/dh…
They're not that good but people will tolerate poor quality.
RealSense was used for industrial operations, I personally was looking into them for packing items in transport containers (specific to the factory involved). Poor quality of depth information would mean jams involving robot capable of goring through industrial enclosures, printers, and maintenance engineers.