Here's a challenge question to folks reading this and learned with the tools of the trade (my apologies in advance for somewhat hijacking the thread): consider this video of an endoscopy: https://www.youtube.com/watch?v=DUVDKoKSEkU -- say, from 3:00 to 5:00. And I have a bunch of movies (i.e., a series of images!) and I want to do a 3d reconstruction of this. It seems super, super difficult... there are free-flowing…
Atlas: End-to-End 3D Scene Reconstruction from Posed Images
21–26 of 26 posts
Re: Atlas: End-to-End 3D Scene Reconstruction from Posed Images
#22I wonder how long it's going to be before we're able to run a significant portion of Youtube video (tourist videos, etc) through something like this, and generate a huge 3d mesh of the world. Combined with Street View data, you'd really have a ton of spaces covered.
I have seen random still images used for this kind of thing: https://nerf-w.github.io/
I haven't heard of any equivalent of EXIF for video. That goes a long way when trying to make sense of random video both for camera settings as well as GPS location if you're trying to correlate multiple videos.
Re: Atlas: End-to-End 3D Scene Reconstruction from Posed Images
#23On a tangential thought, it's interesting to me that a company (magicleap) that has raised several billion dollars generates so little value compared to other companies its size that this is the most notable output from them in a year and I thought it was a phd project until I looked at the project owner. Anyways, it's a very interesting project and thanks for sharing.
The company has a cool name and the product area is divisive. Some say it is vapourware and nobody wants Oculus Rift style VR. Others are gung-ho. It's like Bitcoin all over again.
Although this tech is being done with AI, it was being done with non-AI approach two decades ago for movies/TV. But it wasn't as if people ported this tech to their smart phones from the SGI desktop monsters of yesteryear.
Re: Atlas: End-to-End 3D Scene Reconstruction from Posed Images
#24Here's a challenge question to folks reading this and learned with the tools of the trade (my apologies in advance for somewhat hijacking the thread): consider this video of an endoscopy: https://www.youtube.com/watch?v=DUVDKoKSEkU -- say, from 3:00 to 5:00. And I have a bunch of movies (i.e., a series of images!) and I want to do a 3d reconstruction of this. It seems super, super difficult... there are free-flowing…
Did they ever make it into real life practice?
Re: Atlas: End-to-End 3D Scene Reconstruction from Posed Images
#25Here's a challenge question to folks reading this and learned with the tools of the trade (my apologies in advance for somewhat hijacking the thread): consider this video of an endoscopy: https://www.youtube.com/watch?v=DUVDKoKSEkU -- say, from 3:00 to 5:00. And I have a bunch of movies (i.e., a series of images!) and I want to do a 3d reconstruction of this. It seems super, super difficult... there are free-flowing…
Even more hijacking, I remember thinking medical applications were going to be the killer apps for VR. I was blown away by these demos almost half a decade ago https://youtu.be/MWGBRsV9omw?t=251 Did they ever make it into real life practice?
https://www.youtube.com/c/okreylos/videos
5 years ago he was active in the Vive VR world
Re: Atlas: End-to-End 3D Scene Reconstruction from Posed Images
#26I wonder how long it's going to be before we're able to run a significant portion of Youtube video (tourist videos, etc) through something like this, and generate a huge 3d mesh of the world. Combined with Street View data, you'd really have a ton of spaces covered.
I believe random videos are too low of a quality. Like this, most of the stuff I've seen uses constrained videos. I have seen random still images used for this kind of thing: https://nerf-w.github.io/ I haven't heard of any equivalent of EXIF for video. That goes a long way when trying to make sense of random video both for camera settings as well as GPS location if you're trying to correlate multiple videos.