Earlier quoted context omitted.
Have a look at the google drive files if you click on the data link; it's a lot more than 2 photos. eg. That dinosaur skeleton is derived from 60 photos. The drumkit comes from ~100. ...so it's not magic, it's very close to what you get from standard photogrammetry. The big part of this is that it isn't representing the scene as block of voxels like some other approaches. > The biggest practical tradeoffs between the…
Also I do think the 3d voxel reconstruction approach and the nerf approach solves different goals. I didn't read the original nerf paper thoroughly but AFAIK the network learns to interpolate between the photos in a beautiful, smooth way, but the voxel representation would allow a lot of other reconstructions.
Re: Neural Databases
#31If the photos span the camera's full position-orientation vector space, I don't see why you can't put the camera anywhere in the scene.