arXiv · 2307.15988
RGB-D-Fusion: Image Conditioned Depth Diffusion of Humanoid Subjects
Abstract
We present RGB-D-Fusion, a multi-modal conditional denoising diffusion probabilistic model to generate high resolution depth maps from low-resolution monocular RGB images of humanoid subjects. RGB-D-Fusion first generates a low-resolution depth map using an image conditioned denoising diffusion probabilistic model and then upsamples the depth map using a second denoising diffusion probabilistic model conditioned on a low-resolution RGB-D image. We further introduce a novel augmentation technique, depth noise augmentation, to increase the robustness of our super-resolution model.
Explore related subjects
Keep this discovery
Sascha Kirch, Valeria Olyunina, Jan Ondřej, Rafael Pagés, Sergio Martin, Clara Pérez-Molina. 2023-07-29. RGB-D-Fusion: Image Conditioned Depth Diffusion of Humanoid Subjects. https://doi.org/10.1109/access.2023.3312017
Cite the original work for its findings. Save a collection to share your selection of sources.