RGB-D-Fusion: 类人主体的图像条件深度扩散
计算机视觉与模式识别
2023-09-25 v1 机器学习
摘要
我们提出RGB-D-Fusion,一种多模态条件去噪扩散概率模型,用于从类人主体的低分辨率单目RGB图像生成高分辨率深度图。RGB-D-Fusion首先使用图像条件去噪扩散概率模型生成低分辨率深度图,随后使用以低分辨率RGB-D图像为条件的第二个去噪扩散概率模型对深度图进行上采样。我们进一步引入一种新颖的增强技术——深度噪声增强,以提升我们超分辨率模型的鲁棒性。
引用
@article{arxiv.2307.15988,
title = {RGB-D-Fusion: Image Conditioned Depth Diffusion of Humanoid Subjects},
author = {Sascha Kirch and Valeria Olyunina and Jan Ondřej and Rafael Pagés and Sergio Martin and Clara Pérez-Molina},
journal= {arXiv preprint arXiv:2307.15988},
year = {2023}
}