中文

RGB-D-Fusion: 类人主体的图像条件深度扩散

计算机视觉与模式识别 2023-09-25 v1 机器学习

摘要

我们提出RGB-D-Fusion,一种多模态条件去噪扩散概率模型,用于从类人主体的低分辨率单目RGB图像生成高分辨率深度图。RGB-D-Fusion首先使用图像条件去噪扩散概率模型生成低分辨率深度图,随后使用以低分辨率RGB-D图像为条件的第二个去噪扩散概率模型对深度图进行上采样。我们进一步引入一种新颖的增强技术——深度噪声增强,以提升我们超分辨率模型的鲁棒性。

关键词

引用

@article{arxiv.2307.15988,
  title  = {RGB-D-Fusion: Image Conditioned Depth Diffusion of Humanoid Subjects},
  author = {Sascha Kirch and Valeria Olyunina and Jan Ondřej and Rafael Pagés and Sergio Martin and Clara Pérez-Molina},
  journal= {arXiv preprint arXiv:2307.15988},
  year   = {2023}
}