中文
相关论文

相关论文: MOSE: Monocular Semantic Reconstruction Using NeRF…

200 篇论文

Estimating a mesh from an unordered set of sparse, noisy 3D points is a challenging problem that requires carefully selected priors. Existing hand-crafted priors, such as smoothness regularizers, impose an undesirable trade-off between…

计算机视觉与模式识别 · 计算机科学 2020-06-03 Abhishek Badki , Orazio Gallo , Jan Kautz , Pradeep Sen

Monocular SLAM has received a lot of attention due to its simple RGB inputs and the lifting of complex sensor constraints. However, existing monocular SLAM systems are designed for bounded scenes, restricting the applicability of SLAM…

计算机视觉与模式识别 · 计算机科学 2024-03-11 Heng Zhou , Zhetao Guo , Shuhong Liu , Lechen Zhang , Qihao Wang , Yuxiang Ren , Mingrui Li

High-fidelity reconstruction of head avatars from monocular videos is highly desirable for virtual human applications, but it remains a challenge in the fields of computer graphics and computer vision. In this paper, we propose a two-phase…

图形学 · 计算机科学 2025-03-31 Pilseo Park , Ze Zhang , Michel Sarkis , Ning Bi , Xiaoming Liu , Yiying Tong

While 3D reconstruction is a well-established and widely explored research topic, semantic 3D reconstruction has only recently witnessed an increasing share of attention from the Computer Vision community. Semantic annotations allow in fact…

计算机视觉与模式识别 · 计算机科学 2017-08-28 Andrea Romanoni , Marco Ciccone , Francesco Visin , Matteo Matteucci

Monocular normal estimation aims to estimate the normal map from a single RGB image of an object under arbitrary lights. Existing methods rely on deep models to directly predict normal maps. However, they often suffer from 3D misalignment:…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Zongrui Li , Xinhua Ma , Minghui Hu , Yunqing Zhao , Yingchen Yu , Qian Zheng , Chang Liu , Xudong Jiang , Song Bai

Recovery of a 3D head model including the complete face and hair regions is still a challenging problem in computer vision and graphics. In this paper, we consider this problem using only a few multi-view portrait images as input. Previous…

计算机视觉与模式识别 · 计算机科学 2021-10-05 Xueying Wang , Yudong Guo , Zhongqi Yang , Juyong Zhang

Monocular 3D reconstruction for categorical objects heavily relies on accurately perceiving each object's pose. While gradient-based optimization in a NeRF framework updates the initial pose, this paper highlights that scale-depth ambiguity…

计算机视觉与模式识别 · 计算机科学 2024-07-16 Yuliang Guo , Abhinav Kumar , Cheng Zhao , Ruoyu Wang , Xinyu Huang , Liu Ren

In recent years, the neural implicit surface has emerged as a powerful representation for multi-view surface reconstruction due to its simplicity and state-of-the-art performance. However, reconstructing smooth and detailed surfaces in…

计算机视觉与模式识别 · 计算机科学 2024-07-12 Yuting Xiao , Jingwei Xu , Zehao Yu , Shenghua Gao

Recent advances in language and vision have demonstrated that scaling up model capacity consistently improves performance across diverse tasks. In 3D visual geometry reconstruction, large-scale training has likewise proven effective for…

计算机视觉与模式识别 · 计算机科学 2025-11-03 Jingnan Gao , Zhe Wang , Xianze Fang , Xingyu Ren , Zhuo Chen , Shengqi Liu , Yuhao Cheng , Jiangjing Lyu , Xiaokang Yang , Yichao Yan

Recovering 3D geometry and textures of individual objects is crucial for many robotics applications, such as manipulation, pose estimation, and autonomous driving. However, decomposing a target object from a complex background is…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Jun Wu , Sicheng Li , Sihui Ji , Yifei Yang , Yue Wang , Rong Xiong , Yiyi Liao

We tackle the problem of monocular 3D reconstruction of articulated objects like humans and animals. We contribute DensePose 3D, a method that can learn such reconstructions in a weakly supervised fashion from 2D image annotations only.…

计算机视觉与模式识别 · 计算机科学 2021-09-02 Roman Shapovalov , David Novotny , Benjamin Graham , Patrick Labatut , Andrea Vedaldi

Self-supervised learning is showing great promise for monocular depth estimation, using geometry as the only source of supervision. Depth networks are indeed capable of learning representations that relate visual appearance to 3D properties…

计算机视觉与模式识别 · 计算机科学 2020-02-28 Vitor Guizilini , Rui Hou , Jie Li , Rares Ambrus , Adrien Gaidon

Neural implicit surface reconstruction using volume rendering techniques has recently achieved significant advancements in creating high-fidelity surfaces from multiple 2D images. However, current methods primarily target scenes with…

计算机视觉与模式识别 · 计算机科学 2025-05-13 Lintao Xiang , Hongpei Zheng , Bailin Deng , Hujun Yin

Neural radiance fields (NeRFs) have enabled high fidelity 3D reconstruction from multiple 2D input views. However, a well-known drawback of NeRFs is the less-than-ideal performance under a small number of views, due to insufficient…

计算机视觉与模式识别 · 计算机科学 2023-03-27 Mikaela Angelina Uy , Ricardo Martin-Brualla , Leonidas Guibas , Ke Li

Neural implicit modeling permits to achieve impressive 3D reconstruction results on small objects, while it exhibits significant limitations in large indoor scenes. In this work, we propose a novel neural implicit modeling method that…

计算机视觉与模式识别 · 计算机科学 2023-09-14 Federico Lincetto , Gianluca Agresti , Mattia Rossi , Pietro Zanuttigh

Recent works on implicit neural representations have made significant strides. Learning implicit neural surfaces using volume rendering has gained popularity in multi-view reconstruction without 3D supervision. However, accurately…

计算机视觉与模式识别 · 计算机科学 2022-11-22 Decai Chen , Peng Zhang , Ingo Feldmann , Oliver Schreer , Peter Eisert

Estimating human pose and shape from monocular images is a long-standing problem in computer vision. Since the release of statistical body models, 3D human mesh recovery has been drawing broader attention. With the same goal of obtaining…

计算机视觉与模式识别 · 计算机科学 2024-01-03 Yating Tian , Hongwen Zhang , Yebin Liu , Limin Wang

We propose MoGe-2, an advanced open-domain geometry estimation model that recovers a metric scale 3D point map of a scene from a single image. Our method builds upon the recent monocular geometry estimation approach, MoGe, which predicts…

计算机视觉与模式识别 · 计算机科学 2025-07-04 Ruicheng Wang , Sicheng Xu , Yue Dong , Yu Deng , Jianfeng Xiang , Zelong Lv , Guangzhong Sun , Xin Tong , Jiaolong Yang

In recent years, reconstructing indoor scene geometry from multi-view images has achieved encouraging accomplishments. Current methods incorporate monocular priors into neural implicit surface models to achieve high-quality reconstructions.…

计算机视觉与模式识别 · 计算机科学 2025-01-03 Yulun Wu , Han Huang , Wenyuan Zhang , Chao Deng , Ge Gao , Ming Gu , Yu-Shen Liu

3D object detection based on roadside cameras is an additional way for autonomous driving to alleviate the challenges of occlusion and short perception range from vehicle cameras. Previous methods for roadside 3D object detection mainly…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Xiahan Chen , Mingjian Chen , Sanli Tang , Yi Niu , Jiang Zhu