English
Related papers

Related papers: LucidFusion: Reconstructing 3D Gaussians with Arbi…

200 papers

Reconstructing dynamic 3D scenes from monocular video has broad applications in AR/VR, robotics, and autonomous navigation, but often fails due to severe motion blur caused by camera and object motion. Existing methods commonly follow a…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Zhijing Wu , Longguang Wang

In this paper, we introduce 3D-GMNet, a deep neural network for 3D object shape reconstruction from a single image. As the name suggests, 3D-GMNet recovers 3D shape as a Gaussian mixture. In contrast to voxels, point clouds, or meshes, a…

Computer Vision and Pattern Recognition · Computer Science 2020-08-18 Kohei Yamashita , Shohei Nobuhara , Ko Nishino

Human perceive the 3D world through 2D observations from limited viewpoints. While recent feed-forward generalizable 3D reconstruction models excel at recovering 3D structures from sparse images, their representations are often confined to…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Mochu Xiang , Zhelun Shen , Xuesong Li , Jiahui Ren , Jing Zhang , Chen Zhao , Shanshan Liu , Haocheng Feng , Jingdong Wang , Yuchao Dai

Recovering physical properties of objects in motion is a core task across scientific and industrial applications. When the relative motion between the object and the sensing apparatus provides sufficient angular coverage, Computerized…

Numerical Analysis · Mathematics 2026-05-19 Daniel Burrows , Can Evren Yarman , Ozan Öktem

Modern feed-forward 3D reconstruction methods like VGGT predict pixel-aligned pointmaps in camera-centric coordinate frames. However, this choice of coordinate frame is not always optimal. We propose instead to predict pointmaps in upright,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-27 Bharath Raj Nagoor Kani , Noah Snavely

In recent years, neural rendering methods such as NeRFs and 3D Gaussian Splatting (3DGS) have made significant progress in scene reconstruction and novel view synthesis. However, they heavily rely on preprocessed camera poses and 3D…

Graphics · Computer Science 2025-07-01 Chenhao Zhang , Yezhi Shen , Fengqing Zhu

Accurate surface reconstruction from unposed images is crucial for efficient 3D object or scene creation. However, it remains challenging, particularly for the joint camera pose estimation. Previous approaches have achieved impressive…

Computer Vision and Pattern Recognition · Computer Science 2026-01-29 Li-Heng Chen , Zi-Xin Zou , Chang Liu , Tianjiao Jing , Yan-Pei Cao , Shi-Sheng Huang , Hongbo Fu , Hua Huang

Recent advances in 3D Gaussian Splatting (3DGS) have enabled high-quality, real-time novel-view synthesis from multi-view images. However, most existing methods assume the object is captured in a single, static pose, resulting in incomplete…

Computer Vision and Pattern Recognition · Computer Science 2025-10-20 Ting-Yu Yen , Yu-Sheng Chiu , Shih-Hsuan Hung , Peter Wonka , Hung-Kuo Chu

We present a fast and efficient volumetric capture and reconstruction system that processes either RGB-D or RGB-only input to generate 3D representations in the form of point clouds and Gaussian splats. For Gaussian splat reconstructions,…

Graphics · Computer Science 2025-12-19 Athanasios Charisoudis , Simone Croci , Lam Kit Yung , Pascal Frossard , Aljosa Smolic

Large Reconstruction Models have made significant strides in the realm of automated 3D content generation from single or multiple input images. Despite their success, these models often produce 3D meshes with geometric inaccuracies,…

Computer Vision and Pattern Recognition · Computer Science 2024-05-27 Ruikai Cui , Xibin Song , Weixuan Sun , Senbo Wang , Weizhe Liu , Shenzhou Chen , Taizhang Shang , Yang Li , Nick Barnes , Hongdong Li , Pan Ji

Real-time multi-agent collaboration for ego-motion estimation and high-fidelity 3D reconstruction is vital for scalable spatial intelligence. However, traditional methods produce sparse, low-detail maps, while recent dense mapping…

Computer Vision and Pattern Recognition · Computer Science 2024-12-16 Xiaohao Xu , Feng Xue , Shibo Zhao , Yike Pan , Sebastian Scherer , Xiaonan Huang

Monocular object pose estimation, as a pivotal task in computer vision and robotics, heavily depends on accurate 2D-3D correspondences, which often demand costly CAD models that may not be readily available. Object 3D reconstruction methods…

Computer Vision and Pattern Recognition · Computer Science 2024-09-05 Luqing Luo , Shichu Sun , Jiangang Yang , Linfang Zheng , Jinwei Du , Jian Liu

Open-vocabulary panoptic reconstruction is essential for advanced robotics perception and simulation. However, existing methods based on 3D Gaussian Splatting (3DGS) often struggle to simultaneously achieve geometric accuracy, coherent…

Robotics · Computer Science 2026-04-14 Xuan Yu , Yuxuan Xie , Changjian Jiang , Shichao Zhai , Rong Xiong , Yu Zhang , Yue Wang

In this paper, we propose an efficient method for robust 3D self-portraits using a single RGBD camera. Benefiting from the proposed PIFusion and lightweight bundle adjustment algorithm, our method can generate detailed 3D self-portraits in…

Computer Vision and Pattern Recognition · Computer Science 2020-04-07 Zhe Li , Tao Yu , Chuanyu Pan , Zerong Zheng , Yebin Liu

Efficiently synthesizing novel views from sparse inputs while maintaining accuracy remains a critical challenge in 3D reconstruction. While advanced techniques like radiance fields and 3D Gaussian Splatting achieve rendering quality and…

Computer Vision and Pattern Recognition · Computer Science 2025-01-22 Chenlu Zhan , Yufei Zhang , Yu Lin , Gaoang Wang , Hongwei Wang

Rigid structure-from-motion (RSfM) and non-rigid structure-from-motion (NRSfM) have long been treated in the literature as separate (different) problems. Inspired by a previous work which solved directly for 3D scene structure by factoring…

Computer Vision and Pattern Recognition · Computer Science 2017-07-18 Pan Ji , Hongdong Li , Yuchao Dai , Ian Reid

We introduce a unified single and multi-view neural implicit 3D reconstruction framework VPFusion. VPFusion attains high-quality reconstruction using both - 3D feature volume to capture 3D-structure-aware context, and pixel-aligned image…

Computer Vision and Pattern Recognition · Computer Science 2022-07-19 Jisan Mahmud , Jan-Michael Frahm

We propose a Pose-Free Large Reconstruction Model (PF-LRM) for reconstructing a 3D object from a few unposed images even with little visual overlap, while simultaneously estimating the relative camera poses in ~1.3 seconds on a single A100…

Computer Vision and Pattern Recognition · Computer Science 2023-11-27 Peng Wang , Hao Tan , Sai Bi , Yinghao Xu , Fujun Luan , Kalyan Sunkavalli , Wenping Wang , Zexiang Xu , Kai Zhang

3D scene reconstruction is fundamental for spatial intelligence applications such as AR, robotics, and digital twins. Traditional multi-view stereo struggles with sparse viewpoints or low-texture regions, while neural rendering approaches,…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Jiaqi Yao , Zhongmiao Yan , Jingyi Xu , Songpengcheng Xia , Yan Xiang , Ling Pei

3D human articulated pose recovery from monocular image sequences is very challenging due to the diverse appearances, viewpoints, occlusions, and also the human 3D pose is inherently ambiguous from the monocular imagery. It is thus critical…

Computer Vision and Pattern Recognition · Computer Science 2017-08-01 Mude Lin , Liang Lin , Xiaodan Liang , Keze Wang , Hui Cheng