中文
相关论文

相关论文: UFO-4D: Unposed Feedforward 4D Reconstruction from…

200 篇论文

We propose a straightforward method that simultaneously reconstructs the 3D facial structure and provides dense alignment. To achieve this, we design a 2D representation called UV position map which records the 3D shape of a complete face…

计算机视觉与模式识别 · 计算机科学 2018-03-22 Yao Feng , Fan Wu , Xiaohu Shao , Yanfeng Wang , Xi Zhou

Reconstructing 3D scenes from sparse, unposed images remains challenging under real-world conditions with varying illumination and transient occlusions. Existing methods rely on scene-specific optimization using appearance embeddings or…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Vinayak Gupta , Chih-Hao Lin , Shenlong Wang , Anand Bhattad , Jia-Bin Huang

Feedforward 3D Gaussian Splatting (3DGS) overcomes the limitations of optimization-based 3DGS by enabling fast and high-quality reconstruction without the need for per-scene optimization. However, existing feedforward approaches typically…

计算机视觉与模式识别 · 计算机科学 2025-08-06 Anran Wu , Long Peng , Xin Di , Xueyuan Dai , Chen Wu , Yang Wang , Xueyang Fu , Yang Cao , Zheng-Jun Zha

Learning model-free object pose estimation for unseen instances remains a fundamental challenge in 3D vision. Existing methods typically fall into two disjoint paradigms: category-level approaches predict absolute poses in a canonical space…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Weihang Li , Lorenzo Garattoni , Fabien Despinoy , Nassir Navab , Benjamin Busam

As generative models become increasingly capable of producing high-fidelity visual content, the demand for efficient, interpretable, and editable image representations has grown substantially. Recent advances in 2D Gaussian Splatting (2DGS)…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Hao Wang , Ashish Bastola , Chaoyi Zhou , Wenhui Zhu , Xiwen Chen , Xuanzhao Dong , Siyu Huang , Abolfazl Razi

Recent trends in SLAM and visual navigation have embraced 3D Gaussians as the preferred scene representation, highlighting the importance of estimating camera poses from a single image using a pre-built Gaussian model. However, existing…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Hao Wang , Linqing Zhao , Xiuwei Xu , Jiwen Lu , Haibin Yan

Understanding and reconstructing the complex geometry and motion of dynamic scenes from video remains a formidable challenge in computer vision. This paper introduces D4RT, a simple yet powerful feedforward model designed to efficiently…

Free-Viewpoint Video (FVV) enables immersive 3D experiences, but efficient compression of dynamic 3D representation remains a major challenge. Existing dynamic 3D Gaussian Splatting methods couple reconstruction with optimization-dependent…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Wenkang Zhang , Yan Zhao , Qiang Wang , Zhixin Xu , Li Song , Zhengxue Cheng

LiDAR-camera fusion enhances 3D panoptic segmentation by leveraging camera images to complement sparse LiDAR scans, but it also introduces a critical failure mode. Under adverse conditions, degradation or failure of the camera sensor can…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Rohit Mohan , Florian Drews , Yakov Miron , Daniele Cattaneo , Abhinav Valada

In this paper, we propose UniGS, a unified map representation and differentiable framework for high-fidelity multimodal 3D reconstruction based on 3D Gaussian Splatting. Our framework integrates a CUDA-accelerated rasterization pipeline…

计算机视觉与模式识别 · 计算机科学 2025-11-14 Yusen Xie , Zhenmin Huang , Jianhao Jiao , Dimitrios Kanoulas , Jun Ma

Neural rendering has demonstrated remarkable success in high-quality 3D neural reconstruction and novel view synthesis with dense input views and accurate poses. However, applying it to extremely sparse, unposed views in unbounded 360{\deg}…

计算机视觉与模式识别 · 计算机科学 2025-04-01 Chong Bao , Xiyu Zhang , Zehao Yu , Jiale Shi , Guofeng Zhang , Songyou Peng , Zhaopeng Cui

As robotics technology advances, dense point cloud maps are increasingly in demand. However, dense reconstruction using a single unmanned aerial vehicle (UAV) suffers from limitations in flight speed and battery power, resulting in slow…

机器人学 · 计算机科学 2023-04-12 Yifei Dong , Shuhui Bu , Kun Li , Lin Chen , Zhenyu Xia , Yu Wang , Pengcheng Han , Xuefeng Cao , Ke Li

Dense image correspondence is central to many applications, such as visual odometry, 3D reconstruction, object association, and re-identification. Historically, dense correspondence has been tackled separately for wide-baseline scenarios…

计算机视觉与模式识别 · 计算机科学 2026-02-11 Yuchen Zhang , Nikhil Keetha , Chenwei Lyu , Bhuvan Jhamb , Yutian Chen , Yuheng Qiu , Jay Karhade , Shreyas Jha , Yaoyu Hu , Deva Ramanan , Sebastian Scherer , Wenshan Wang

Feed-forward 3D reconstruction from sparse, low-resolution (LR) images is a crucial capability for real-world applications, such as autonomous driving and embodied AI. However, existing methods often fail to recover fine texture details.…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Xinyuan Hu , Changyue Shi , Chuxiao Yang , Minghao Chen , Jiajun Ding , Tao Wei , Chen Wei , Zhou Yu , Min Tan

Recently, Gaussian splatting has received more and more attention in the field of static scene rendering. Due to the low computational overhead and inherent flexibility of explicit representations, plane-based explicit methods are popular…

计算机视觉与模式识别 · 计算机科学 2024-12-30 Jiawei Xu , Zexin Fan , Jian Yang , Jin Xie

Feed-forward 3D Gaussian splatting (3DGS) models have gained significant popularity due to their ability to generate scenes immediately without needing per-scene optimization. Although omnidirectional images are becoming more popular since…

计算机视觉与模式识别 · 计算机科学 2025-03-28 Suyoung Lee , Jaeyoung Chung , Kihoon Kim , Jaeyoo Huh , Gunhee Lee , Minsoo Lee , Kyoung Mu Lee

Understanding dynamic 3D environments is essential for safe autonomous driving, particularly when reasoning about human-centric, nonrigid agents. However, existing weakly supervised occupancy prediction frameworks predominantly assume…

计算机视觉与模式识别 · 计算机科学 2026-05-28 Yang Gao , Wuyang Li , Po-Chien Luan , Alexandre Alahi

In this paper we present an autonomous system for acquiring close-range high-resolution images that maximize the quality of a later-on 3D reconstruction with respect to coverage, ground resolution and 3D uncertainty. In contrast to previous…

计算机视觉与模式识别 · 计算机科学 2016-05-09 Christian Mostegel , Markus Rumpler , Friedrich Fraundorfer , Horst Bischof

We propose WildSplatter, a feed-forward 3D Gaussian Splatting (3DGS) model for unconstrained images with unknown camera parameters and varying lighting conditions. 3DGS is an effective scene representation that enables high-quality,…

计算机视觉与模式识别 · 计算机科学 2026-04-24 Yuki Fujimura , Takahiro Kushida , Kazuya Kitano , Takuya Funatomi , Yasuhiro Mukaigawa

We present FLARE, a feed-forward model designed to infer high-quality camera poses and 3D geometry from uncalibrated sparse-view images (i.e., as few as 2-8 inputs), which is a challenging yet practical setting in real-world applications.…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Shangzhan Zhang , Jianyuan Wang , Yinghao Xu , Nan Xue , Christian Rupprecht , Xiaowei Zhou , Yujun Shen , Gordon Wetzstein