English
Related papers

Related papers: FlowMap: High-Quality Camera Poses, Intrinsics, an…

200 papers

Gaussian Splatting has become a leading reconstruction technique, known for its high-quality novel view synthesis and detailed reconstruction. However, most existing methods require dense, calibrated views. Reconstructing from free sparse…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Yibin Zhao , Yihan Pan , Jun Nan , Liwei Chen , Jianjun Yi

Monocular camera calibration is a key precondition for numerous 3D vision applications. Despite considerable advancements, existing methods often hinge on specific assumptions and struggle to generalize across varied real-world scenarios,…

Computer Vision and Pattern Recognition · Computer Science 2024-05-27 Xiankang He , Guangkai Xu , Bo Zhang , Hao Chen , Ying Cui , Dongyan Guo

Video denoising aims at removing noise from videos to recover clean ones. Some existing works show that optical flow can help the denoising by exploiting the additional spatial-temporal clues from nearby frames. However, the flow estimation…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Jiezhang Cao , Qin Wang , Jingyun Liang , Yulun Zhang , Kai Zhang , Radu Timofte , Luc Van Gool

Three-dimensional reconstruction in scenes with extreme depth variations remains challenging due to inconsistent supervisory signals between near-field and far-field regions. Existing methods fail to simultaneously address inaccurate depth…

Computer Vision and Pattern Recognition · Computer Science 2025-11-14 Yu Deng , Baozhu Zhao , Junyan Su , Xiaohan Zhang , Qi Liu

We present GSLoc: a new visual localization method that performs dense camera alignment using 3D Gaussian Splatting as a map representation of the scene. GSLoc backpropagates pose gradients over the rendering pipeline to align the rendered…

Robotics · Computer Science 2024-10-10 Kazii Botashev , Vladislav Pyatov , Gonzalo Ferrer , Stamatios Lefkimmiatis

Diffusion probabilistic models (DPMs) are a key component in modern generative models. DPM-solvers have achieved reduced latency and enhanced quality significantly, but have posed challenges to find the exact inverse (i.e., finding the…

Computer Vision and Pattern Recognition · Computer Science 2023-12-01 Seongmin Hong , Kyeonghyun Lee , Suh Yoon Jeon , Hyewon Bae , Se Young Chun

Structure-from-motion (SfM) is a long-standing problem in the computer vision community, which aims to reconstruct the camera poses and 3D structure of a scene from a set of unconstrained 2D images. Classical frameworks solve this problem…

Computer Vision and Pattern Recognition · Computer Science 2023-12-08 Jianyuan Wang , Nikita Karaev , Christian Rupprecht , David Novotny

While neural 3D reconstruction has advanced substantially, its performance significantly degrades with sparse-view data, which limits its broader applicability, since SfM is often unreliable in sparse-view scenarios where feature matches…

Computer Vision and Pattern Recognition · Computer Science 2025-07-08 Zhiwen Fan , Wenyan Cong , Kairun Wen , Kevin Wang , Jian Zhang , Xinghao Ding , Danfei Xu , Boris Ivanovic , Marco Pavone , Georgios Pavlakos , Zhangyang Wang , Yue Wang

Image enhancement holds extensive applications in real-world scenarios due to complex environments and limitations of imaging devices. Conventional methods are often constrained by their tailored models, resulting in diminished robustness…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Yixuan Zhu , Wenliang Zhao , Ao Li , Yansong Tang , Jie Zhou , Jiwen Lu

We present a dense simultaneous localization and mapping (SLAM) method that uses 3D Gaussians as a scene representation. Our approach enables interactive-time reconstruction and photo-realistic rendering from real-world single-camera RGBD…

Computer Vision and Pattern Recognition · Computer Science 2024-03-25 Vladimir Yugay , Yue Li , Theo Gevers , Martin R. Oswald

We present a real-time tracking SLAM system that unifies efficient camera tracking with photorealistic feature-enriched mapping using 3D Gaussian Splatting (3DGS). Our main contribution is integrating dense feature rasterization into the…

Computer Vision and Pattern Recognition · Computer Science 2026-03-23 Christopher Thirgood , Oscar Mendez , Erin Ling , Jon Storey , Simon Hadfield

We present an algorithm to estimate depth in dynamic video scenes. We propose to learn and infer depth in videos from appearance, motion, occlusion boundaries, and geometric context of the scene. Using our method, depth can be estimated…

Computer Vision and Pattern Recognition · Computer Science 2015-10-27 S. Hussain Raza , Omar Javed , Aveek Das , Harpreet Sawhney , Hui Cheng , Irfan Essa

Achieving high-fidelity 3D reconstruction from monocular video remains challenging due to the inherent limitations of traditional methods like Structure-from-Motion (SfM) and monocular SLAM in accurately capturing scene details. While…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Yue Hu , Rong Liu , Meida Chen , Peter Beerel , Andrew Feng

Diffusion models (DMs) have demonstrated remarkable success in real-world image super-resolution (SR), yet their reliance on time-consuming multi-step sampling largely hinders their practical applications. While recent efforts have…

Computer Vision and Pattern Recognition · Computer Science 2026-05-13 Jiaqi Xu , Wenbo Li , Haoze Sun , Fan Li , Zhixin Wang , Long Peng , Jingjing Ren , Haoran Yang , Xiaowei Hu , Renjing Pei , Pheng-Ann Heng

Reconstructing a 3D scene from unordered images is pivotal in computer vision and robotics, with applications spanning crowd-sourced mapping and beyond. While global Structure-from-Motion (SfM) techniques are scalable and fast, they often…

Computer Vision and Pattern Recognition · Computer Science 2024-10-17 Linfei Pan , Marc Pollefeys , Dániel Baráth

Self-supervised monocular depth estimation methods have been increasingly given much attention due to the benefit of not requiring large, labelled datasets. Such self-supervised methods require high-quality salient features and consequently…

Computer Vision and Pattern Recognition · Computer Science 2024-03-28 Xiaotong Guo , Huijie Zhao , Shuwei Shao , Xudong Li , Baochang Zhang

Though there exists a reasonable forward model for blur based on optical physics, recovering depth from a collection of defocused images remains a computationally challenging optimization problem. In this paper, we show that with…

Computer Vision and Pattern Recognition · Computer Science 2026-02-27 Holly Jackson , Caleb Adams , Ignacio Lopez-Francos , Benjamin Recht

We propose MFT -- Multi-Flow dense Tracker -- a novel method for dense, pixel-level, long-term tracking. The approach exploits optical flows estimated not only between consecutive frames, but also for pairs of frames at logarithmically…

Computer Vision and Pattern Recognition · Computer Science 2023-11-13 Michal Neoral , Jonáš Šerých , Jiří Matas

Accurately and efficiently modeling dynamic scenes and motions is considered so challenging a task due to temporal dynamics and motion complexity. To address these challenges, we propose DynMF, a compact and efficient representation that…

Computer Vision and Pattern Recognition · Computer Science 2024-12-06 Agelos Kratimenos , Jiahui Lei , Kostas Daniilidis

We propose DrivingForward, a feed-forward Gaussian Splatting model that reconstructs driving scenes from flexible surround-view input. Driving scene images from vehicle-mounted cameras are typically sparse, with limited overlap, and the…

Computer Vision and Pattern Recognition · Computer Science 2024-12-24 Qijian Tian , Xin Tan , Yuan Xie , Lizhuang Ma