English
Related papers

Related papers: FastViDAR: Real-Time Omnidirectional Depth Estimat…

200 papers

Multi-view 3D reconstruction remains a core challenge in computer vision, particularly in applications requiring accurate and scalable representations across diverse perspectives. Current leading methods such as DUSt3R employ a…

Computer Vision and Pattern Recognition · Computer Science 2025-03-21 Jianing Yang , Alexander Sax , Kevin J. Liang , Mikael Henaff , Hao Tang , Ang Cao , Joyce Chai , Franziska Meier , Matt Feiszli

Monocular depth estimation (MDE) plays a pivotal role in various computer vision applications, such as robotics, augmented reality, and autonomous driving. Despite recent advancements, existing methods often fail to meet key requirements…

Computer Vision and Pattern Recognition · Computer Science 2025-09-29 Andrii Litvynchuk , Ivan Livinsky , Anand Ravi , Nima Kalantari , Andrii Tsarov

In this paper, we propose a novel end-to-end deep neural network model for omnidirectional depth estimation from a wide-baseline multi-view stereo setup. The images captured with ultra wide field-of-view (FOV) cameras on an omnidirectional…

Computer Vision and Pattern Recognition · Computer Science 2019-08-20 Changhee Won , Jongbin Ryu , Jongwoo Lim

Omnidirectional multi-view stereo (MVS) vision is attractive for its ultra-wide field-of-view (FoV), enabling machines to perceive 360{\deg} 3D surroundings. However, the existing solutions require expensive dense depth labels for…

Computer Vision and Pattern Recognition · Computer Science 2023-02-23 Zisong Chen , Chunyu Lin , Lang Nie , Kang Liao , Yao Zhao

We design a multiscopic vision system that utilizes a low-cost monocular RGB camera to acquire accurate depth estimation. Unlike multi-view stereo with images captured at unconstrained camera poses, the proposed system controls the motion…

Computer Vision and Pattern Recognition · Computer Science 2021-08-21 Weihao Yuan , Rui Fan , Michael Yu Wang , Qifeng Chen

High dynamic range (HDR) novel view synthesis (NVS) aims to reconstruct HDR scenes from multi-exposure low dynamic range (LDR) images. Existing HDR pipelines heavily rely on known camera poses, well-initialized dense point clouds, and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Dingqiang Ye , Jiacong Xu , Jianglu Ping , Yuxiang Guo , Chao Fan , Vishal M. Patel

Despite recent progress in 3D Gaussian-based head avatar modeling, efficiently generating high fidelity avatars remains a challenge. Current methods typically rely on extensive multi-view capture setups or monocular videos with per-identity…

Computer Vision and Pattern Recognition · Computer Science 2026-02-02 Xinya Ji , Sebastian Weiss , Manuel Kansy , Jacek Naruniec , Xun Cao , Barbara Solenthaler , Derek Bradley

Depth from a monocular video can enable billions of devices and robots with a single camera to see the world in 3D. In this paper, we present an approach with a differentiable flow-to-depth layer for video depth estimation. The model…

Computer Vision and Pattern Recognition · Computer Science 2020-03-04 Jiaxin Xie , Chenyang Lei , Zhuwen Li , Li Erran Li , Qifeng Chen

In this paper, we present the first pinhole-fisheye framework for heterogeneous multi-view depth estimation, PFDepth. Our key insight is to exploit the complementary characteristics of pinhole and fisheye imagery (undistorted vs. distorted,…

Computer Vision and Pattern Recognition · Computer Science 2025-10-01 Zhiwei Zhang , Ruikai Xu , Weijian Zhang , Zhizhong Zhang , Xin Tan , Jingyu Gong , Yuan Xie , Lizhuang Ma

Accurately capturing dynamic scenes with wide-ranging motion and light intensity is crucial for many vision applications. However, acquiring high-speed high dynamic range (HDR) video is challenging because the camera's frame rate restricts…

Image and Video Processing · Electrical Eng. & Systems 2024-04-26 Caixin Wang , Jie Zhang , Matthew A. Wilson , Ralph Etienne-Cummings

Many cameras implement auto-focus functionality. However, they typically require the user to manually identify the location to be focused on. While such an approach works for temporally-sparse autofocusing functionality (e.g., photo…

Computer Vision and Pattern Recognition · Computer Science 2017-11-10 Wolfgang Fuhl , Thiago Santini , Enkelejda Kasneci

Image fusion seeks to integrate complementary information from multiple sources into a single, superior image. While traditional methods are fast, they lack adaptability and performance. Conversely, deep learning approaches achieve…

Computer Vision and Pattern Recognition · Computer Science 2026-02-25 Ran Zhang , Xuanhua He , Liu Liu

We propose DeepFusion, a modular multi-modal architecture to fuse lidars, cameras and radars in different combinations for 3D object detection. Specialized feature extractors take advantage of each modality and can be exchanged easily,…

Computer Vision and Pattern Recognition · Computer Science 2022-09-28 Florian Drews , Di Feng , Florian Faion , Lars Rosenbaum , Michael Ulrich , Claudius Gläser

Panorama has a full FoV (360$^\circ\times$180$^\circ$), offering a more complete visual description than perspective images. Thanks to this characteristic, panoramic depth estimation is gaining increasing traction in 3D vision. However, due…

Computer Vision and Pattern Recognition · Computer Science 2025-11-11 Haodong Li , Wangguangdong Zheng , Jing He , Yuhao Liu , Xin Lin , Xin Yang , Ying-Cong Chen , Chunchao Guo

We tackle active view selection in novel view synthesis and 3D reconstruction. Existing methods like FisheRF and ActiveNeRF select the next best view by minimizing uncertainty or maximizing information gain in 3D, but they require…

Computer Vision and Pattern Recognition · Computer Science 2025-06-25 Zirui Wang , Yash Bhalgat , Ruining Li , Victor Adrian Prisacariu

Monocular metric depth estimation (MMDE) is a core challenge in computer vision, playing a pivotal role in real-world applications that demand accurate spatial understanding. Although prior works have shown promising zero-shot performance…

Computer Vision and Pattern Recognition · Computer Science 2026-04-09 Girish Chandar Ganesan , Yuliang Guo , Liu Ren , Xiaoming Liu

Automated driving systems use multi-modal sensor suites to ensure the reliable, redundant and robust perception of the operating domain, for example camera and LiDAR. An accurate extrinsic calibration is required to fuse the camera and…

Computer Vision and Pattern Recognition · Computer Science 2023-06-26 Jack Borer , Jeremy Tschirner , Florian Ölsner , Stefan Milz

Deep learning has been used to demonstrate end-to-end neural network learning for autonomous vehicle control from raw sensory input. While LiDAR sensors provide reliably accurate information, existing end-to-end driving solutions are mainly…

Robotics · Computer Science 2021-05-21 Zhijian Liu , Alexander Amini , Sibo Zhu , Sertac Karaman , Song Han , Daniela Rus

Joint Alignment (JA) of images aims to align a collection of images into a unified coordinate frame, such that semantically-similar features appear at corresponding spatial locations. Most existing approaches often require long training…

Computer Vision and Pattern Recognition · Computer Science 2025-10-30 Omri Hirsch , Ron Shapira Weber , Shira Ifergane , Oren Freifeld

Depth cameras allow to set up reliable solutions for people monitoring and behavior understanding, especially when unstable or poor illumination conditions make unusable common RGB sensors. Therefore, we propose a complete framework for the…

Computer Vision and Pattern Recognition · Computer Science 2018-09-03 Guido Borghi , Matteo Fabbri , Roberto Vezzani , Simone Calderara , Rita Cucchiara