English
Related papers

Related papers: Continuous Pose for Monocular Cameras in Neural Im…

200 papers

Dense 3D reconstruction from RGB images traditionally assumes static camera pose estimates. This assumption has endured, even as recent works have increasingly focused on real-time methods for mobile devices. However, the assumption of a…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Noah Stier , Baptiste Angles , Liang Yang , Yajie Yan , Alex Colburn , Ming Chuang

This paper proposes a self-supervised monocular image-to-depth prediction framework that is trained with an end-to-end photometric loss that handles not only 6-DOF camera motion but also 6-DOF moving object instances. Self-supervision is…

Computer Vision and Pattern Recognition · Computer Science 2022-08-10 Houssem Boulahbal , Adrian Voicila , Andrew Comport

We design a multiscopic vision system that utilizes a low-cost monocular RGB camera to acquire accurate depth estimation for robotic applications. Unlike multi-view stereo with images captured at unconstrained camera poses, the proposed…

Computer Vision and Pattern Recognition · Computer Science 2020-01-24 Weihao Yuan , Rui Fan , Michael Yu Wang , Qifeng Chen

In many robotic applications, the environment setting in which the 6-DoF pose estimation of a known, rigid object and its subsequent grasping is to be performed, remains nearly unchanging and might even be known to the robot in advance. In…

Computer Vision and Pattern Recognition · Computer Science 2022-07-28 Rohan Pratap Singh , Iori Kumagai , Antonio Gabas , Mehdi Benallegue , Yusuke Yoshiyasu , Fumio Kanehiro

Recent years have witnessed the remarkable success of implicit neural representation methods. The recent work Local Implicit Image Function (LIIF) has achieved satisfactory performance for continuous image representation, where pixel values…

Computer Vision and Pattern Recognition · Computer Science 2023-09-26 Zongyao He , Zhi Jin

The accuracy of monocular 3D human pose estimation depends on the viewpoint from which the image is captured. While freely moving cameras, such as on drones, provide control over this viewpoint, automatically positioning them at the…

Computer Vision and Pattern Recognition · Computer Science 2020-06-19 Sena Kiciroglu , Helge Rhodin , Sudipta N. Sinha , Mathieu Salzmann , Pascal Fua

Collaborative mapping of unknown environments can be done faster and more robustly than a single robot. However, a collaborative approach requires a distributed paradigm to be scalable and deal with communication issues. This work presents…

Robotics · Computer Science 2025-08-08 Mahboubeh Asadi , Kourosh Zareinia , Sajad Saeedi

Neural Radiance Fields (NeRFs) are trained using a set of camera poses and associated images as input to estimate density and color values for each position. The position-dependent density learning is of particular interest for…

Computer Vision and Pattern Recognition · Computer Science 2023-04-24 Miriam Jäger , Patrick Hübner , Dennis Haitz , Boris Jutzi

With the popularity of monocular videos generated by video sharing and live broadcasting applications, reconstructing and editing dynamic scenes in stationary monocular cameras has become a special but anticipated technology. In contrast to…

Computer Vision and Pattern Recognition · Computer Science 2024-02-02 Weixing Xie , Xiao Dong , Yong Yang , Qiqin Lin , Jingze Chen , Junfeng Yao , Xiaohu Guo

Neural Radiance Fields (NeRF) can be optimized to obtain high-fidelity 3D scene reconstructions of objects and large-scale scenes. However, NeRFs require accurate camera parameters as input -- inaccurate camera parameters result in blurry…

Computer Vision and Pattern Recognition · Computer Science 2023-09-01 Keunhong Park , Philipp Henzler , Ben Mildenhall , Jonathan T. Barron , Ricardo Martin-Brualla

In this paper, a learning-based approach is proposed for optimizing downlink beamforming in multiple-input multiple-output (MIMO) systems that employ continuous aperture arrays (CAPAs) at both the base station (BS) and the user. Beamforming…

Signal Processing · Electrical Eng. & Systems 2026-03-18 Shiyong Chen , Jia Guo , Shengqian Han

This paper presents a novel tightly-coupled monocular visual-inertial Simultaneous Localization and Mapping algorithm, which provides accurate and robust localization within the globally consistent map in real time on a standard CPU. This…

Robotics · Computer Science 2021-02-24 Meixiang Quan , Songhao Piao , Minglang Tan , Shi-Sheng Huang

Unsupervised learning for monocular camera motion and 3D scene understanding has gained popularity over traditional methods, relying on epipolar geometry or non-linear optimization. Notably, deep learning can overcome many issues of…

Computer Vision and Pattern Recognition · Computer Science 2022-03-15 Claudio Cimarelli , Hriday Bavle , Jose Luis Sanchez-Lopez , Holger Voos

Neural Radiance Fields (NeRF) have demonstrated impressive performance in novel view synthesis. However, NeRF and most of its variants still rely on traditional complex pipelines to provide extrinsic and intrinsic camera parameters, such as…

Computer Vision and Pattern Recognition · Computer Science 2023-12-15 Qingsong Yan , Qiang Wang , Kaiyong Zhao , Jie Chen , Bo Li , Xiaowen Chu , Fei Deng

Object localization, and more specifically object pose estimation, in large industrial spaces such as warehouses and production facilities, is essential for material flow operations. Traditional approaches rely on artificial artifacts…

Computer Vision and Pattern Recognition · Computer Science 2023-10-24 Hazem Youssef , Frederik Polachowski , Jérôme Rutinowski , Moritz Roidl , Christopher Reining

Dynamic scene reconstruction for autonomous driving enables vehicles to perceive and interpret complex scene changes more precisely. Dynamic Neural Radiance Fields (NeRFs) have recently shown promising capability in scene modeling. However,…

Computer Vision and Pattern Recognition · Computer Science 2025-05-15 Yue Wen , Liang Song , Yijia Liu , Siting Zhu , Yanzi Miao , Lijun Han , Hesheng Wang

In this paper, we tackle the problem of estimating the depth of a scene from a monocular video sequence. In particular, we handle challenging scenarios, such as non-translational camera motion and dynamic scenes, where traditional structure…

Computer Vision and Pattern Recognition · Computer Science 2015-11-20 Miaomiao Liu , Mathieu Salzmann , Xuming He

Multi-camera dynamic Augmented Reality (AR) applications require a camera pose estimation to leverage individual information from each camera in one common system. This can be achieved by combining contextual information, such as markers or…

Computer Vision and Pattern Recognition · Computer Science 2026-03-25 Shiyu Li , Hannah Schieber , Kristoffer Waldow , Benjamin Busam , Julian Kreimeier , Daniel Roth

Fiducial markers can encode rich information about the environment and can aid Visual SLAM (VSLAM) approaches in reconstructing maps with practical semantic information. Current marker-based VSLAM approaches mainly utilize markers for…

Robotics · Computer Science 2023-12-27 Ali Tourani , Hriday Bavle , Jose Luis Sanchez-Lopez , Rafael Munoz Salinas , Holger Voos

Unsupervised Learning based monocular visual odometry (VO) has lately drawn significant attention for its potential in label-free leaning ability and robustness to camera parameters and environmental variations. However, partially due to…

Computer Vision and Pattern Recognition · Computer Science 2019-03-18 Yang Li , Yoshitaka Ushiku , Tatsuya Harada