English
Related papers

Related papers: RGB-Only Supervised Camera Parameter Optimization …

200 papers

Camera motion estimation is a key technique for 3D scene reconstruction and Simultaneous localization and mapping (SLAM). To make it be feasibly achieved, previous works usually assume slow camera motions, which limits its usage in many…

Computer Vision and Pattern Recognition · Computer Science 2018-12-10 Zunjie Zhu , Feng Xu

Purpose: Surgical scene understanding plays a critical role in the technology stack of tomorrow's intervention-assisting systems in endoscopic surgeries. For this, tracking the endoscope pose is a key component, but remains challenging due…

Computer Vision and Pattern Recognition · Computer Science 2023-04-18 Michel Hayoz , Christopher Hahne , Mathias Gallardo , Daniel Candinas , Thomas Kurmann , Maximilian Allan , Raphael Sznitman

Visual simultaneous localization and mapping (SLAM) plays a critical role in autonomous robotic systems, especially where accurate and reliable measurements are essential for navigation and sensing. In feature-based SLAM, the quantityand…

Robotics · Computer Science 2025-09-03 Haolan Zhang , Chenghao Li , Thanh Nguyen Canh , Lijun Wang , Nak Young Chong

Learning scene flow from a monocular camera still remains a challenging task due to its ill-posedness as well as lack of annotated data. Self-supervised methods demonstrate learning scene flow estimation from unlabeled data, yet their…

Computer Vision and Pattern Recognition · Computer Science 2022-05-04 Bayram Bayramli , Junhwa Hur , Hongtao Lu

We present ObjectMatch, a semantic and object-centric camera pose estimator for RGB-D SLAM pipelines. Modern camera pose estimators rely on direct correspondences of overlapping regions between frames; however, they cannot align camera…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Can Gümeli , Angela Dai , Matthias Nießner

We propose a new multi-instance dynamic RGB-D SLAM system using an object-level octree-based volumetric representation. It can provide robust camera tracking in dynamic environments and at the same time, continuously estimate geometric,…

Robotics · Computer Science 2019-03-25 Binbin Xu , Wenbin Li , Dimos Tzoumanikas , Michael Bloesch , Andrew Davison , Stefan Leutenegger

Radar is usually more robust than the camera in severe driving scenarios, e.g., weak/strong lighting and bad weather. However, unlike RGB images captured by a camera, the semantic information from the radar signals is noticeably difficult…

Computer Vision and Pattern Recognition · Computer Science 2021-02-11 Yizhou Wang , Zhongyu Jiang , Xiangyu Gao , Jenq-Neng Hwang , Guanbin Xing , Hui Liu

We present a novel non-rigid reconstruction method using a moving RGB-D camera. Current approaches use only non-rigid part of the scene and completely ignore the rigid background. Non-rigid parts often lack sufficient geometric and…

Computer Vision and Pattern Recognition · Computer Science 2018-05-31 Shafeeq Elanattil , Peyman Moghadam , Sridha Sridharan , Clinton Fookes , Mark Cox

We tackle the task of multi-view, multi-person 3D human pose estimation from a limited number of uncalibrated depth cameras. Recently, many approaches have been proposed for 3D human pose estimation from multi-view RGB cameras. However,…

Computer Vision and Pattern Recognition · Computer Science 2024-01-30 Yu-Jhe Li , Yan Xu , Rawal Khirodkar , Jinhyung Park , Kris Kitani

The frame rates of most 3D LIDAR sensors used in intelligent vehicles are substantially lower than current cameras installed in the same vehicle. This research suggests using a mono camera to virtually enhance the frame rate of LIDARs,…

Computer Vision and Pattern Recognition · Computer Science 2023-02-13 Zoltan Rozsa , Tamas Sziranyi

Neural Radiance Fields (NeRF) can be optimized to obtain high-fidelity 3D scene reconstructions of objects and large-scale scenes. However, NeRFs require accurate camera parameters as input -- inaccurate camera parameters result in blurry…

Computer Vision and Pattern Recognition · Computer Science 2023-09-01 Keunhong Park , Philipp Henzler , Ben Mildenhall , Jonathan T. Barron , Ricardo Martin-Brualla

Scene flow describes the motion of 3D objects in real world and potentially could be the basis of a good feature for 3D action recognition. However, its use for action recognition, especially in the context of convolutional neural networks…

Computer Vision and Pattern Recognition · Computer Science 2017-03-28 Pichao Wang , Wanqing Li , Zhimin Gao , Yuyao Zhang , Chang Tang , Philip Ogunbona

This work presents a novel dense RGB-D SLAM approach for dynamic planar environments that enables simultaneous multi-object tracking, camera localisation and background reconstruction. Previous dynamic SLAM methods either rely on semantic…

Robotics · Computer Science 2022-10-19 Ran Long , Christian Rauch , Tianwei Zhang , Vladimir Ivan , Tin Lun Lam , Sethu Vijayakumar

In this paper, a robust RGB-D SLAM system is proposed to utilize the structural information in indoor scenes, allowing for accurate tracking and efficient dense mapping on a CPU. Prior works have used the Manhattan World (MW) assumption to…

Computer Vision and Pattern Recognition · Computer Science 2021-03-30 Raza Yunus , Yanyan Li , Federico Tombari

To address the challenge of short-term object pose tracking in dynamic environments with monocular RGB input, we introduce a large-scale synthetic dataset OmniPose6D, crafted to mirror the diversity of real-world conditions. We additionally…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Yunzhi Lin , Yipu Zhao , Fu-Jen Chu , Xingyu Chen , Weiyao Wang , Hao Tang , Patricio A. Vela , Matt Feiszli , Kevin Liang

In dynamic environments, performance of visual SLAM techniques can be impaired by visual features taken from moving objects. One solution is to identify those objects so that their visual features can be removed for localization and…

Computer Vision and Pattern Recognition · Computer Science 2020-08-04 Jonathan Vincent , Mathieu Labbé , Jean-Samuel Lauzon , François Grondin , Pier-Marc Comtois-Rivet , François Michaud

With the popularity of monocular videos generated by video sharing and live broadcasting applications, reconstructing and editing dynamic scenes in stationary monocular cameras has become a special but anticipated technology. In contrast to…

Computer Vision and Pattern Recognition · Computer Science 2024-02-02 Weixing Xie , Xiao Dong , Yong Yang , Qiqin Lin , Jingze Chen , Junfeng Yao , Xiaohu Guo

This paper considers the final approach phase of visual-closed-loop grasping where the RGB-D camera is no longer able to provide valid depth information. Many current robotic grasping controllers are not closed-loop and therefore fail for…

Robotics · Computer Science 2020-03-02 Jesse Haviland , Feras Dayoub , Peter Corke

Motion segmentation from a single moving camera presents a significant challenge in the field of computer vision. This challenge is compounded by the unknown camera movements and the lack of depth information of the scene. While deep…

Computer Vision and Pattern Recognition · Computer Science 2024-06-28 Yuxiang Huang , Yuhao Chen , John Zelek

We present MonoGS++, a novel fast and accurate Simultaneous Localization and Mapping (SLAM) method that leverages 3D Gaussian representations and operates solely on RGB inputs. While previous 3D Gaussian Splatting (GS)-based methods largely…

Computer Vision and Pattern Recognition · Computer Science 2025-04-04 Renwu Li , Wenjing Ke , Dong Li , Lu Tian , Emad Barsoum
‹ Prev 1 4 5 6 7 8 10 Next ›