English
Related papers

Related papers: Semantic Nearest Neighbor Fields Monocular Edge Vi…

200 papers

Detecting and segmenting moving objects from a moving monocular camera is challenging in the presence of unknown camera motion, diverse object motions and complex scene structures. Most existing methods rely on a single motion cue to…

Computer Vision and Pattern Recognition · Computer Science 2024-05-06 Yuxiang Huang , Yuhao Chen , John Zelek

Accurate perception of objects in the environment is important for improving the scene understanding capability of SLAM systems. In robotic and augmented reality applications, object maps with semantic and metric information show attractive…

Robotics · Computer Science 2023-11-21 Xiao Han , Houxuan Liu , Yunchao Ding , Lu Yang

We introduce ZeroVO, a novel visual odometry (VO) algorithm that achieves zero-shot generalization across diverse cameras and environments, overcoming limitations in existing methods that depend on predefined or static camera calibration…

Computer Vision and Pattern Recognition · Computer Science 2025-06-10 Lei Lai , Zekai Yin , Eshed Ohn-Bar

We propose a learning-based method that solves monocular stereo and can be extended to fuse depth information from multiple target frames. Given two unconstrained images from a monocular camera with known intrinsic calibration, our network…

Computer Vision and Pattern Recognition · Computer Science 2019-09-13 Kaixuan Wang , Shaojie Shen

Bird's-Eye-View (BEV) maps have emerged as one of the most powerful representations for scene understanding due to their ability to provide rich spatial context while being easy to interpret and process. Such maps have found use in many…

Computer Vision and Pattern Recognition · Computer Science 2022-03-01 Nikhil Gosala , Abhinav Valada

Monocular simultaneous localization and mapping (SLAM) is emerging in advanced driver assistance systems and autonomous driving, because a single camera is cheap and easy to install. Conventional monocular SLAM has two major challenges…

Computer Vision and Pattern Recognition · Computer Science 2022-12-16 Jinkyu Lee , Muhyun Back , Sung Soo Hwang , Il Yong Chun

Curriculum Learning (CL), drawing inspiration from natural learning patterns observed in humans and animals, employs a systematic approach of gradually introducing increasingly complex training data during model development. Our work…

Robotics · Computer Science 2024-12-16 Assaf Lahiany , Oren Gal

Augmented reality (AR) displays become more and more popular recently, because of its high intuitiveness for humans and high-quality head-mounted display have rapidly developed. To achieve such displays with augmented information, highly…

Computer Vision and Pattern Recognition · Computer Science 2015-06-22 Kuan-Wen Chen , Chun-Hsin Wang , Xiao Wei , Qiao Liang , Ming-Hsuan Yang , Chu-Song Chen , Yi-Ping Hung

Transparent object perception is indispensable for numerous robotic tasks. However, accurately segmenting and estimating the depth of transparent objects remain challenging due to complex optical properties. Existing methods primarily delve…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Jiangyuan Liu , Hongxuan Ma , Yuxin Guo , Yuhao Zhao , Chi Zhang , Wei Sui , Wei Zou

Combining cameras and inertial measurement units (IMUs) has been proven effective in motion tracking, as these two sensing modalities offer complementary characteristics that are suitable for fusion. While most works focus on global-shutter…

Computer Vision and Pattern Recognition · Computer Science 2018-10-15 Yonggen Ling , Linchao Bao , Zequn Jie , Fengming Zhu , Ziyang Li , Shanmin Tang , Yongsheng Liu , Wei Liu , Tong Zhang

This paper presents an end-to-end multi-modal learning approach for monocular Visual-Inertial Odometry (VIO), which is specifically designed to exploit sensor complementarity in the light of sensor degradation scenarios. The proposed…

Computer Vision and Pattern Recognition · Computer Science 2020-07-16 Kashmira Shinde , Jongseok Lee , Matthias Humt , Aydin Sezgin , Rudolph Triebel

We introduce a novel task of 3D visual grounding in monocular RGB images using language descriptions with both appearance and geometry information. Specifically, we build a large-scale dataset, Mono3DRefer, which contains 3D object targets…

Computer Vision and Pattern Recognition · Computer Science 2023-12-14 Yang Zhan , Yuan Yuan , Zhitong Xiong

Semantic segmentation in videos has been a focal point of recent research. However, existing models encounter challenges when faced with unfamiliar categories. To address this, we introduce the Open Vocabulary Video Semantic Segmentation…

Multimedia · Computer Science 2024-12-13 Xinhao Li , Yun Liu , Guolei Sun , Min Wu , Le Zhang , Ce Zhu

In this paper, we propose a monocular visual localization pipeline leveraging semantic and depth cues. We apply semantic consistency evaluation to rank the image retrieval results and a practical clustering technique to reject estimation…

Computer Vision and Pattern Recognition · Computer Science 2020-05-26 Huanhuan Fan , Yuhao Zhou , Ang Li , Shuang Gao , Jijunnan Li , Yandong Guo

In the field of Simultaneous Localization and Mapping (SLAM), researchers have always pursued better performance in terms of accuracy and time cost. Traditional algorithms typically rely on fundamental geometric elements in images to…

Robotics · Computer Science 2024-03-05 Zhang Zhihe

We present COMO, a real-time monocular mapping and odometry system that encodes dense geometry via a compact set of 3D anchor points. Decoding anchor point projections into dense geometry via per-keyframe depth covariance functions…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Eric Dexheimer , Andrew J. Davison

Vision-based localization in a prior map is of crucial importance for autonomous vehicles. Given a query image, the goal is to estimate the camera pose corresponding to the prior map, and the key is the registration problem of camera images…

Computer Vision and Pattern Recognition · Computer Science 2022-10-11 Xingyu Chen , Jianru Xue , Shanmin Pang

A reliable sense-and-avoid system is critical to enabling safe autonomous operation of unmanned aircraft. Existing sense-and-avoid methods often require specialized sensors that are too large or power intensive for use on small unmanned…

Computer Vision and Pattern Recognition · Computer Science 2021-11-04 John Mern , Kyle Julian , Rachael E. Tompa , Mykel J. Kochenderfer

Monocular visual navigation methods have seen significant advances in the last decade, recently producing several real-time solutions for autonomously navigating small unmanned aircraft systems without relying on GPS. This is critical for…

Computer Vision and Pattern Recognition · Computer Science 2020-09-24 Kyung Kim , Robert C. Leishman , Scott L. Nykl

In this paper we introduce MLINE-VINS, a novel monocular visual-inertial odometry (VIO) system that leverages line features and Manhattan Word assumption. Specifically, for line matching process, we propose a novel geometric line optical…

Robotics · Computer Science 2025-03-04 Chao Ye , Haoyuan Li , Weiyang Lin , Xianqiang Yang
‹ Prev 1 8 9 10 Next ›