中文
相关论文

相关论文: BRIGHT-VO: Brightness-Guided Hybrid Transformer fo…

200 篇论文

We present a modified velocity-obstacle (VO) algorithm that uses probabilistic partial observations of the environment to compute velocities and navigate a robot to a target. Our system uses commodity visual sensors, including a mono-camera…

机器人学 · 计算机科学 2021-06-10 Jing Liang , Yi-Ling Qiao , Tianrui Guan , Dinesh Manocha

Monocular visual odometry (VO) has attracted extensive research attention by providing real-time vehicle motion from cost-effective camera images. However, state-of-the-art optimization-based monocular VO methods suffer from the scale…

计算机视觉与模式识别 · 计算机科学 2022-03-14 Sen Zhang , Jing Zhang , Dacheng Tao

Recent approaches to VO have significantly improved performance by using deep networks to predict optical flow between video frames. However, existing methods still suffer from noisy and inconsistent flow matching, making it difficult to…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Zhaoxing Zhang , Junda Cheng , Gangwei Xu , Xiaoxiang Wang , Can Zhang , Xin Yang

Odometry is of key importance for localization in the absence of a map. There is considerable work in the area of visual odometry (VO), and recent advances in deep learning have brought novel approaches to VO, which directly learn salient…

计算机视觉与模式识别 · 计算机科学 2020-03-06 Wei Wang , Muhamad Risqi U. Saputra , Peijun Zhao , Pedro Gusmao , Bo Yang , Changhao Chen , Andrew Markham , Niki Trigoni

It is typically challenging for visual or visual-inertial odometry systems to handle the problems of dynamic scenes and pure rotation. In this work, we design a novel visual-inertial odometry (VIO) system called RD-VIO to handle both of…

机器人学 · 计算机科学 2024-02-19 Jinyu Li , Xiaokun Pan , Gan Huang , Ziyang Zhang , Nan Wang , Hujun Bao , Guofeng Zhang

We propose a novel monocular visual odometry (VO) system called UnDeepVO in this paper. UnDeepVO is able to estimate the 6-DoF pose of a monocular camera and the depth of its view by using deep neural networks. There are two salient…

计算机视觉与模式识别 · 计算机科学 2018-02-22 Ruihao Li , Sen Wang , Zhiqiang Long , Dongbing Gu

In recent years, unsupervised deep learning approaches have received significant attention to estimate the depth and visual odometry (VO) from unlabelled monocular image sequences. However, their performance is limited in challenging…

计算机视觉与模式识别 · 计算机科学 2020-11-03 Yasin Almalioglu , Angel Santamaria-Navarro , Benjamin Morrell , Ali-akbar Agha-mohammadi

With the increasing application of robots, stable and efficient Visual Odometry (VO) algorithms are becoming more and more important. Based on the Fourier Mellin Transformation (FMT) algorithm, the extended Fourier Mellin Transformation…

机器人学 · 计算机科学 2023-07-20 Wenqing Jiang , Chengqian Li , Jinyue Cao , Sören Schwertfeger

Vision-based odometry has been widely adopted in autonomous driving owing to its low cost and lightweight setup; however, its performance often degrades in complex outdoor urban environments. To address these challenges, we propose…

机器人学 · 计算机科学 2025-09-29 Zhixin Zhang , Liang Zhao , Pawel Ladosz

This paper presents an self-supervised deep learning network for monocular visual inertial odometry (named DeepVIO). DeepVIO provides absolute trajectory estimation by directly merging 2D optical flow feature (OFF) and Inertial Measurement…

机器人学 · 计算机科学 2019-07-01 Liming Han , Yimin Lin , Guoguang Du , Shiguo Lian

This paper overviews different pose representations and metric functions in visual odometry (VO) networks. The performance of VO networks heavily relies on how their architecture encodes the information. The choice of pose representation…

计算机视觉与模式识别 · 计算机科学 2024-01-12 Olaya Álvarez-Tuñón , Yury Brodskiy , Erdal Kayacan

Successful visual navigation depends upon capturing images that contain sufficient useful information. In this letter, we explore a data-driven approach to account for environmental lighting changes, improving the quality of images for use…

机器人学 · 计算机科学 2022-07-12 Justin Tomasi , Brandon Wagstaff , Steven L. Waslander , Jonathan Kelly

We present MVTOP, a novel transformer-based method for multi-view rigid object pose estimation. Through an early fusion of the view-specific features, our method can resolve pose ambiguities that would be impossible to solve with a single…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Lukas Ranftl , Felix Brendel , Bertram Drost , Carsten Steger

Deep learning approaches for Visual-Inertial Odometry (VIO) have proven successful, but they rarely focus on incorporating robust fusion strategies for dealing with imperfect input sensory data. We propose a novel end-to-end selective…

计算机视觉与模式识别 · 计算机科学 2019-03-06 Changhao Chen , Stefano Rosa , Yishu Miao , Chris Xiaoxuan Lu , Wei Wu , Andrew Markham , Niki Trigoni

Autonomous navigation for legged robots in complex and dynamic environments relies on robust simultaneous localization and mapping (SLAM) systems to accurately map surroundings and localize the robot, ensuring safe and efficient operation.…

Unsupervised learning for monocular camera motion and 3D scene understanding has gained popularity over traditional methods, relying on epipolar geometry or non-linear optimization. Notably, deep learning can overcome many issues of…

计算机视觉与模式识别 · 计算机科学 2022-03-15 Claudio Cimarelli , Hriday Bavle , Jose Luis Sanchez-Lopez , Holger Voos

Vision Transformers (ViT) have advanced computer vision, yet their efficacy in complex tasks like driving remains less explored. This study enhances ViT by integrating human eye gaze, captured via eye-tracking, to increase prediction…

计算机视觉与模式识别 · 计算机科学 2025-01-13 Sharath Koorathota , Nikolas Papadopoulos , Jia Li Ma , Shruti Kumar , Xiaoxiao Sun , Arunesh Mittal , Patrick Adelman , Paul Sajda

Understanding model decisions is crucial in medical imaging, where interpretability directly impacts clinical trust and adoption. Vision Transformers (ViTs) have demonstrated state-of-the-art performance in diagnostic imaging; however,…

计算机视觉与模式识别 · 计算机科学 2025-10-15 Leili Barekatain , Ben Glocker

The problem of tracking self-motion as well as motion of objects in the scene using information from a camera is known as multi-body visual odometry and is a challenging task. This paper proposes a robust solution to achieve accurate…

机器人学 · 计算机科学 2020-07-29 Jun Zhang , Mina Henein , Robert Mahony , Viorela Ila

In this paper we propose a framework for integrating map-based relocalization into online direct visual odometry. To achieve map-based relocalization for direct methods, we integrate image features into Direct Sparse Odometry (DSO) and rely…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Mariia Gladkova , Rui Wang , Niclas Zeller , Daniel Cremers