中文
相关论文

相关论文: Learning-based Image Enhancement for Visual Odomet…

200 篇论文

In the field of Simultaneous Localization and Mapping (SLAM), researchers have always pursued better performance in terms of accuracy and time cost. Traditional algorithms typically rely on fundamental geometric elements in images to…

机器人学 · 计算机科学 2024-03-05 Zhang Zhihe

Robust feature matching forms the backbone for most Visual Simultaneous Localization and Mapping (vSLAM), visual odometry, 3D reconstruction, and Structure from Motion (SfM) algorithms. However, recovering feature matches from texture-poor…

计算机视觉与模式识别 · 计算机科学 2023-08-03 Shenbagaraj Kannapiran , Nalin Bendapudi , Ming-Yuan Yu , Devarth Parikh , Spring Berman , Ankit Vora , Gaurav Pandey

Accurate and robust localization is a fundamental need for mobile agents. Visual-inertial odometry (VIO) algorithms exploit the information from camera and inertial sensors to estimate position and translation. Recent deep learning based…

计算机视觉与模式识别 · 计算机科学 2022-09-20 Zheming Tu , Changhao Chen , Xianfei Pan , Ruochen Liu , Jiarui Cui , Jun Mao

We introduce a novel monocular visual odometry (VO) system, NeRF-VO, that integrates learning-based sparse visual odometry for low-latency camera tracking and a neural radiance scene representation for fine-detailed dense reconstruction and…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Jens Naumann , Binbin Xu , Stefan Leutenegger , Xingxing Zuo

Visual recognition under adverse conditions is a very important and challenging problem of high practical value, due to the ubiquitous existence of quality distortions during image acquisition, transmission, or storage. While deep neural…

计算机视觉与模式识别 · 计算机科学 2019-04-04 Ding Liu , Bowen Cheng , Zhangyang Wang , Haichao Zhang , Thomas S. Huang

Interventional magnetic resonance imaging (i-MRI) for surgical guidance could help visualize the interventional process such as deep brain stimulation (DBS), improving the surgery performance and patient outcome. Different from…

计算机视觉与模式识别 · 计算机科学 2022-04-13 Ruiyang Zhao , Zhao He , Tao Wang , Suhao Qiu , Pawel Herman , Yanle Hu , Chencheng Zhang , Dinggang Shen , Bomin Sun , Guang-Zhong Yang , Yuan Feng

High dynamic range (HDR) imaging provides the capability of handling real world lighting as opposed to the traditional low dynamic range (LDR) which struggles to accurately represent images with higher dynamic range. However, most imaging…

计算机视觉与模式识别 · 计算机科学 2019-09-05 Demetris Marnerides , Thomas Bashford-Rogers , Jonathan Hatchett , Kurt Debattista

Visual Inertial Odometry (VIO) is one of the most established state estimation methods for mobile platforms. However, when visual tracking fails, VIO algorithms quickly diverge due to rapid error accumulation during inertial data…

机器人学 · 计算机科学 2023-06-13 Russell Buchanan , Varun Agrawal , Marco Camurri , Frank Dellaert , Maurice Fallon

Detection of moving objects is an essential capability in dealing with dynamic environments. Most moving object detection algorithms have been designed for color images without depth. For robotic navigation where real-time RGB-D data is…

计算机视觉与模式识别 · 计算机科学 2020-09-21 Haram Kim , Pyojin Kim , H. Jin Kim

Vision-based path following allows robots to autonomously repeat manually taught paths. Stereo Visual Teach and Repeat (VT\&R) accomplishes accurate and robust long-range path following in unstructured outdoor environments across changing…

机器人学 · 计算机科学 2020-03-09 Mona Gridseth , Timothy D. Barfoot

Deep operator networks (DeepONets, DONs) offer a distinct advantage over traditional neural networks in their ability to be trained on multi-resolution data. This property becomes especially relevant in real-world scenarios where…

机器学习 · 计算机科学 2023-10-05 Katarzyna Michałowska , Somdatta Goswami , George Em Karniadakis , Signe Riemer-Sørensen

Effectively localizing an agent in a realistic, noisy setting is crucial for many embodied vision tasks. Visual Odometry (VO) is a practical substitute for unreliable GPS and compass sensors, especially in indoor environments. While…

计算机视觉与模式识别 · 计算机科学 2023-05-02 Marius Memmel , Roman Bachmann , Amir Zamir

In the recent past, algorithms based on Convolutional Neural Networks (CNNs) have achieved significant milestones in object recognition. With large examples of each object class, standard datasets train well for inter-class variability.…

计算机视觉与模式识别 · 计算机科学 2018-06-11 Shrinivasan Sankar , Adrien Bartoli

Vision-and-Language Navigation (VLN) has long been constrained by the limited diversity and scalability of simulator-curated datasets, which fail to capture the complexity of real-world environments. To overcome this limitation, we…

计算机视觉与模式识别 · 计算机科学 2026-03-11 Mingfei Han , Haihong Hao , Liang Ma , Kamila Zhumakhanova , Ekaterina Radionova , Jingyi Zhang , Xiaojun Chang , Xiaodan Liang , Ivan Laptev

With the success of deep learning based approaches in tackling challenging problems in computer vision, a wide range of deep architectures have recently been proposed for the task of visual odometry (VO) estimation. Most of these proposed…

机器人学 · 计算机科学 2018-04-16 Ganesh Iyer , J. Krishna Murthy , Gunshi Gupta , K. Madhava Krishna , Liam Paull

This thesis report studies methods to solve Visual Question-Answering (VQA) tasks with a Deep Learning framework. As a preliminary step, we explore Long Short-Term Memory (LSTM) networks used in Natural Language Processing (NLP) to tackle…

计算与语言 · 计算机科学 2016-10-11 Issey Masuda , Santiago Pascual de la Puente , Xavier Giro-i-Nieto

This work proposes a novel deep network architecture to solve the camera Ego-Motion estimation problem. A motion estimation network generally learns features similar to Optical Flow (OF) fields starting from sequences of images. This OF can…

计算机视觉与模式识别 · 计算机科学 2018-02-16 Gabriele Costante , Thomas A. Ciarfuglia

Deep neural networks (DNNs) are reshaping the field of information processing. With their exponential growth challenging existing electronic hardware, optical neural networks (ONNs) are emerging to process DNN tasks in the optical domain…

Classical monocular vSLAM/VO methods suffer from the scale ambiguity problem. Hybrid approaches solve this problem by adding deep learning methods, for example by using depth maps which are predicted by a CNN. We suggest that it is better…

计算机视觉与模式识别 · 计算机科学 2019-04-18 Robin Kreuzig , Matthias Ochs , Rudolf Mester

This project aims to develop a robust video surveillance system, which can segment videos into smaller clips based on the detection of activities. It uses CCTV footage, for example, to record only major events-like the appearance of a…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Shahran Rahman Alve