中文
相关论文

相关论文: A Novel Georeferenced Dataset for Stereo Visual Od…

200 篇论文

A 360{\deg} perception of scene geometry is essential for automated driving, notably for parking and urban driving scenarios. Typically, it is achieved using surround-view fisheye cameras, focusing on the near-field area around the vehicle.…

计算机视觉与模式识别 · 计算机科学 2021-04-12 Varun Ravi Kumar , Marvin Klingner , Senthil Yogamani , Markus Bach , Stefan Milz , Tim Fingscheidt , Patrick Mäder

Drones are increasingly used in fields like industry, medicine, research, disaster relief, defense, and security. Technical challenges, such as navigation in GPS-denied environments, hinder further adoption. Research in visual odometry is…

机器人学 · 计算机科学 2024-04-30 Olivier Brochu Dufour , Abolfazl Mohebbi , Sofiane Achiche

Stereo matching is an important problem in computer vision which has drawn tremendous research attention for decades. Recent years, data-driven methods with convolutional neural networks (CNNs) are continuously pushing stereo matching to…

计算机视觉与模式识别 · 计算机科学 2021-01-27 Ju He , Enyu Zhou , Liusheng Sun , Fei Lei , Chenyang Liu , Wenxiu Sun

Retrieving the missing dimension information in acoustic images from 2D forward-looking sonar is a well-known problem in the field of underwater robotics. There are works attempting to retrieve 3D information from a single image which…

计算机视觉与模式识别 · 计算机科学 2022-08-02 Yusheng Wang , Yonghoon Ji , Hiroshi Tsuchiya , Hajime Asama , Atsushi Yamashita

Stereo matching plays a crucial role in enabling depth perception for autonomous driving and robotics. While recent years have witnessed remarkable progress in stereo matching algorithms, largely driven by learning-based methods and…

计算机视觉与模式识别 · 计算机科学 2025-09-17 Xianda Guo , Chenming Zhang , Ruilin Wang , Youmin Zhang , Wenzhao Zheng , Matteo Poggi , Hao Zhao , Qin Zou , Long Chen

We present SetDiff, a geometry-grounded multi-view diffusion framework that enhances novel-view renderings produced by 3D Gaussian Splatting. Our method integrates explicit 3D priors, pixel-aligned coordinate maps and pose-aware Plucker ray…

计算机视觉与模式识别 · 计算机科学 2026-03-16 Farhad G. Zanjani , Hong Cai , Amirhossein Habibian

Understanding human instructions is essential for enabling smooth human-robot interaction. In this work, we focus on object grounding, i.e., localizing an object of interest in a visual scene (e.g., an image) based on verbal human…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Joel Alberto Santos , Zongwei Wu , Xavier Alameda-Pineda , Radu Timofte

This paper proposes a novel framework for real-time localization and egomotion tracking of a vehicle in a reference map. The core idea is to map the semantic objects observed by the vehicle and register them to their corresponding objects…

机器人学 · 计算机科学 2022-09-30 Jacqueline Ankenbauer , Kaveh Fathian , Jonathan P. How

Stereo depth estimation relies on optimal correspondence matching between pixels on epipolar lines in the left and right images to infer depth. In this work, we revisit the problem from a sequence-to-sequence correspondence perspective to…

计算机视觉与模式识别 · 计算机科学 2021-08-27 Zhaoshuo Li , Xingtong Liu , Nathan Drenkow , Andy Ding , Francis X. Creighton , Russell H. Taylor , Mathias Unberath

Self-supervised monocular depth estimation has become an appealing solution to the lack of ground truth labels, but its reconstruction loss often produces over-smoothed results across object boundaries and is incapable of handling occlusion…

计算机视觉与模式识别 · 计算机科学 2021-08-24 Hyesong Choi , Hunsang Lee , Sunkyung Kim , Sunok Kim , Seungryong Kim , Kwanghoon Sohn , Dongbo Min

We present a novel real-time visual odometry framework for a stereo setup of a depth and high-resolution event camera. Our framework balances accuracy and robustness against computational efficiency towards strong performance in challenging…

机器人学 · 计算机科学 2022-02-08 Yi-Fan Zuo , Jiaqi Yang , Jiaben Chen , Xia Wang , Yifu Wang , Laurent Kneip

Active stereo systems are used in many robotic applications that require 3D information. These depth sensors, however, suffer from stereo artefacts and do not provide dense depth estimates.In this work, we present the first self-supervised…

计算机视觉与模式识别 · 计算机科学 2022-01-21 Frederik Warburg , Daniel Hernandez-Juarez , Juan Tarrio , Alexander Vakhitov , Ujwal Bonde , Pablo F. Alcantarilla

Radar odometry is crucial for robust localization in challenging environments; however, the sparsity of reliable returns and distinctive noise characteristics impede its performance. This paper introduces geometrically-constrained…

机器人学 · 计算机科学 2026-04-06 Wooseong Yang , Dongjae Lee , Minwoo Jung , Ayoung Kim

Modern cameras are equipped with a wide array of sensors that enable recording the geospatial context of an image. Taking advantage of this, we explore depth estimation under the assumption that the camera is geocalibrated, a problem we…

计算机视觉与模式识别 · 计算机科学 2021-09-22 Scott Workman , Hunter Blanton

Robust perception is critical for autonomous driving, especially under adverse weather and lighting conditions that commonly occur in real-world environments. In this paper, we introduce the Stereo Image Dataset (SID), a large-scale…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Zaid A. El-Shair , Abdalmalek Abu-raddaha , Aaron Cofield , Hisham Alawneh , Mohamed Aladem , Yazan Hamzeh , Samir A. Rawashdeh

Robot manipulation of unknown objects in unstructured environments is a challenging problem due to the variety of shapes, materials, arrangements and lighting conditions. Even with large-scale real-world data collection, robust perception…

机器人学 · 计算机科学 2021-07-01 Thomas Kollar , Michael Laskey , Kevin Stone , Brijen Thananjeyan , Mark Tjersland

Visual localization is the task of estimating camera pose in a known scene, which is an essential problem in robotics and computer vision. However, long-term visual localization is still a challenge due to the environmental appearance…

机器人学 · 计算机科学 2022-12-02 Yuxuan Chen , Timothy D. Barfoot

Although the number of camera-based sensors mounted on vehicles has recently increased dramatically, robust and accurate object velocity detection is difficult. Additionally, it is still common to use radar as a fusion system. We have…

计算机视觉与模式识别 · 计算机科学 2020-12-02 Toru Saito , Toshimi Okubo , Naoki Takahashi

Despite learning-based visual odometry (VO) has shown impressive results in recent years, the pretrained networks may easily collapse in unseen environments. The large domain gap between training and testing data makes them difficult to…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Shunkai Li , Xin Wu , Yingdian Cao , Hongbin Zha

Robotic grasping is a cornerstone capability of embodied systems. Many methods directly output grasps from partial information without modeling the geometry of the scene, leading to suboptimal motion and even collisions. To address these…