中文
相关论文

相关论文: WildCross: A Cross-Modal Large Scale Benchmark for…

200 篇论文

Spatial visual perception is a fundamental requirement in physical-world applications like autonomous driving and robotic manipulation, driven by the need to interact with 3D environments. Capturing pixel-aligned metric depth using RGB-D…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Bin Tan , Changjiang Sun , Xiage Qin , Hanat Adai , Zelin Fu , Tianxiang Zhou , Han Zhang , Yinghao Xu , Xing Zhu , Yujun Shen , Nan Xue

Road detection or traversability analysis has been a key technique for a mobile robot to traverse complex off-road scenes. The problem has been mainly formulated in early works as a binary classification one, e.g. associating pixels with…

计算机视觉与模式识别 · 计算机科学 2021-03-08 Biao Gao , Shaochi Hu , Xijun Zhao , Huijing Zhao

We present TartanGround, a large-scale, multi-modal dataset to advance the perception and autonomy of ground robots operating in diverse environments. This dataset, collected in various photorealistic simulation environments includes…

机器人学 · 计算机科学 2025-07-31 Manthan Patel , Fan Yang , Yuheng Qiu , Cesar Cadena , Sebastian Scherer , Marco Hutter , Wenshan Wang

Tactile perception is essential for human interaction with the environment and is becoming increasingly crucial in robotics. Tactile sensors like the BioTac mimic human fingertips and provide detailed interaction data. Despite its utility…

机器人学 · 计算机科学 2026-04-29 Wadhah Zai El Amri , Malte Kuhlmann , Nicolás Navarro-Guerrero

Reliable localization is crucial for navigation in forests, where GPS is often degraded and LiDAR measurements are repetitive, occluded, and structurally complex. These conditions weaken the assumptions of traditional urban-centric…

机器人学 · 计算机科学 2026-02-13 Minwoo Jung , Nived Chebrolu , Lucas Carvalho de Lima , Haedam Oh , Maurice Fallon , Ayoung Kim

While a great variety of 3D cameras have been introduced in recent years, most publicly available datasets for object recognition and pose estimation focus on one single camera. In this work, we present a dataset of 32 scenes that have been…

机器人学 · 计算机科学 2020-09-30 Till Grenzdörffer , Martin Günther , Joachim Hertzberg

People detection methods are highly sensitive to the perpetual occlusions among the targets. As multi-camera set-ups become more frequently encountered, joint exploitation of the across views information would allow for improved detection…

计算机视觉与模式识别 · 计算机科学 2017-07-31 Tatjana Chavdarova , Pierre Baqué , Stéphane Bouquet , Andrii Maksai , Cijo Jose , Louis Lettry , Pascal Fua , Luc Van Gool , François Fleuret

Safe highway autonomy for heavy trucks remains an open and unsolved challenge: due to long braking distances, scene understanding of hundreds of meters is required for anticipatory planning and to allow safe braking margins. However,…

计算机视觉与模式识别 · 计算机科学 2026-04-01 Filippo Ghilotti , Edoardo Palladin , Samuel Brucker , Adam Sigal , Mario Bijelic , Felix Heide

Robotic and animal mapping systems share many of the same objectives and challenges, but differ in one key aspect: where much of the research in robotic mapping has focused on solving the data association problem, the grid cell neurons…

计算机视觉与模式识别 · 计算机科学 2018-10-24 Huu Le , Michael Milford

Navigating robots safely and efficiently in crowded and complex environments remains a significant challenge. However, due to the dynamic and intricate nature of these settings, planning efficient and collision-free paths for robots to…

机器人学 · 计算机科学 2024-10-22 Zhuanglei Wen , Mingze Dong , Xiai Chen

In indoor environments, multi-robot visual (RGB-D) mapping and exploration hold immense potential for application in domains such as domestic service and logistics, where deploying multiple robots in the same environment can significantly…

机器人学 · 计算机科学 2024-11-06 Sai Krishna Ghanta , Ramviyas Parasuraman

Autonomous driving datasets are essential for validating the progress of intelligent vehicle algorithms, which include localization, perception, and prediction. However, existing datasets are predominantly focused on structured urban…

计算机视觉与模式识别 · 计算机科学 2025-05-29 Chenfeng Wei , Qi Wu , Si Zuo , Jiahua Xu , Boyang Zhao , Zeyu Yang , Guotao Xie , Shenhong Wang

Recently, significant progress has been made in single-view depth estimation thanks to increasingly large and diverse depth datasets. However, these datasets are largely limited to specific application domains (e.g. indoor, autonomous…

计算机视觉与模式识别 · 计算机科学 2021-04-01 Yifan Wang , Linjie Luo , Xiaohui Shen , Xing Mei

Robots and autonomous systems need to know where they are within a map to navigate effectively. Thus, simultaneous localization and mapping or SLAM is a common building block of robot navigation systems. When building a map via a SLAM…

机器人学 · 计算机科学 2021-03-18 Luca Di Giammarino , Irvin Aloise , Cyrill Stachniss , Giorgio Grisetti

Reliable depth estimation under real optical conditions remains a core challenge for camera vision in systems such as autonomous robotics and augmented reality. Despite recent progress in depth estimation and depth-of-field rendering,…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Nisarg K. Trivedi , Vinayak A. Belludi , Li-Yun Wang

Deploying robots at scale demands robustness to the long tail of everyday situations. The countless variations in scene layout, object geometry, and task specifications that characterize real environments are vast and underrepresented in…

Estimating the 3D pose of desktop objects is crucial for applications such as robotic manipulation. Many existing approaches to this problem require a depth map of the object for both training and prediction, which restricts them to opaque,…

计算机视觉与模式识别 · 计算机科学 2020-05-20 Xingyu Liu , Rico Jonschkowski , Anelia Angelova , Kurt Konolige

Autonomous robotic systems operating in human environments must understand their surroundings to make accurate and safe decisions. In crowded human scenes with close-up human-robot interaction and robot navigation, a deep understanding…

计算机视觉与模式识别 · 计算机科学 2023-03-14 Edward Vendrow , Duy Tho Le , Jianfei Cai , Hamid Rezatofighi

Place recognition is a key module in robotic navigation. The existing line of studies mostly focuses on visual place recognition to recognize previously visited places solely based on their appearance. In this paper, we address structural…

机器人学 · 计算机科学 2021-09-29 Giseop Kim , Sunwook Choi , Ayoung Kim

Monocular RGB cameras mounted on drones are widely used for wildlife monitoring, yet most analytical pipelines remain confined to two-dimensional image space, leaving geometric information in video underexploited. We present WildLIFT, a…

计算机视觉与模式识别 · 计算机科学 2026-04-28 Vandita Shukla , Fabio Remondino , Blair Costelloe , Benjamin Risse