English
Related papers

Related papers: X-View: Graph-Based Semantic Multi-View Localizati…

200 papers

Global visual geolocation predicts where an image was captured on Earth. Since images vary in how precisely they can be localized, this task inherently involves a significant degree of ambiguity. However, existing approaches are…

Computer Vision and Pattern Recognition · Computer Science 2024-12-10 Nicolas Dufour , David Picard , Vicky Kalogeiton , Loic Landrieu

Skeleton-based human action recognition has been drawing more interest recently due to its low sensitivity to appearance changes and the accessibility of more skeleton data. However, even the 3D skeletons captured in practice are still…

Computer Vision and Pattern Recognition · Computer Science 2022-09-26 Cunling Bian , Wei Feng , Fanbo Meng , Song Wang

Visual recognition models have achieved unprecedented success in various tasks. While researchers aim to understand the underlying mechanisms of these models, the growing demand for deployment in safety-critical areas like autonomous…

Computer Vision and Pattern Recognition · Computer Science 2026-03-12 Qiyang Wan , Chengzhi Gao , Ruiping Wang , Xilin Chen

In recent years, developing AI for robotics has raised much attention. The interaction of vision and language of robots is particularly difficult. We consider that giving robots an understanding of visual semantics and language semantics…

Robotics · Computer Science 2021-05-26 Cheng Yu Tsai , Mu-Chun Su

Cross-view matching refers to the problem of finding the closest match for a given query ground view image to one from a database of aerial images. If the aerial images are geotagged, then the closest matching aerial image can be used to…

Robotics · Computer Science 2020-11-17 Deeksha Dixit , Surabhi Verma , Pratap Tokekar

In this paper, we propose a general graph optimization based framework for localization, which can accommodate different types of measurements with varying measurement time intervals. Special emphasis will be on range-based localization.…

Robotics · Computer Science 2020-01-27 Xu Fang , Chen Wang , Thien-Minh Nguyen , Lihua Xie

Comprehensive perception of human beings is the prerequisite to ensure the safety of human-robot interaction. Currently, prevailing visual sensing approach typically involves a single static camera, resulting in a restricted and occluded…

Robotics · Computer Science 2024-03-20 Yuanjiong Ying , Xian Huang , Wei Dong

This paper describes a method of global localization based on graph-theoretic association of instances between a query and the prior map. The proposed framework employs correspondence matching based on the maximum clique problem (MCP). The…

Robotics · Computer Science 2023-06-07 Shigemichi Matsuzaki , Kenji Koide , Shuji Oishi , Masashi Yokozuka , Atsuhiko Banno

Navigational signs enable humans to navigate unfamiliar environments without maps. This work studies how robots can similarly exploit signs for mapless navigation in the open world. A central challenge lies in interpreting signs: real-world…

Robotics · Computer Science 2026-02-16 Nicky Zimmerman , Joel Loo , Benjamin Koh , Zishuo Wang , David Hsu

Simultaneous localization and mapping (SLAM) in slowly varying scenes is important for long-term robot task completion. Failing to detect scene changes may lead to inaccurate maps and, ultimately, lost robots. Classical SLAM algorithms…

We introduce a new large-scale dataset for the advancement of object detection techniques and overhead object detection research. This satellite imagery dataset enables research progress pertaining to four key computer vision frontiers. We…

Computer Vision and Pattern Recognition · Computer Science 2018-02-23 Darius Lam , Richard Kuzma , Kevin McGee , Samuel Dooley , Michael Laielli , Matthew Klaric , Yaroslav Bulatov , Brendan McCord

Until open-world foundation models match the performance of specialized approaches, deep learning systems remain dependent on task- and sensor-specific data availability. To bridge the gap between available datasets and deployment domains,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Frank Bieder , Hendrik Königshof , Haohao Hu , Fabian Immel , Yinzhe Shen , Jan-Hendrik Pauls , Christoph Stiller

Several deployment locations of mobile robotic systems are human made (i.e. urban firefighter, building inspection, property security) and the manager may have access to domain-specific knowledge about the place, which can provide semantic…

Robotics · Computer Science 2021-06-21 Rafael Gomes Braga , Sina Karimi , Ulrich Dah-Achinanon , Ivanka Iordanova , David St-Onge

Visual information displays are typically composed of multiple visualizations that are used to facilitate an understanding of the underlying data. A common example are dashboards, which are frequently used in domains such as finance,…

Human-Computer Interaction · Computer Science 2021-09-20 Yngve S. Kristiansen , Laura Garrison , Stefan Bruckner

The task of UAV-view geo-localization is to estimate the localization of a query satellite/drone image by matching it against a reference dataset consisting of drone/satellite images. Though tremendous strides have been made in feature…

Computer Vision and Pattern Recognition · Computer Science 2024-07-04 Jie Shao , LingHao Jiang

Robust visual localization for urban vehicles remains challenging and unsolved. The limitation of computation efficiency and memory size has made it harder for large-scale applications. Since semantic information serves as a stable and…

Robotics · Computer Science 2020-10-14 Ziwei Liao , Jieqi Shi , Xianyu Qi , Xiaoyu Zhang , Wei Wang , Yijia He , Ran Wei , Xiao Liu

This article introduces a novel method for object-level relocalization of robotic systems. It determines the pose of a camera sensor by robustly associating the object detections in the current frame with 3D objects in a lightweight…

Robotics · Computer Science 2024-08-16 Yutong Wang , Chaoyang Jiang , Xieyuanli Chen

Supervised keypoint localization methods rely on large manually labeled image datasets, where objects can deform, articulate, or occlude. However, creating such large keypoint labels is time-consuming and costly, and is often error-prone…

Computer Vision and Pattern Recognition · Computer Science 2023-03-31 Xingzhe He , Gaurav Bharaj , David Ferman , Helge Rhodin , Pablo Garrido

The increasing complexity of machine learning models in computer vision, particularly in face verification, requires the development of explainable artificial intelligence (XAI) to enhance interpretability and transparency. This study…

Computer Vision and Pattern Recognition · Computer Science 2025-01-13 Miriam Doh , Caroline Mazini Rodrigues , N. Boutry , L. Najman , Matei Mancas , Bernard Gosselin

Accurate multispectral image matching presents significant challenges due to non-linear intensity variations across spectral modalities, extreme viewpoint changes, and the scarcity of labeled datasets. Current state-of-the-art methods are…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Ismail Can Yagmur , Hasan F. Ates , Bahadir K. Gunturk