中文
相关论文

相关论文: LFTag: A Scalable Visual Fiducial System with Low …

200 篇论文

Fast and efficient identify a large number of RFID tags in the region of interest is a critical issue in various RFID applications. In this paper, a novel sub-frame-based algorithm with a time-efficient frame size adjustment strategy to…

网络与互联网体系结构 · 计算机科学 2018-05-10 Jian Su , Kexiong Liu , Haipeng Chen , Yu Han

This paper presents L-VITeX, a lightweight visual intuition system for terrain exploration designed for resource-constrained robots and swarms. L-VITeX aims to provide a hint of Regions of Interest (RoIs) without computationally expensive…

机器人学 · 计算机科学 2024-10-11 Antar Mazumder , Zarin Anjum Madhiha

Recognizing places using Lidar in large-scale environments is challenging due to the sparse nature of point cloud data. In this paper we present BVMatch, a Lidar-based frame-to-frame place recognition framework, that is capable of…

计算机视觉与模式识别 · 计算机科学 2025-06-26 Lun Luo , Si-Yuan Cao , Bin Han , Hui-Liang Shen , Junwei Li

With the advancement of communication and security technologies, it has become crucial to have robustness of embedded biometric systems. This paper presents the realization of such technologies which demands reliable and error-free…

计算机视觉与模式识别 · 计算机科学 2012-04-20 Aamir Khan , Muhammad Farhan , Asar Ali

Fast and Relaxed Vector Fitting (FRVF) is a frequency-domain system identification approach that has been widely adopted in electrical system modelling, while its application to mechanical systems has remained relatively unexplored. In this…

信号处理 · 电气工程与系统科学 2026-05-18 Beatrice E. Bauret Martínez , Gabriele Dessena , Marco Civera , Oscar E. Bonilla-Manrique

We propose an accurate and interpretable fine-grained cross-view localization method that estimates the 3 Degrees of Freedom (DoF) pose of a ground-level image by matching its local features with a reference aerial image. Unlike prior…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Zimin Xia , Chenghao Xu , Alexandre Alahi

Reliable localization is one of the most important parts of an MAV system. Localization in an indoor GPS-denied environment is a relatively difficult problem. Current vision based algorithms track optical features to calculate odometry. We…

机器人学 · 计算机科学 2017-09-18 Manash Pratim Das , Gaurav Gardi , Jayanta Mukhopadhyay

Few-shot Video Object Detection (FSVOD) addresses the challenge of detecting novel objects in videos with limited labeled examples, overcoming the constraints of traditional detection methods that require extensive training data. This task…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Yogesh Kumar , Anand Mishra

For several tasks, ranging from manipulation to inspection, it is beneficial for robots to localize a target object in their surroundings. In this paper, we propose an approach that utilizes coarse point clouds obtained from miniaturized…

机器人学 · 计算机科学 2025-08-01 Giammarco Caroleo , Alessandro Albini , Daniele De Martini , Timothy D. Barfoot , Perla Maiolino

Large-scale medical biobanks provide imaging data complemented by extensive tabular information, such as clinical measurements or demographics. However, this abundance of tabular attributes does not reflect real-world datasets, where only a…

计算机视觉与模式识别 · 计算机科学 2026-05-27 Marta Hasny , Laura Daza , Keno Bressem , Maxime Di Folco , Julia Schnabel

Although the recent image-based 3D object detection methods using Pseudo-LiDAR representation have shown great capabilities, a notable gap in efficiency and accuracy still exist compared with LiDAR-based methods. Besides, over-reliance on…

计算机视觉与模式识别 · 计算机科学 2021-01-01 Peixuan Li , Shun Su , Huaici Zhao

Metrics for Visual Grounding (VG) in Visual Question Answering (VQA) systems primarily aim to measure a system's reliance on relevant parts of the image when inferring an answer to the given question. Lack of VG has been a common problem…

计算机视觉与模式识别 · 计算机科学 2024-01-17 Daniel Reich , Felix Putze , Tanja Schultz

Most deepfake detection methods focus on detecting spatial and/or spatio-temporal changes in facial attributes and are centered around the binary classification task of detecting whether a video is real or fake. This is because available…

计算机视觉与模式识别 · 计算机科学 2023-07-18 Zhixi Cai , Shreya Ghosh , Abhinav Dhall , Tom Gedeon , Kalin Stefanov , Munawar Hayat

Traditional vision-based autonomous driving systems often face difficulties in navigating complex environments when relying solely on single-image inputs. To overcome this limitation, incorporating temporal data such as past image frames or…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Tuong Do , Binh X. Nguyen , Quang D. Tran , Erman Tjiputra , Te-Chuan Chiu , Anh Nguyen

Data-driven visual odometry (VO) is a critical subroutine for autonomous edge robotics, and recent progress in the field has produced highly accurate point predictions in complex environments. However, emerging autonomous edge robotics…

计算机视觉与模式识别 · 计算机科学 2023-03-07 Alex C. Stutts , Danilo Erricolo , Theja Tulabandhula , Amit Ranjan Trivedi

In this paper, we propose a fixed-size object encoding method (FOE-VRD) to improve performance of visual relationship detection tasks. Comparing with previous methods, FOE-VRD has an important feature, i.e., it uses one fixed-size vector to…

计算机视觉与模式识别 · 计算机科学 2020-06-01 Hengyue Pan , Xin Niu , Rongchun Li , Siqi Shen , Yong Dou

Estimating relative camera poses between images has been a central problem in computer vision. Methods that find correspondences and solve for the fundamental matrix offer high precision in most cases. Conversely, methods predicting pose…

计算机视觉与模式识别 · 计算机科学 2024-03-06 Chris Rockwell , Nilesh Kulkarni , Linyi Jin , Jeong Joon Park , Justin Johnson , David F. Fouhey

Consistent localization of cooperative multi-robot systems during navigation presents substantial challenges. This paper proposes a fault-tolerant, multi-modal localization framework for multi-robot systems on matrix Lie groups. We…

机器人学 · 计算机科学 2025-05-05 Mahboubeh Zarei , Robin Chhabra

Multi-agents rely on accurate poses to share and align observations, enabling a collaborative perception of the environment. However, traditional GNSS-based localization often fails in GNSS-denied environments, making consistent feature…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Wenkai Lin , Qiming Xia , Wen Li , Xun Huang , Chenglu Wen

Vision-Language-Action (VLA) models process visual inputs independently at each timestep, discarding valuable temporal information inherent in robotic manipulation tasks. This frame-by-frame processing makes models vulnerable to visual…

计算机视觉与模式识别 · 计算机科学 2025-11-17 Chenghao Liu , Jiachen Zhang , Chengxuan Li , Zhimu Zhou , Shixin Wu , Songfang Huang , Huiling Duan