English
Related papers

Related papers: TransFusionOdom: Interpretable Transformer-based L…

200 papers

In recent years, multiple Light Detection and Ranging (LiDAR) systems have grown in popularity due to their enhanced accuracy and stability from the increased field of view (FOV). However, integrating multiple LiDARs can be challenging,…

Robotics · Computer Science 2023-11-08 Minwoo Jung , Sangwoo Jung , Ayoung Kim

Visual-LiDAR odometry is a critical component for autonomous system localization, yet achieving high accuracy and strong robustness remains a challenge. Traditional approaches commonly struggle with sensor misalignment, fail to fully…

Computer Vision and Pattern Recognition · Computer Science 2025-09-09 Mengmeng Liu , Michael Ying Yang , Jiuming Liu , Yunpeng Zhang , Jiangtao Li , Sander Oude Elberink , George Vosselman , Hao Cheng

Multimodal fusion frameworks for Human Action Recognition (HAR) using depth and inertial sensor data have been proposed over the years. In most of the existing works, fusion is performed at a single level (feature level or decision level),…

Machine Learning · Computer Science 2019-10-28 Zeeshan Ahmad , Naimul Khan

Human-machine interaction has been around for several decades now, with new applications emerging every day. One of the major goals that remain to be achieved is designing an interaction similar to how a human interacts with another human.…

Human-Computer Interaction · Computer Science 2022-12-27 Tauheed Khan Mohd , Nicole Nguyen , Ahmad Y Javaid

In recent years, Onboard Self Localization (OSL) methods based on cameras or Lidar have achieved many significant progresses. However, some issues such as estimation drift and feature-dependence still remain inherent limitations. On the…

Robotics · Computer Science 2020-10-26 Thien-Minh Nguyen , Shenghai Yuan , Muqing Cao , Yang Lyu , Thien Hoang Nguyen , Lihua Xie

Multimodal visual information fusion aims to integrate the multi-sensor data into a single image which contains more complementary information and less redundant features. However the complementary information is hard to extract, especially…

Computer Vision and Pattern Recognition · Computer Science 2024-06-18 Hui Li , Xiao-Jun Wu

For many years, there has been an impressive progress on visual odometry applied to mobile robots and drones. However, the visual perception is still in the spotlight as a challenging field because the vision sensor has some problems in…

Robotics · Computer Science 2021-09-09 Sungjae Shin , Eungchang Lee , Junho Choi , Hyun Myung

Recent years have witnessed the remarkable progress of 3D multi-modality object detection methods based on the Bird's-Eye-View (BEV) perspective. However, most of them overlook the complementary interaction and guidance between LiDAR and…

Computer Vision and Pattern Recognition · Computer Science 2024-11-04 Xiaotian Li , Baojie Fan , Jiandong Tian , Huijie Fan

The research introduces a reproducible framework for transforming raw, heterogeneous sensor streams into aligned, semantically meaningful representations for multimodal human activity recognition. Grounded in the Carnegie Mellon University…

Applications · Statistics 2026-05-05 Yiyao Yang , Yasemin Gulbahar

Multimodal sensor fusion is an essential capability for autonomous robots, enabling object detection and decision-making in the presence of failing or uncertain inputs. While recent fusion methods excel in normal environmental conditions,…

Computer Vision and Pattern Recognition · Computer Science 2025-08-25 Edoardo Palladin , Roland Dietze , Praveen Narayanan , Mario Bijelic , Felix Heide

Spatiotemporal fusion aims to improve both the spatial and temporal resolution of remote sensing images, thus facilitating time-series analysis at a fine spatial scale. However, there are several important issues that limit the application…

Computer Vision and Pattern Recognition · Computer Science 2024-02-27 Houcai Guo , Dingqi Ye , Lorenzo Bruzzone

The correct ego-motion estimation basically relies on the understanding of correspondences between adjacent LiDAR scans. However, given the complex scenarios and the low-resolution LiDAR, finding reliable structures for identifying…

Computer Vision and Pattern Recognition · Computer Science 2022-03-01 Yan Xu , Junyi Lin , Jianping Shi , Guofeng Zhang , Xiaogang Wang , Hongsheng Li

Multi-modal methods based on camera and LiDAR sensors have garnered significant attention in the field of 3D detection. However, many prevalent works focus on single or partial stage fusion, leading to insufficient feature extraction and…

Computer Vision and Pattern Recognition · Computer Science 2025-08-19 Zhiwei Ning , Zhaojiang Liu , Xuanang Gao , Yifan Zuo , Jie Yang , Yuming Fang , Wei Liu

Learning multi-modal representations is an essential step towards real-world robotic applications, and various multi-modal fusion models have been developed for this purpose. However, we observe that existing models, whose objectives are…

Machine Learning · Computer Science 2021-06-22 Chenzhuang Du , Tingle Li , Yichen Liu , Zixin Wen , Tianyu Hua , Yue Wang , Hang Zhao

In robotic navigation, maintaining precise pose estimation and navigation in complex and dynamic environments is crucial. However, environmental challenges such as smoke, tunnels, and adverse weather can significantly degrade the…

Robotics · Computer Science 2025-07-25 Chenglong Qian , Yang Xu , Xiufang Shi , Jiming Chen , Liang Li

Currently, the improvement of LiDAR poses estimation accuracy is an urgent need for mobile robots. Research indicates that diverse LiDAR points have different influences on the accuracy of pose estimation. This study aimed to select a good…

Robotics · Computer Science 2022-08-17 Zeyu Wan , Yu Zhang , Bin He , Zhuofan Cui , Weichen Dai , Lipu Zhou , Guoquan Huang

Visual odometry and Simultaneous Localization And Mapping (SLAM) has been studied as one of the most important tasks in the areas of computer vision and robotics, to contribute to autonomous navigation and augmented reality systems. In case…

Robotics · Computer Science 2023-11-08 Seongwook Yoon , Jaehyun Kim , Sanghoon Sull

The integration of data from diverse sensor modalities (e.g., camera and LiDAR) constitutes a prevalent methodology within the ambit of autonomous driving scenarios. Recent advancements in efficient point cloud transformers have underscored…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Yutao Zhu , Xiaosong Jia , Xinyu Yang , Junchi Yan

For 3D object detection, both camera and lidar have been demonstrated to be useful sensory devices for providing complementary information about the same scenery with data representations in different modalities, e.g., 2D RGB image vs 3D…

Computer Vision and Pattern Recognition · Computer Science 2023-11-08 Xinhao Xiang , Jiawei Zhang

Combining multiple LiDARs enables a robot to maximize its perceptual awareness of environments and obtain sufficient measurements, which is promising for simultaneous localization and mapping (SLAM). This paper proposes a system to achieve…

Robotics · Computer Science 2021-05-06 Jianhao Jiao , Haoyang Ye , Yilong Zhu , Ming Liu
‹ Prev 1 4 5 6 7 8 10 Next ›