中文
相关论文

相关论文: Masked Video and Body-worn IMU Autoencoder for Ego…

200 篇论文

We propose a multi-sensor fusion method for capturing challenging 3D human motions with accurate consecutive local poses and global trajectories in large-scale scenarios, only using single LiDAR and 4 IMUs, which are set up conveniently and…

计算机视觉与模式识别 · 计算机科学 2023-04-11 Yiming Ren , Chengfeng Zhao , Yannan He , Peishan Cong , Han Liang , Jingyi Yu , Lan Xu , Yuexin Ma

This paper presents a novel inertial localization framework named Egocentric Action-aware Inertial Localization (EAIL), which leverages egocentric action cues from head-mounted IMU signals to localize the target individual within a 3D point…

计算机视觉与模式识别 · 计算机科学 2025-07-29 Mingfang Zhang , Ryo Yonetani , Yifei Huang , Liangyang Ouyang , Ruicong Liu , Yoichi Sato

Robust and accurate proprioceptive state estimation of the main body is crucial for legged robots to execute tasks in extreme environments where exteroceptive sensors, such as LiDARs and cameras, may become unreliable. In this paper, we…

机器人学 · 计算机科学 2025-07-29 Yibin Wu , Jian Kuang , Shahram Khorshidi , Xiaoji Niu , Lasse Klingbeil , Maren Bennewitz , Heiner Kuhlmann

Similar to humans, robots benefit from interacting with their environment through a number of different sensor modalities, such as vision, touch, sound. However, learning from different sensor modalities is difficult, because the learning…

机器人学 · 计算机科学 2019-10-10 Martina Zambelli , Antoine Cully , Yiannis Demiris

This work explores the effectiveness of masked image modelling for learning representations of retinal OCT images. To this end, we leverage Masked Autoencoders (MAE), a simple and scalable method for self-supervised learning, to obtain a…

计算机视觉与模式识别 · 计算机科学 2024-05-24 Theodoros Pissas , Pablo Márquez-Neila , Sebastian Wolf , Martin Zinkernagel , Raphael Sznitman

Human activity recognition (HAR) on smartglasses has various use cases, including health/fitness tracking and input for context-aware AI assistants. However, current approaches for egocentric activity recognition suffer from low performance…

计算机视觉与模式识别 · 计算机科学 2025-04-25 Akhil Padmanabha , Saravanan Govindarajan , Hwanmun Kim , Sergio Ortiz , Rahul Rajan , Doruk Senkal , Sneha Kadetotad

This article presents a method to automatically detect and classify climbing activities using inertial measurement units (IMUs) attached to the wrists, feet and pelvis of the climber. The IMUs record limb acceleration and angular velocity.…

Wearable inertial motion capture (MoCap) provides a portable, occlusion-free, and privacy-preserving alternative to camera-based systems, but its accuracy depends on tightly attached sensors - an intrusive and uncomfortable requirement for…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Jiawei Fang , Ruonan Zheng , Xiaoxia Gao , Shifan Jiang , Anjun Chen , Qi Ye , Shihui Guo

Motion capture (MoCap) data from wearable Inertial Measurement Units (IMUs) is vital for applications in sports science, but its utility is often compromised by missing data. Despite numerous imputation techniques, a systematic performance…

机器学习 · 计算机科学 2025-07-15 Mahmoud Bekhit , Ahmad Salah , Ahmed Salim Alrawahi , Tarek Attia , Ahmed Ali , Esraa Eldesokey , Ahmed Fathalla

Masked Autoencoders (MAE) play a pivotal role in learning potent representations, delivering outstanding results across various 3D perception tasks essential for autonomous driving. In real-world driving scenarios, it's commonplace to…

计算机视觉与模式识别 · 计算机科学 2024-08-26 Jian Zou , Tianyu Huang , Guanglei Yang , Zhenhua Guo , Tao Luo , Chun-Mei Feng , Wangmeng Zuo

Human action recognition has been widely used in many fields of life, and many human action datasets have been published at the same time. However, most of the multi-modal databases have some shortcomings in the layout and number of…

计算机视觉与模式识别 · 计算机科学 2022-02-08 Xin Chao , Zhenjie Hou , Yujian Mo

The lack of large-scale, labeled data sets impedes progress in developing robust and generalized predictive models for on-body sensor-based human activity recognition (HAR). Labeled data in human activity recognition is scarce and hard to…

计算机视觉与模式识别 · 计算机科学 2020-08-05 Hyeokhyen Kwon , Catherine Tong , Harish Haresamudram , Yan Gao , Gregory D. Abowd , Nicholas D. Lane , Thomas Ploetz

We present an extension to masked autoencoders (MAE) which improves on the representations learnt by the model by explicitly encouraging the learning of higher scene-level features. We do this by: (i) the introduction of a perceptual…

计算机视觉与模式识别 · 计算机科学 2023-03-29 Samyakh Tukra , Frederick Hoffman , Ken Chatfield

Recently, action recognition has been dominated by transformer-based methods, thanks to their spatiotemporal contextual aggregation capacities. However, despite the significant progress achieved on scene-related datasets, they do not…

计算机视觉与模式识别 · 计算机科学 2025-10-24 Peiqin Zhuang , Lei Bai , Yichao Wu , Ding Liang , Luping Zhou , Yali Wang , Wanli Ouyang

Marker-based and marker-less optical skeletal motion-capture methods use an outside-in arrangement of cameras placed around a scene, with viewpoints converging on the center. They often create discomfort by possibly needed marker suits, and…

计算机视觉与模式识别 · 计算机科学 2016-09-26 Helge Rhodin , Christian Richardt , Dan Casas , Eldar Insafutdinov , Mohammad Shafiei , Hans-Peter Seidel , Bernt Schiele , Christian Theobalt

Current video-based Masked Autoencoders (MAEs) primarily focus on learning effective spatiotemporal representations from a visual perspective, which may lead the model to prioritize general spatial-temporal patterns but often overlook…

计算机视觉与模式识别 · 计算机科学 2025-02-13 Shihab Aaqil Ahamed , Malitha Gunawardhana , Liel David , Michael Sidorov , Daniel Harari , Muhammad Haris Khan

Missing input sequences are common in medical imaging data, posing a challenge for deep learning models reliant on complete input data. In this work, inspired by MultiMAE [2], we develop a masked autoencoder (MAE) paradigm for multi-modal,…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Ayhan Can Erdur , Christian Beischl , Daniel Scholz , Jiazhen Pan , Benedikt Wiestler , Daniel Rueckert , Jan C Peeken

In the human activity recognition research area, prior studies predominantly concentrate on leveraging advanced algorithms on public datasets to enhance recognition performance, little attention has been paid to executing real-time kitchen…

信号处理 · 电气工程与系统科学 2024-09-11 Mengxi Liu , Sungho Suh , Juan Felipe Vargas , Bo Zhou , Agnes Grünerbl , Paul Lukowicz

Event cameras are an interesting visual exteroceptive sensor that reacts to brightness changes rather than integrating absolute image intensities. Owing to this design, the sensor exhibits strong performance in situations of challenging…

计算机视觉与模式识别 · 计算机科学 2024-08-05 Runze Yuan , Tao Liu , Zijia Dai , Yi-Fan Zuo , Laurent Kneip

Imitation learning is a powerful paradigm for robot skill acquisition, yet conventional demonstration methods--such as kinesthetic teaching and teleoperation--are cumbersome, hardware-heavy, and disruptive to workflows. Recently, passive…

机器人学 · 计算机科学 2025-09-30 Rohan Walia , Yusheng Wang , Ralf Römer , Masahiro Nishio , Angela P. Schoellig , Jun Ota