中文
相关论文

相关论文: Robust Multimodal Learning Framework For Intake Ge…

200 篇论文

In motion tracking of connected multi-body systems Inertial Measurement Units (IMUs) are used in a wide variety of applications, since they provide a low-cost easy-to-use method for orientation estimation. However, in indoor environments or…

系统与控制 · 电气工程与系统科学 2021-08-11 Dustin Lehmann , Daniel Laidig , Raphael Deimel , Thomas Seel

Binge eating disorder (BED) is the most prevalent eating disorder. However, current diagnostic frameworks remain largely grounded in symptom-based criteria rather than underlying biological mechanisms, thereby limiting early detection and…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Lin Zhao , Qiaohui Gao , Elizabeth Martin , Kurt P. Schulz , Tom Hildebrandt , Robyn Sysko , Tianming Liu , Xiaobo Li

Continuous and multimodal stress detection has been performed recently through wearable devices and machine learning algorithms. However, a well-known and important challenge of working on physiological signals recorded by conventional…

机器学习 · 计算机科学 2021-07-30 Arman Iranfar , Adriana Arza , David Atienza

This paper proposes a novel multi-modal transformer network for detecting actions in untrimmed videos. To enrich the action features, our transformer network utilizes a new multi-modal attention mechanism that computes the correlations…

计算机视觉与模式识别 · 计算机科学 2023-06-01 Matthew Korban , Scott T. Acton , Peter Youngs

Inertial odometry (IO) using strap-down inertial measurement units (IMUs) is critical in many robotic applications where precise orientation and position tracking are essential. Prior kinematic motion model-based IO methods often use a…

机器人学 · 计算机科学 2024-05-16 Yuheng Qiu , Chen Wang , Can Xu , Yutian Chen , Xunfei Zhou , Youjie Xia , Sebastian Scherer

Stable and robust robotic grasping is essential for current and future robot applications. In recent works, the use of large datasets and supervised learning has enhanced speed and precision in antipodal grasping. However, these methods…

机器人学 · 计算机科学 2025-02-28 Boya Zhang , Iris Andrussow , Andreas Zell , Georg Martius

As wearable and mobile devices become increasingly embedded in daily life, they offer a practical way to continuously sense human motion in the wild. But inertial signals are highly dependent on the sensing setup, including body location,…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Baiyu Chen , Zechen Li , Wilson Wongso , Lihuan Li , Xiachong Lin , Hao Xue , Benjamin Tag , Flora Salim

This paper proposes a hybrid learning and optimization framework for mobile manipulators for complex and physically interactive tasks. The framework exploits an admittance-type physical interface to obtain intuitive and simplified human…

机器人学 · 计算机科学 2022-08-02 Jianzhuang Zhao , Alberto Giammarino , Edoardo Lamon , Juan M. Gandarias , Elena De Momi , Arash Ajoudani

Simultaneously using multimodal inputs from multiple sensors to train segmentors is intuitively advantageous but practically challenging. A key challenge is unimodal bias, where multimodal segmentors over rely on certain modalities, causing…

计算机视觉与模式识别 · 计算机科学 2025-05-19 Xu Zheng , Haiwei Xue , Jialei Chen , Yibo Yan , Lutao Jiang , Yuanhuiyi Lyu , Kailun Yang , Linfeng Zhang , Xuming Hu

Accurate and reliable estimation of biases of low-cost Inertial Measurement Units (IMU) is a key factor to maintain the resilience of Visual-Inertial Odometry (VIO), particularly when visual tracking fails in challenging areas. In such…

计算机视觉与模式识别 · 计算机科学 2025-10-20 Yang Yi , Kunqing Wang , Jinpu Zhang , Zhen Tan , Xiangke Wang , Hui Shen , Dewen Hu

Contact-rich manipulation requires reliable estimation of extrinsic contacts-the interactions between a grasped object and its environment which provide essential contextual information for planning, control, and policy learning. However,…

机器人学 · 计算机科学 2026-02-03 Zhengtong Xu , Yuki Shirai

Multimodal learning integrates complementary information from different modalities such as image, text, and audio to improve model performance, but its success relies on large-scale labeled data, which is costly to obtain. Active learning…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Yuqiao Zeng , Xu Wang , Tengfei Liang , Yiqing Hao , Yi Jin , Hui Yu

Wireless communications at high-frequency bands with large antenna arrays face challenges in beam management, which can potentially be improved by multimodality sensing information from cameras, LiDAR, radar, and GPS. In this paper, we…

信号处理 · 电气工程与系统科学 2023-09-22 Yu Tian , Qiyang Zhao , Zine el abidine Kherroubi , Fouzi Boukhalfa , Kebin Wu , Faouzi Bader

We present a method for gesture detection and localisation based on multi-scale and multi-modal deep learning. Each visual modality captures spatial information at a particular spatial scale (such as motion of the upper body or a hand), and…

计算机视觉与模式识别 · 计算机科学 2015-07-21 Natalia Neverova , Christian Wolf , Graham W. Taylor , Florian Nebout

Wearable sensor-based Human Action Recognition (HAR) has made significant strides in recent times. However, the accuracy performance of wearable sensor-based HAR is currently still lagging behind that of visual modalities-based systems,…

多媒体 · 计算机科学 2024-05-21 Jianyuan Ni , Hao Tang , Anne H. H. Ngu , Gaowen Liu , Yan Yan

Inertial measurement units (IMUs) are fundamental sensing components in multi-source integrated navigation systems, and their performance directly determines the accuracy and reliability of solutions. However, the precision of low-cost IMUs…

信号处理 · 电气工程与系统科学 2026-05-19 Jiarui Lv , Feng Zhu , Xiaohong Zhang

We present an efficient approach for leveraging the knowledge from multiple modalities in training unimodal 3D convolutional neural networks (3D-CNNs) for the task of dynamic hand gesture recognition. Instead of explicitly combining…

计算机视觉与模式识别 · 计算机科学 2025-10-13 Mahdi Abavisani , Hamid Reza Vaezi Joze , Vishal M. Patel

In precision sports such as archery, athletes' performance depends on both biomechanical stability and psychological resilience. Traditional motion analysis systems are often expensive and intrusive, limiting their use in natural training…

机器学习 · 计算机科学 2025-11-19 Xianghe Liu , Jiajia Liu , Chuxian Xu , Minghan Wang , Hongbo Peng , Tao Sun , Jiaqi Xu

This paper presents a novel deep learning framework for robotic arm manipulation that integrates multimodal inputs using a late-fusion strategy. Unlike traditional end-to-end or reinforcement learning approaches, our method processes image…

机器学习 · 计算机科学 2025-04-07 Sathish Kumar , Swaroop Damodaran , Naveen Kumar Kuruba , Sumit Jha , Arvind Ramanathan

Food recognition is an important task for a variety of applications, including managing health conditions and assisting visually impaired people. Several food recognition studies have focused on generic types of food or specific cuisines,…

计算机视觉与模式识别 · 计算机科学 2022-04-21 Şeymanur Aktı , Marwa Qaraqe , Hazım Kemal Ekenel