中文
相关论文

相关论文: MobiDiary: Autoregressive Action Captioning with W…

200 篇论文

Reliable and robust user identification and authentication are important and often necessary requirements for many digital services. It becomes paramount in social virtual reality (VR) to ensure trust, specifically in digital encounters…

机器学习 · 计算机科学 2023-10-27 Christian Schell , Andreas Hotho , Marc Erich Latoschik

Recent advances in text-driven human motion generation enable models to synthesize realistic motion sequences from natural language descriptions. However, most existing approaches assume identity-neutral motion and generate movements using…

计算机视觉与模式识别 · 计算机科学 2026-04-29 Wenqi Jia , Zekun Li , Abhay Mittal , Chengcheng Tang , Chuan Guo , Lezi Wang , James Matthew Rehg , Lingling Tao , Size An

Efficient querying and retrieval of healthcare data is posing a critical challenge today with numerous connected devices continuously generating petabytes of images, text, and internet of things (IoT) sensor data. One approach to…

机器学习 · 计算机科学 2023-02-28 Sazia Mahfuz , Farhana Zulkernine

Autoregressive (AR) models have demonstrated significant success in the realm of text-to-image generation. However, they usually face two major challenges. Firstly, the generated images may not always meet the quality standards expected by…

计算机视觉与模式识别 · 计算机科学 2026-04-03 Kai Dong , Tingting Bai

We perform classification of activities of daily living (ADL) using a Frequency-Modulated Continuous Waveform (FMCW) radar. In particular, we consider contiguous motions that are inseparable in time. Both the micro-Doppler signature and…

信号处理 · 电气工程与系统科学 2019-12-18 Moeness G. Amin , Ronny G. Guendel

Inertial Measurement Unit (IMU) sensors are widely employed for Human Activity Recognition (HAR) due to their portability, energy efficiency, and growing research interest. However, a significant challenge for IMU-HAR models is achieving…

计算机视觉与模式识别 · 计算机科学 2024-06-28 Qi Qiu , Tao Zhu , Furong Duan , Kevin I-Kai Wang , Liming Chen , Mingxing Nie , Mingxing Nie

Human Activity Recognition (HAR) describes the machines ability to recognize human actions. Nowadays, most people on earth are health conscious, so people are more interested in tracking their daily activities using Smartphones or Smart…

机器学习 · 计算机科学 2022-05-23 Sanku Satya Uday , Satti Thanuja Pavani , T. Jaya Lakshmi , Rohit Chivukula

Ensuring the safety and well-being of elderly and vulnerable populations in assisted living environments is a critical concern. Computer vision presents an innovative and powerful approach to predicting health risks through video…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Yixuan Wang , Paul Stynes , Pramod Pathak , Cristina Muntean

Anatomical movements of the human body can change the channel state information (CSI) of wireless signals in an indoor environment. These changes in the CSI signals can be used for human activity recognition (HAR), which is a predominant…

人机交互 · 计算机科学 2022-05-04 Hojjat Salehinejad , Shahrokh Valaee

Remote monitoring of motor functions is a powerful approach for health assessment, especially among the elderly population or among subjects affected by pathologies that negatively impact their walking capabilities. This is further…

信号处理 · 电气工程与系统科学 2022-05-18 Antonio Bevilacqua , Lisa Alcock , Brian Caulfield , Eran Gazit , Clint Hansen , Neil Ireson , Georgiana Ifrim

The research on human activity recognition has provided novel solutions to many applications like healthcare, sports, and user profiling. Considering the complex nature of human activities, it is still challenging even after effective and…

人机交互 · 计算机科学 2023-07-11 Ranjit Kolkar , Geetha V

It is indisputable that physical activity is vital for an individual's health and wellness. However, a global prevalence of physical inactivity has induced significant personal and socioeconomic implications. In recent years, a significant…

人工智能 · 计算机科学 2023-01-04 Asterios Bampakis , Sofia Yfantidou , Athena Vakali

Audio captioning aims to generate text descriptions of audio clips. In the real world, many objects produce similar sounds. How to accurately recognize ambiguous sounds is a major challenge for audio captioning. In this work, inspired by…

音频与语音处理 · 电气工程与系统科学 2023-05-30 Xubo Liu , Qiushi Huang , Xinhao Mei , Haohe Liu , Qiuqiang Kong , Jianyuan Sun , Shengchen Li , Tom Ko , Yu Zhang , Lilian H. Tang , Mark D. Plumbley , Volkan Kılıç , Wenwu Wang

Wearable sensors such as Inertial Measurement Units (IMUs) are often used to assess the performance of human exercise. Common approaches use handcrafted features based on domain expertise or automatically extracted features using time…

计算机视觉与模式识别 · 计算机科学 2023-07-11 Ashish Singh , Antonio Bevilacqua , Timilehin B. Aderinola , Thach Le Nguyen , Darragh Whelan , Martin O'Reilly , Brian Caulfield , Georgiana Ifrim

This study achieved bidirectional translation between descriptions and actions using small paired data from different modalities. The ability to mutually generate descriptions and actions is essential for robots to collaborate with humans…

机器人学 · 计算机科学 2022-09-27 Minori Toyoda , Kanata Suzuki , Yoshihiko Hayashi , Tetsuya Ogata

Human activity recognition (HAR) is a key challenge in pervasive computing and its solutions have been presented based on various disciplines. Specifically, for HAR in a smart space without privacy and accessibility issues, data streams…

机器学习 · 计算机科学 2023-12-04 Hyunju Kim , Heesuk Son , Dongman Lee

Motion, speech, and sound effects are fundamental elements of human-centric videos, yet their heterogeneous temporal characteristics make joint generation highly challenging. Existing audio-video generation models often fail to maintain…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Shihao Cheng , Jiaxu Zhang , Quanyue Song , Shansong Liu , Zhizhi Guo , Xiaolei Zhang , Chi Zhang , Xuelong Li , Zhigang Tu

While the widely available embedded sensors in smartphones and other wearable devices make it easier to obtain data of human activities, recognizing different types of human activities from sensor-based data remains a difficult research…

信号处理 · 电气工程与系统科学 2024-08-15 Taoran Sheng , Manfred Huber

In line with the human capacity to perceive the world by simultaneously processing and integrating high-dimensional inputs from multiple modalities like vision and audio, we propose a novel model, MAiVAR-T (Multimodal Audio-Image to Video…

计算机视觉与模式识别 · 计算机科学 2023-08-08 Muhammad Bilal Shaikh , Douglas Chai , Syed Mohammed Shamsul Islam , Naveed Akhtar

Video-based Emotional Reaction Intensity (ERI) estimation measures the intensity of subjects' reactions to stimuli along several emotional dimensions from videos of the subject as they view the stimuli. We propose a multi-modal architecture…

计算机视觉与模式识别 · 计算机科学 2023-05-10 Yini Fang , Liang Wu , Frederic Jumelle , Bertram Shi
‹ 上一页 1 8 9 10 下一页 ›