中文
相关论文

相关论文: Multimodal Group Activity Dataset for Classroom En…

200 篇论文

We present a novel LLM-based pipeline for creating contextual descriptions of human body poses in images using only auxiliary attributes. This approach facilitates the creation of the MPII Pose Descriptions dataset, which includes natural…

计算机视觉与模式识别 · 计算机科学 2024-10-30 Muhammad Saif Ullah Khan , Muhammad Ferjad Naeem , Federico Tombari , Luc Van Gool , Didier Stricker , Muhammad Zeshan Afzal

Group activity recognition is a hot topic in computer vision. Recognizing activities through group relationships plays a vital role in group activity recognition. It holds practical implications in various scenarios, such as video analysis,…

计算机视觉与模式识别 · 计算机科学 2023-07-26 Chuanchuan Wang , Ahmad Sufril Azlan Mohamed

We describe the DeepMind Kinetics human action video dataset. The dataset contains 400 human action classes, with at least 400 video clips for each action. Each clip lasts around 10s and is taken from a different YouTube video. The actions…

Communicating in noisy, multi-talker environments is challenging, especially for people with hearing impairments. Egocentric video data can potentially be used to identify a user's conversation partners, which could be used to inform…

计算机视觉与模式识别 · 计算机科学 2024-06-13 Tobias Dorszewski , Søren A. Fuglsang , Jens Hjortkjær

We introduce the Visual Experience Dataset (VEDB), a compilation of over 240 hours of egocentric video combined with gaze- and head-tracking data that offers an unprecedented view of the visual world as experienced by human observers. The…

In this study, we present a comprehensive public dataset for driver drowsiness detection, integrating multimodal signals of facial, behavioral, and biometric indicators. Our dataset includes 3D facial video using a depth camera, IR camera…

计算机视觉与模式识别 · 计算机科学 2025-07-21 Morteza Bodaghi , Majid Hosseini , Raju Gottumukkala , Ravi Teja Bhupatiraju , Iftikhar Ahmad , Moncef Gabbouj

Multi-agent behavior modeling aims to understand the interactions that occur between agents. We present a multi-agent dataset from behavioral neuroscience, the Caltech Mouse Social Interactions (CalMS21) Dataset. Our dataset consists of…

Sustained effort is essential for realizing the benefits of intelligent tutoring systems (ITS), yet many learners disengage or underuse available practice time. We introduce engagement forecasting as a supervised prediction task based on…

机器学习 · 计算机科学 2026-05-14 Eric S. Qiu , Danielle R. Thomas , Boyuan Guo , Vincent Aleven , Conrad Borchers

While recommender systems with multi-modal item representations (image, audio, and text), have been widely explored, learning recommendations from multi-modal user interactions (e.g., clicks and speech) remains an open problem. We study the…

信息检索 · 计算机科学 2024-05-08 Simone Borg Bruun , Krisztian Balog , Maria Maistro

This paper presents VisioPhysioENet, a novel multimodal system that leverages visual and physiological signals to detect learner engagement. It employs a two-level approach for extracting both visual and physiological features. For visual…

计算机视觉与模式识别 · 计算机科学 2025-08-21 Alakhsimar Singh , Kanav Goyal , Nischay Verma , Puneet Kumar , Xiaobai Li , Amritpal Singh

Modeling tap or click sequences of users on a mobile device can improve our understandings of interaction behavior and offers opportunities for UI optimization by recommending next element the user might want to click on. We analyzed a…

机器学习 · 计算机科学 2021-08-12 Xin Zhou , Yang Li

We present a method to study engagement level uniformity in a class of students. We validate our method by comparing two semesters taught using different methods in a physics and mathematics course. The first semester used conventional…

物理教育 · 物理学 2016-11-11 George C. Cardoso

The integration of information across multiple modalities and across time is a promising way to enhance the emotion recognition performance of affective systems. Much previous work has focused on instantaneous emotion recognition. The 2018…

图像与视频处理 · 电气工程与系统科学 2018-05-07 Didan Deng , Yuqian Zhou , Jimin Pi , Bertram E. Shi

Crowd density level estimation is an essential aspect of crowd safety since it helps to identify areas of probable overcrowding and required conditions. Nowadays, AI systems can help in various sectors. Here for safety purposes or many for…

密码学与安全 · 计算机科学 2024-05-14 Mahira Arefin , Md. Anwar Hussen Wadud , Anichur Rahman

This notebook paper describes our system for the untrimmed classification task in the ActivityNet challenge 2016. We investigate multiple state-of-the-art approaches for action recognition in long, untrimmed videos. We exploit hand-crafted…

计算机视觉与模式识别 · 计算机科学 2017-04-13 Yi Zhu , Shawn Newsam , Zaikun Xu

Measuring online behavioural student engagement often relies on simple count indicators or retrospective, predictive methods, which present challenges for real-time application. To address these limitations, we reconceptualise an existing…

计算机与社会 · 计算机科学 2025-07-17 Laura J. Johnston , Jim E. Griffin , Ioanna Manolopoulou , Takoua Jendoubi

Accurately detecting student behavior in classroom videos can aid in analyzing their classroom performance and improving teaching effectiveness. However, the current accuracy rate in behavior detection is low. To address this challenge, we…

计算机视觉与模式识别 · 计算机科学 2024-09-10 Fan Yang , Tao Wang , Xiaofei Wang

Hand pose estimation from egocentric video has broad implications across various domains, including human-computer interaction, assistive technologies, activity recognition, and robotics, making it a topic of significant research interest.…

计算机视觉与模式识别 · 计算机科学 2024-09-12 Olga Taran , Damian M. Manzone , Jose Zariffa

Complex activity recognition can benefit from understanding the steps that compose them. Current datasets, however, are annotated with one label only, hindering research in this direction. In this paper, we describe a new dataset for…

An adaptive guidance system that supports equipment operators requires a comprehensive model, which involves a variety of user behaviors that considers different skill and knowledge levels, as well as rapid-changing task situations. In the…

人机交互 · 计算机科学 2020-09-17 Chen Long-fei , Yuichi Nakamura , Kazuaki Kondo
‹ 上一页 1 8 9 10 下一页 ›