中文
相关论文

相关论文: AVCAffe: A Large Scale Audio-Visual Dataset of Cog…

200 篇论文

Everyday speech conveys far more than words, it reflects who we are, how we feel, and the circumstances surrounding our interactions. Yet, most existing speech datasets are acted, limited in scale, and fail to capture the expressive…

音频与语音处理 · 电气工程与系统科学 2025-11-04 Zongyang Du , Shreeram Suresh Chandra , Ismail Rasim Ulgen , Aurosweta Mahapatra , Ali N. Salman , Carlos Busso , Berrak Sisman

The COVID-19 pandemic is having a dramatic impact on societies and economies around the world. With various measures of lockdowns and social distancing in place, it becomes important to understand emotional responses on a large scale. In…

计算与语言 · 计算机科学 2020-05-15 Bennett Kleinberg , Isabelle van der Vegt , Maximilian Mozes

Earable acoustic sensing offers a powerful and non-invasive modality for capturing fine-grained auditory and physiological signals directly from the ear canal, enabling continuous and context-aware monitoring of cognitive states. As earable…

人机交互 · 计算机科学 2025-12-23 Xijia Wei , Ting Dang , Khaldoon Al-Naimi , Yang Liu , Fahim Kawsar , Alessandro Montanari

Despite impressive advancements in multilingual corpora collection and model training, developing large-scale deployments of multilingual models still presents a significant challenge. This is particularly true for language tasks that are…

Systems like ChatGPT and Claude assist billions through proactive dialogue-offering unsolicited, task-relevant information. Drawing on Cognitive Load Theory, we study how cognitive load shapes performance in AI-assisted knowledge work. We…

人工智能 · 计算机科学 2026-03-10 Brandon Lepine , Juho Kim , Pamela Mishkin , Matthew Beane

Oral presentation skills are a critical component of higher education, yet comprehensive datasets capturing real-world student performance across multiple modalities remain scarce. To address this gap, we present SOPHIAS (Student Oral…

人机交互 · 计算机科学 2026-01-13 Alvaro Becerra , Ruth Cobos , Roberto Daza

Large Language Models (LLMs) are increasingly engaged in emotionally vulnerable conversations that extend beyond information seeking to moments of personal distress. As they adopt affective tones and simulate empathy, they risk creating the…

计算与语言 · 计算机科学 2026-01-23 Sewon Kim , Jiwon Kim , Seungwoo Shin , Hyejin Chung , Daeun Moon , Yejin Kwon , Hyunsoo Yoon

The ability to monitor audience reactions is critical when delivering presentations. However, current videoconferencing platforms offer limited solutions to support this. This work leverages recent advances in affect sensing to capture and…

人机交互 · 计算机科学 2021-02-01 Prasanth Murali , Javier Hernandez , Daniel McDuff , Kael Rowan , Jina Suh , Mary Czerwinski

In simultaneous interpreting, an interpreter renders a source speech into another language with a very short lag, much sooner than sentences are finished. In order to understand and later reproduce this dynamic and complex task…

计算与语言 · 计算机科学 2025-06-06 Dávid Javorský , Ondřej Bojar , François Yvon

Action Quality Assessment (AQA) -- the task of quantifying how well an action is performed -- has great potential for detecting errors in gym weight training, where accurate feedback is critical to prevent injuries and maximize gains.…

计算机视觉与模式识别 · 计算机科学 2026-04-06 Hao Yin , Lijun Gu , Paritosh Parmar , Lin Xu , Tianxiao Guo , Xiujin Liu , Weiwei Fu , Yang Zhang , Tianyou Zheng

Human affect recognition is an essential part of natural human-computer interaction. However, current methods are still in their infancy, especially for in-the-wild data. In this work, we introduce our submission to the Affective Behavior…

计算机视觉与模式识别 · 计算机科学 2020-11-03 Felix Kuhnke , Lars Rumberg , Jörn Ostermann

Recognising expressive behaviours in face videos is a long-standing challenge in Affective Computing. Despite significant advancements in recent years, it still remains a challenge to build a robust and reliable system for naturalistic and…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Mani Kumar Tellamekala , Shashank Jaiswal , Thomas Smith , Timur Alamev , Gary McKeown , Anthony Brown , Michel Valstar

The detection and localization of highly realistic deepfake audio-visual content are challenging even for the most advanced state-of-the-art methods. While most of the research efforts in this domain are focused on detecting high-quality…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Zhixi Cai , Shreya Ghosh , Aman Pankaj Adatia , Munawar Hayat , Abhinav Dhall , Tom Gedeon , Kalin Stefanov

Wearables equipped with pervasive sensors enable us to monitor physiological and behavioral signals in our everyday life. We propose the WellAff system able to recognize affective states for wellbeing support. It also includes health care…

Dynamic facial expression recognition (FER) databases provide important data support for affective computing and applications. However, most FER databases are annotated with several basic mutually exclusive emotional categories and contain…

计算机视觉与模式识别 · 计算机科学 2023-08-15 Yuanyuan Liu , Wei Dai , Chuanxu Feng , Wenbin Wang , Guanghao Yin , Jiabei Zeng , Shiguang Shan

Since its introduction in 2018, EPIC-KITCHENS has attracted attention as the largest egocentric video benchmark, offering a unique viewpoint on people's interaction with objects, their attention, and even intention. In this paper, we detail…

This paper investigates the RWKV model's efficacy in content moderation through targeted experimentation. We introduce a novel dataset specifically designed for distillation into smaller models, enhancing content moderation practices. This…

计算与语言 · 计算机科学 2024-09-09 Umut Yildirim , Rohan Dutta , Burak Yildirim , Atharva Vaidya

This paper introduces QAConv, a new question answering (QA) dataset that uses conversations as a knowledge source. We focus on informative conversations, including business emails, panel discussions, and work channels. Unlike open-domain…

计算与语言 · 计算机科学 2022-04-18 Chien-Sheng Wu , Andrea Madotto , Wenhao Liu , Pascale Fung , Caiming Xiong

Voice data is increasingly being used in modern digital communications, yet there is still a lack of comprehensive tools for automated voice analysis and characterization. To this end, we developed the VANPY (Voice Analysis in Python)…

声音 · 计算机科学 2025-05-06 Gregory Koushnir , Michael Fire , Galit Fuhrmann Alpert , Dima Kagan

Facial expression recognition (FER) in the wild is crucial for building reliable human-computer interactive systems. However, annotations of large scale datasets in FER has been a key challenge as these datasets suffer from noise due to…

计算机视觉与模式识别 · 计算机科学 2021-07-27 Darshan Gera , S Balasubramanian