English
Related papers

Related papers: Technical Report for Valence-Arousal Estimation in…

200 papers

Affective computing is confronted to high inter-subject variability, in both emotional and physiological responses to a given stimulus. In a stimuli-shared framework, that is to say for different subjects who watch the same stimuli,…

Neurons and Cognition · Quantitative Biology 2018-09-25 Ayoub Hajlaoui , Mohamed Chetouani , Slim Essid

The DimABSA task requires fine-grained sentiment intensity prediction for restaurant reviews, including scores for Valence and Arousal dimensions for each Aspect Term. In this study, we propose a Coarse-to-Fine In-context Learning(CFICL)…

Computation and Language · Computer Science 2024-12-30 Senbin Zhu , Hanjie Zhao , Xingren Wang , Shanhong Liu , Yuxiang Jia , Hongying Zan

We introduce Multimodal Matching based on Valence and Arousal (MMVA), a tri-modal encoder framework designed to capture emotional content across images, music, and musical captions. To support this framework, we expand the…

Sound · Computer Science 2025-11-21 Suhwan Choi , Kyu Won Kim , Myungjoo Kang

The WASSA 2017 EmoInt shared task has the goal to predict emotion intensity values of tweet messages. Given the text of a tweet and its emotion category (anger, joy, fear, and sadness), the participants were asked to build a system that…

Computation and Language · Computer Science 2020-03-17 Egor Lakomkin , Chandrakant Bothe , Stefan Wermter

Emotion recognition is inherently ambiguous, with uncertainty arising both from rater disagreement and from discrepancies across modalities such as speech and text. There is growing interest in modeling rater ambiguity using label…

Audio and Speech Processing · Electrical Eng. & Systems 2026-01-27 Jingyao Wu , Grace Lin , Yinuo Song , Rosalind Picard

The emotion recognition has attracted more attention in recent decades. Although significant progress has been made in the recognition technology of the seven basic emotions, existing methods are still hard to tackle compound emotion…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Sunan Li , Hailun Lian , Cheng Lu , Yan Zhao , Tianhua Qi , Hao Yang , Yuan Zong , Wenming Zheng

This paper presents a light-weight and accurate deep neural model for audiovisual emotion recognition. To design this model, the authors followed a philosophy of simplicity, drastically limiting the number of parameters to learn from the…

Artificial Intelligence · Computer Science 2018-08-09 Valentin Vielzeuf , Corentin Kervadec , Stéphane Pateux , Alexis Lechervy , Frédéric Jurie

Automatic emotion recognition (ER) has recently gained lot of interest due to its potential in many real-world applications. In this context, multimodal approaches have been shown to improve performance (over unimodal approaches) by…

Computer Vision and Pattern Recognition · Computer Science 2022-09-20 R Gnana Praveen , Eric Granger , Patrick Cardinal

We introduce EmoLoom-2B, a lightweight and reproducible pipeline that turns small language models under 2B parameters into fast screening candidates for joint emotion classification and Valence-Arousal-Dominance prediction. To ensure…

Computation and Language · Computer Science 2026-02-17 Zilin Li , Weiwei Xu , Xuanbo Lu , Zheda Liu

The ACM Multimedia 2023 Computational Paralinguistics Challenge addresses two different problems for the first time in a research competition under well-defined conditions: In the Emotion Share Sub-Challenge, a regression on speech has to…

As an extensive research in the field of natural language processing (NLP), aspect-based sentiment analysis (ABSA) is the task of predicting the sentiment expressed in a text relative to the corresponding aspect. Unfortunately, most…

Computation and Language · Computer Science 2023-01-10 Nankai Lin , Yingwen Fu , Xiaotian Lin , Aimin Yang , Shengyi Jiang

Speech emotion recognition (SER), particularly for naturally expressed emotions, remains a challenging computational task. Key challenges include the inherent subjectivity in emotion annotation and the imbalanced distribution of emotion…

Sound · Computer Science 2025-06-03 Tiantian Feng , Thanathai Lertpetchpun , Dani Byrd , Shrikanth Narayanan

Most existing text-to-image person retrieval methods usually assume that the training image-text pairs are perfectly aligned; however, the noisy correspondence(NC) issue (i.e., incorrect or unreliable alignment) exists due to poor image…

Computer Vision and Pattern Recognition · Computer Science 2025-02-11 Runqing Zhang , Xue Zhou

In recent times, there has been significant interest in the machine recognition of human emotions, due to the suite of applications to which this knowledge can be applied. A number of different modalities, such as speech or facial…

Human-Computer Interaction · Computer Science 2018-03-06 Jonny O'Dwyer , Ronan Flynn , Niall Murray

When recognizing emotions from speech, we encounter two common problems: how to optimally capture emotion-relevant information from the speech signal and how to best quantify or categorize the noisy subjective emotion labels.…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-04 Sofoklis Kakouros , Themos Stafylakis , Ladislav Mosner , Lukas Burget

Current embodied VLM evaluation relies on static, expert-defined, manually annotated benchmarks that exhibit severe redundancy and coverage imbalance. This labor intensive paradigm drains computational and annotation resources, inflates…

Computation and Language · Computer Science 2026-02-03 Shuai Zhang , Jiayu Hu , Zijie Chen , Zeyuan Ding , Yi Zhang , Yingji Zhang , Ziyi Zhou , Junwei Liao , Shengjie Zhou , Yong Dai , Zhenzhong Lan , Xiaozhu Ju

Ambivalence/hesitancy recognition in unconstrained videos is a challenging problem due to the subtle, multimodal, and context-dependent nature of this behavioral state. In this paper, a multimodal approach for video-level…

Computer Vision and Pattern Recognition · Computer Science 2026-03-16 Elena Ryumina , Alexandr Axyonov , Dmitry Sysoev , Timur Abdulkadirov , Kirill Almetov , Yulia Morozova , Dmitry Ryumin

In this study, we focus on continuous emotion recognition using body motion and speech signals to estimate Activation, Valence, and Dominance (AVD) attributes. Semi-End-To-End network architecture is proposed where both extracted features…

Human-Computer Interaction · Computer Science 2020-11-03 Berkay Köprü , Engin Erzin

We present a model to predict fine-grained emotions along the continuous dimensions of valence, arousal, and dominance (VAD) with a corpus with categorical emotion annotations. Our model is trained by minimizing the EMD (Earth Mover's…

Computation and Language · Computer Science 2021-09-13 Sungjoon Park , Jiseon Kim , Seonghyeon Ye , Jaeyeol Jeon , Hee Young Park , Alice Oh

Automated affective computing in the wild setting is a challenging problem in computer vision. Existing annotated databases of facial expressions in the wild are small and mostly cover discrete emotions (aka the categorical model). There…

Computer Vision and Pattern Recognition · Computer Science 2018-02-06 Ali Mollahosseini , Behzad Hasani , Mohammad H. Mahoor
‹ Prev 1 8 9 10 Next ›