English
Related papers

Related papers: Affective Music Information Retrieval

200 papers

Continuous valence-arousal estimation in real-world environments is challenging due to inconsistent modality reliability and interaction-dependent variability in audio-visual signals. Existing approaches primarily focus on modeling temporal…

Multimedia · Computer Science 2026-03-13 Yubeen Lee , Sangeun Lee , Junyeop Cha , Eunil Park

Understanding emotions and expressions is a task of interest across multiple disciplines, especially for improving user experiences. Contrary to the common perception, it has been shown that emotions are not discrete entities but instead…

Computer Vision and Pattern Recognition · Computer Science 2024-04-24 Niklas Wagner , Felix Mätzler , Samed R. Vossberg , Helen Schneider , Svetlana Pavlitska , J. Marius Zöllner

In recent decades, neuroscientific and psychological research has traced direct relationships between taste and auditory perceptions. This article explores multimodal generative models capable of converting taste information into music,…

Sound · Computer Science 2025-09-01 Matteo Spanio , Massimiliano Zampini , Antonio Rodà , Franco Pierucci

Emotion Representation Mapping (ERM) has the goal to convert existing emotion ratings from one representation format into another one, e.g., mapping Valence-Arousal-Dominance annotations for words or sentences into Ekman's Basic Emotions…

Computation and Language · Computer Science 2018-06-26 Sven Buechel , Udo Hahn

Images shared online strongly influence emotions and public well-being. Understanding the emotions an image elicits is therefore vital for fostering healthier and more sustainable digital communities, especially during public crises. We…

Computer Vision and Pattern Recognition · Computer Science 2025-12-19 Fanhang Man , Xiaoyue Chen , Huandong Wang , Baining Zhao , Han Li , Xinlei Chen

This paper presents the computational details of our emotion model, EEGS, and also provides an overview of a three-stage validation methodology used for the evaluation of our model, which can also be applicable for other computational…

Artificial Intelligence · Computer Science 2020-11-06 Suman Ojha , Jonathan Vitale , Mary-Anne Williams

Recently, the representation of emotions in the Valence, Arousal and Dominance (VAD) space has drawn enough attention. However, the complex nature of emotions and the subjective biases in self-reported values of VAD make the emotion model…

Human-Computer Interaction · Computer Science 2024-01-17 Mohammad Asif , Noman Ali , Sudhakar Mishra , Anushka Dandawate , Uma Shanker Tiwary

In this paper we present a new approach for the generation of multi-instrument symbolic music driven by musical emotion. The principal novelty of our approach centres on conditioning a state-of-the-art transformer based on continuous-valued…

Audio and Speech Processing · Electrical Eng. & Systems 2022-05-10 Serkan Sulun , Matthew E. P. Davies , Paula Viana

We present Affect2MM, a learning method for time-series emotion prediction for multimedia content. Our goal is to automatically capture the varying emotions depicted by characters in real-life human-centric situations and behaviors. We use…

Computer Vision and Pattern Recognition · Computer Science 2021-03-12 Trisha Mittal , Puneet Mathur , Aniket Bera , Dinesh Manocha

Group Emotion Recognition (GER) aims to infer collective affect in social environments such as classrooms, crowds, and public events. Many existing approaches rely on explicit individual-level processing, including cropped faces, person…

Computer Vision and Pattern Recognition · Computer Science 2026-04-06 Anderson Augusma , Dominique Vaufreydaz , Fédérique Letué

Multimodal Large Language Models (MLLMs) excel in Open-Vocabulary (OV) emotion recognition but often neglect fine-grained acoustic modeling. Existing methods typically use global audio encoders, failing to capture subtle, local temporal…

Multimedia · Computer Science 2026-03-24 Liyun Zhang , Xuanmeng Sha , Shuqiong Wu , Fengkai Liu

Generative AI models for music and the arts in general are increasingly complex and hard to understand. The field of eXplainable AI (XAI) seeks to make complex and opaque AI models such as neural networks more understandable to people. One…

Sound · Computer Science 2024-02-06 Nick Bryan-Kinns , Bingyuan Zhang , Songyan Zhao , Berker Banar

Emotion recognition (ER) technology is an integral part for developing innovative applications such as drowsiness detection and health monitoring that plays a pivotal role in contemporary society. This study delves into ER using…

Human-Computer Interaction · Computer Science 2024-02-07 Haseeb ur Rahman Abbasi , Zeeshan Rashid , Muhammad Majid , Syed Muhammad Anwar

Cross-subject EEG-based emotion recognition (EER) remains challenging due to strong inter-subject variability, which induces substantial distribution shifts in EEG signals, as well as the high complexity of emotion-related neural…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Weiwei Wu , Yueyang Li , Yuhu Shi , Weiming Zeng , Lang Qin , Yang Yang , Ke Zhou , Zhiguo Zhang , Wai Ting Siok , Nizhuan Wang

In this paper, we propose a novel framework for recognizing both discrete and dimensional emotions. In our framework, deep features extracted from foundation models are used as robust acoustic and visual representations of raw video. Three…

Audio and Speech Processing · Electrical Eng. & Systems 2023-09-18 Haotian Wang , Yuxuan Xi , Hang Chen , Jun Du , Yan Song , Qing Wang , Hengshun Zhou , Chenxi Wang , Jiefeng Ma , Pengfei Hu , Ya Jiang , Shi Cheng , Jie Zhang , Yuzhe Weng

This paper outlines a machine learning-enabled speaker-centric Emotion AI approach capable of predicting audience-affective engagement and vocal attractiveness in asynchronous video-based learning, relying solely on speaker-side affective…

Human-Computer Interaction · Computer Science 2026-03-20 Hung-Yue Suen , Kuo-En Hung , Fan-Hsun Tseng

We propose MoodNet - A Deep Convolutional Neural Network based architecture to effectively predict the emotion associated with a piece of music given its audio and lyrical content.We evaluate different architectures consisting of varying…

Audio and Speech Processing · Electrical Eng. & Systems 2018-11-15 Aniruddha Bhattacharya , K. V. Kadambari

Generative models of expressive piano performance are usually assessed by comparing their predictions to a reference human performance. A generative algorithm is taken to be better than competing ones if it produces performances that are…

Electroencephalography (EEG)-based emotion recognition plays a critical role in affective computing and emerging decision-support systems, yet remains challenging due to high-dimensional, noisy, and subject-dependent signals. This study…

Machine Learning · Computer Science 2026-02-09 S M Rakib UI Karim , Wenyi Lu , Diponkor Bala , Rownak Ara Rasul , Sean Goggins

Electroencephalogram (EEG)-based emotion decoding can objectively quantify people's emotional state and has broad application prospects in human-computer interaction and early detection of emotional disorders. Recently emerging deep…

Human-Computer Interaction · Computer Science 2024-11-08 Xinke Shen , Runmin Gan , Kaixuan Wang , Shuyi Yang , Qingzhu Zhang , Quanying Liu , Dan Zhang , Sen Song