中文
相关论文

相关论文: EigenEmo: Spectral Utterance Representation Using …

200 篇论文

Tensor decomposition is an important tool for multiway data analysis. In practice, the data is often sparse yet associated with rich temporal information. Existing methods, however, often under-use the time information and ignore the…

机器学习 · 计算机科学 2023-10-31 Zheng Wang , Shikai Fang , Shibo Li , Shandian Zhe

Microblog, an online-based broadcast medium, is a widely used forum for people to share their thoughts and opinions. Recently, Emotion Recognition (ER) from microblogs is an inspiring research topic in diverse areas. In the machine learning…

机器学习 · 计算机科学 2023-01-10 Md. Ahsan Habib , M. A. H. Akhand , Md. Abdus Samad Kamal

Emotion classification in text is typically performed with neural network models which learn to associate linguistic units with emotions. While this often leads to good predictive performance, it does only help to a limited degree to…

计算与语言 · 计算机科学 2022-05-17 Felix Casel , Amelie Heindl , Roman Klinger

Recent advancements in transformer-based speech representation models have greatly transformed speech processing. However, there has been limited research conducted on evaluating these models for speech emotion recognition (SER) across…

计算与语言 · 计算机科学 2023-08-21 Anant Singh , Akshat Gupta

Empirical mode decomposition (EMD) has developed into a prominent tool for adaptive, scale-based signal analysis in various fields like robotics, security and biomedical engineering. Since the dramatic increase in amount of data puts…

计算机视觉与模式识别 · 计算机科学 2021-06-30 Jin Zhang , Fan Feng , Pere Marti-Puig , Cesar F. Caiafa , Zhe Sun , Feng Duan , Jordi Solé-Casals

Major Depressive Disorder (MDD) is a highly prevalent mental health condition, and a deeper understanding of its neurocognitive foundations is essential for identifying how core functions such as emotional and self-referential processing…

Decoding visual experience from brain activity has advanced substantially, but cur- rent brain-to-text systems largely recover semantic content while discarding affect. Additionally, language models can generate emotional text when prompted…

机器学习 · 计算机科学 2026-05-19 Bilal A. Mohammed , Lin Gu , Ruogo Fang

In the domain of human-computer interaction, accurately recognizing and interpreting human emotions is crucial yet challenging due to the complexity and subtlety of emotional expressions. This study explores the potential for detecting a…

多媒体 · 计算机科学 2025-05-13 Jiehui Jia , Huan Zhang , Jinhua Liang

Decoding emotional states from human brain activity plays an important role in brain-computer interfaces. Existing emotion decoding methods still have two main limitations: one is only decoding a single emotion category from a brain…

信号处理 · 电气工程与系统科学 2022-11-07 Kaicheng Fu , Changde Du , Shengpei Wang , Huiguang He

Human emotional expression is inherently dynamic, complex, and fluid, characterized by smooth transitions in intensity throughout verbal communication. However, the modeling of such intensity fluctuations has been largely overlooked by…

声音 · 计算机科学 2024-10-01 Jingyi Xu , Hieu Le , Zhixin Shu , Yang Wang , Yi-Hsuan Tsai , Dimitris Samaras

Speech Emotion Recognition (SER) has been traditionally formulated as a classification task. However, emotions are generally a spectrum whose distribution varies from situation to situation leading to poor Out-of-Domain (OOD) performance.…

声音 · 计算机科学 2024-07-23 Hazim Bukhari , Soham Deshmukh , Hira Dhamyal , Bhiksha Raj , Rita Singh

Recent advancements in speech-language models have yielded significant improvements in speech tokenization and synthesis. However, effectively mapping the complex, multidimensional attributes of speech into discrete tokens remains…

Multimodal emotion recognition in conversation (MERC) refers to identifying and classifying human emotional states by combining data from multiple different modalities (e.g., audio, images, text, video, etc.). Most existing multimodal…

计算与语言 · 计算机科学 2025-08-13 Yuntao Shou , Tao Meng , Wei Ai , Keqin Li

Text is the major method that is used for communication now a days, each and every day lots of text are created. In this paper the text data is used for the classification of the emotions. Emotions are the way of expression of the persons…

计算与语言 · 计算机科学 2019-01-11 Naveenkumar K S , Vinayakumar R , Soman KP

In this work, we present a method which determines optimal multi-step dynamic mode decomposition (DMD) models via entropic regression, which is a nonlinear information flow detection algorithm. Motivated by the higher-order DMD (HODMD)…

机器学习 · 统计学 2024-06-19 Christopher W. Curtis , Erik Bollt , Daniel Jay Alford-Lago

Advancements in spoken language processing have driven the development of spoken language models (SLMs), designed to achieve universal audio understanding by jointly learning text and audio representations for a wide range of tasks.…

计算与语言 · 计算机科学 2025-10-31 Pedro Corrêa , João Lima , Victor Moreno , Lucas Ueda , Paula Dornhofer Paro Costa

Harmonic instability occurs frequently in the power electronic converter system. This paper leverages multi-resolution dynamic mode decomposition (MR-DMD) as a data-driven diagnostic tool for the system stability of power electronic…

信号处理 · 电气工程与系统科学 2024-04-16 Rui Kong , Subham Sahoo , Yongjie Liu , Frede Blaabjerg

Emotional and cognitive factors are essential for understanding mental health disorders. However, existing methods often treat multi-modal data as classification tasks, limiting interpretability especially for emotion and cognition.…

多媒体 · 计算机科学 2026-03-03 Zhiyuan Zhou , Yanrong Guo , Shijie Hao

PESTalk is a novel method for generating 3D facial animations with personalized emotional styles directly from speech. It overcomes key limitations of existing approaches by introducing a Dual-Stream Emotion Extractor (DSEE) that captures…

图形学 · 计算机科学 2025-12-08 Tianshun Han , Benjia Zhou , Ajian Liu , Yanyan Liang , Du Zhang , Zhen Lei , Jun Wan

Electromyography-to-Speech (ETS) conversion has demonstrated its potential for silent speech interfaces by generating audible speech from Electromyography (EMG) signals during silent articulations. ETS models usually consist of an EMG…

声音 · 计算机科学 2024-05-15 Zhao Ren , Kevin Scheck , Qinhan Hou , Stefano van Gogh , Michael Wand , Tanja Schultz
‹ 上一页 1 8 9 10 下一页 ›