English
Related papers

Related papers: Speech-Based Estimation of Schizophrenia Severity …

200 papers

This paper presents a novel multimodal framework to distinguish between different symptom classes of subjects in the schizophrenia spectrum and healthy controls using audio, video, and text modalities. We implemented Convolution Neural…

Audio and Speech Processing · Electrical Eng. & Systems 2024-09-17 Gowtham Premananth , Yashish M. Siriwardena , Philip Resnik , Sonia Bansal , Deanna L. Kelly , Carol Espy-Wilson

Millions of people suffer from mental health conditions, yet many remain undiagnosed or receive delayed care due to limited clinical resources and labor-intensive assessment methods. While most machine-assisted approaches focus on…

Audio and Speech Processing · Electrical Eng. & Systems 2025-11-06 Gowtham Premananth , Philip Resnik , Sonia Bansal , Deanna L. Kelly , Carol Espy-Wilson

This study focuses on how different modalities of human communication can be used to distinguish between healthy controls and subjects with schizophrenia who exhibit strong positive symptoms. We developed a multi-modal schizophrenia…

Signal Processing · Electrical Eng. & Systems 2024-04-22 Gowtham Premananth , Yashish M. Siriwardena , Philip Resnik , Carol Espy-Wilson

Multimodal schizophrenia assessment systems have gained traction over the last few years. This work introduces a schizophrenia assessment system to discern between prominent symptom classes of schizophrenia and predict an overall…

Audio and Speech Processing · Electrical Eng. & Systems 2024-11-19 Gowtham Premananth , Carol Espy-Wilson

Voice disorders negatively impact the quality of daily life in various ways. However, accurately recognizing the category of pathological features from raw audio remains a considerable challenge due to the limited dataset. A promising…

Sound · Computer Science 2024-10-08 Lipeng Shen , Yifan Xiong , Dongyue Guo , Wei Mo , Lingyu Yu , Hui Yang , Yi Lin

A speech emotion recognition algorithm based on multi-feature and Multi-lingual fusion is proposed in order to resolve low recognition accuracy caused by lack of large speech dataset and low robustness of acoustic features in the…

Computation and Language · Computer Science 2020-01-17 Chunyi Wang

Studies on schizophrenia assessments using deep learning typically treat it as a classification task to detect the presence or absence of the disorder, oversimplifying the condition and reducing its clinical applicability. This traditional…

Audio and Speech Processing · Electrical Eng. & Systems 2025-09-29 Gowtham Premananth , Philip Resnik , Sonia Bansal , Deanna L. Kelly , Carol Espy-Wilson

The 1st SpeechWellness Challenge conveys the need for speech-based suicide risk assessment in adolescents. This study investigates a multimodal approach for this challenge, integrating automatic transcription with WhisperX, linguistic…

Computation and Language · Computer Science 2025-05-27 Ambre Marie , Ilias Maoudj , Guillaume Dardenne , Gwenolé Quellec

Advances in artificial intelligence (AI) and deep learning have improved diagnostic capabilities in healthcare, yet limited interpretability continues to hinder clinical adoption. Schizophrenia, a complex disorder with diverse symptoms…

Audio and Speech Processing · Electrical Eng. & Systems 2025-11-06 Gowtham Premananth , Carol Espy-Wilson

Schizophrenia is a debilitating, chronic mental disorder that significantly impacts an individual's cognitive abilities, behavior, and social interactions. It is characterized by subtle morphological changes in the brain, particularly in…

Computer Vision and Pattern Recognition · Computer Science 2024-09-04 Nagur Shareef Shaik , Teja Krishna Cherukuri , Vince Calhoun , Dong Hye Ye

We present our preliminary work to determine if patient's vocal acoustic, linguistic, and facial patterns could predict clinical ratings of depression severity, namely Patient Health Questionnaire depression scale (PHQ-8). We proposed a…

Computer Vision and Pattern Recognition · Computer Science 2017-12-01 Aven Samareh , Yan Jin , Zhangyang Wang , Xiangyu Chang , Shuai Huang

Both functional and structural magnetic resonance imaging (fMRI and sMRI) are widely used for the diagnosis of mental disorder. However, combining complementary information from these two modalities is challenging due to their…

Image and Video Processing · Electrical Eng. & Systems 2024-04-02 Ziyu Zhou , Anton Orlichenko , Gang Qu , Zening Fu , Vince D Calhoun , Zhengming Ding , Yu-Ping Wang

Key features of mental illnesses are reflected in speech. Our research focuses on designing a multimodal deep learning structure that automatically extracts salient features from recorded speech samples for predicting various mental…

Machine Learning · Computer Science 2020-04-15 Habibeh Naderi , Behrouz Haji Soleimani , Stan Matwin

Robust speech recognition is a key prerequisite for semantic feature extraction in automatic aphasic speech analysis. However, standard one-size-fits-all automatic speech recognition models perform poorly when applied to aphasic speech. One…

Audio and Speech Processing · Electrical Eng. & Systems 2023-03-21 Matthew Perez , Zakaria Aldeneh , Emily Mower Provost

Schizophrenia is a severe yet treatable mental disorder, it is diagnosed using a multitude of primary and secondary symptoms. Diagnosis and treatment for each individual depends on the severity of the symptoms, therefore there is a need for…

Human-Computer Interaction · Computer Science 2023-10-26 Niki Maria Foteinopoulou , Ioannis Patras

Cognitive impairment detection through spontaneous speech is a promising avenue for early diagnosis of Alzheimer's disease (AD) and mild cognitive impairment (MCI), where timely intervention can significantly improve patient outcomes. The…

Sound · Computer Science 2025-02-19 Yifan Gao , Long Guo , Hong Liu

When the parameters of Bayesian Short-time Spectral Amplitude (STSA) estimator for speech enhancement are selected based on the characteristics of the human auditory system, the gain function of the estimator becomes more flexible. Although…

Sound · Computer Science 2025-12-18 Suman Samui

We present two multimodal fusion-based deep learning models that consume ASR transcribed speech and acoustic data simultaneously to classify whether a speaker in a structured diagnostic task has Alzheimer's Disease and to what degree,…

Computation and Language · Computer Science 2021-07-01 Morteza Rohanian , Julian Hough , Matthew Purver

Alzheimer's disease (AD) is a progressive neurodegenerative disorder and the leading cause of dementia, affecting memory, reasoning, communication, and daily functioning. Early diagnosis is particularly important, as timely intervention may…

Sound · Computer Science 2026-05-26 Loukas Ilias , Dimitris Askounis

Speech emotion recognition is crucial in human-computer interaction, but extracting and using emotional cues from audio poses challenges. This paper introduces MFHCA, a novel method for Speech Emotion Recognition using Multi-Spatial Fusion…

Sound · Computer Science 2024-04-23 Xinxin Jiao , Liejun Wang , Yinfeng Yu
‹ Prev 1 2 3 10 Next ›