English
Related papers

Related papers: Speech based Depression Severity Level Classificat…

200 papers

Linguistic features have shown promising applications for detecting various cognitive impairments. To improve detection accuracies, increasing the amount of data or the number of linguistic features have been two applicable approaches.…

Computation and Language · Computer Science 2019-03-29 Zining Zhu , Jekaterina Novikova , Frank Rudzicz

We present an approach to learn a dense pixel-wise labeling from image-level tags. Each image-level tag imposes constraints on the output labeling of a Convolutional Neural Network (CNN) classifier. We propose Constrained CNN (CCNN), a…

Computer Vision and Pattern Recognition · Computer Science 2015-10-20 Deepak Pathak , Philipp Krähenbühl , Trevor Darrell

Machine learning models for speech-based depression classification offer promise for health care applications. Despite growing work on depression classification, little is understood about how the length of speech-input impacts model…

Computation and Language · Computer Science 2025-01-03 Tomasz Rutowski , Amir Harati , Yang Lu , Elizabeth Shriberg

Traditional screening practices for anxiety and depression pose an impediment to monitoring and treating these conditions effectively. However, recent advances in NLP and speech modelling allow textual, acoustic, and hand-crafted…

Sound · Computer Science 2023-01-02 Brian Diep , Marija Stanojevic , Jekaterina Novikova

Respiratory diseases remain major global health challenges, and traditional auscultation is often limited by subjectivity, environmental noise, and inter-clinician variability. This study presents an explainable multimodal deep learning…

Sound · Computer Science 2025-12-02 S M Asiful Islam Saky , Md Rashidul Islam , Md Saiful Arefin , Shahaba Alam

Deep neural networks have recently been shown to achieve highly competitive performance in many computer vision tasks due to their abilities of exploring in a much larger hypothesis space. However, since most deep architectures like stacked…

Computation and Language · Computer Science 2018-02-06 Zixiang Ding , Rui Xia , Jianfei Yu , Xiang Li , Jian Yang

In this work, we propose a training algorithm for an audio-visual automatic speech recognition (AV-ASR) system using deep recurrent neural network (RNN).First, we train a deep RNN acoustic model with a Connectionist Temporal Classification…

Computer Vision and Pattern Recognition · Computer Science 2016-11-10 Abhinav Thanda , Shankar M Venkatesan

Acoustic Scene Classification (ASC) is a challenging task, as a single scene may involve multiple events that contain complex sound patterns. For example, a cooking scene may contain several sound sources including silverware clinking,…

Audio and Speech Processing · Electrical Eng. & Systems 2019-09-20 Weimin Wang , Weiran Wang , Ming Sun , Chao Wang

Depression, a prevalent and serious mental health issue, affects approximately 3.8\% of the global population. Despite the existence of effective treatments, over 75\% of individuals in low- and middle-income countries remain untreated,…

Computation and Language · Computer Science 2024-07-19 Shengjie Li , Yinhao Xiao

We present a comprehensive study of deep bidirectional long short-term memory (LSTM) recurrent neural network (RNN) based acoustic models for automatic speech recognition (ASR). We study the effect of size and depth and train models of up…

Neural and Evolutionary Computing · Computer Science 2019-08-06 Albert Zeyer , Patrick Doetsch , Paul Voigtlaender , Ralf Schlüter , Hermann Ney

Previous studies have shown the correlation between sensor data collected from mobile phones and human depression states. Compared to the traditional self-assessment questionnaires, the passive data collected from mobile phones is easier to…

Acoustic-to-Word recognition provides a straightforward solution to end-to-end speech recognition without needing external decoding, language model re-scoring or lexicon. While character-based models offer a natural solution to the…

Audio and Speech Processing · Electrical Eng. & Systems 2018-08-22 Shruti Palaskar , Florian Metze

Dementia, a prevalent neurodegenerative condition, is a major manifestation of Alzheimer's disease (AD). As the condition progresses from mild to severe, it significantly impairs the individual's ability to perform daily tasks…

Machine Learning · Computer Science 2023-11-06 Md Gulzar Hussain , Ye Shiren

Conventionally, the manner of articulations in speech signal are derived using discriminative signal processing techniques or deep learning approaches. However, training such complex systems involves feature extraction, phoneme force…

Audio and Speech Processing · Electrical Eng. & Systems 2018-11-06 Pradeep R , Sreenivasa Rao K

Major depressive disorder (MDD) is a complex psychiatric disorder that affects the lives of hundreds of millions of individuals around the globe. Even today, researchers debate if morphological alterations in the brain are linked to MDD,…

Quantitative Methods · Quantitative Biology 2025-01-27 Roberto Goya-Maldonado , Tracy Erwin-Grabner , Ling-Li Zeng , Christopher R. K. Ching , Andre Aleman , Alyssa R. Amod , Zeynep Basgoze , Francesco Benedetti , Bianca Besteher , Katharina Brosch , Robin Bülow , Romain Colle , Colm G. Connolly , Emmanuelle Corruble , Baptiste Couvy-Duchesne , Kathryn Cullen , Udo Dannlowski , Christopher G. Davey , Annemiek Dols , Jan Ernsting , Jennifer W. Evans , Lukas Fisch , Paola Fuentes-Claramonte , Ali Saffet Gonul , Ian H. Gotlib , Hans J. Grabe , Nynke A. Groenewold , Dominik Grotegerd , Tim Hahn , J. Paul Hamilton , Laura K. M. Han , Ben J. Harrison , Tiffany C. Ho , Neda Jahanshad , Alec J. Jamieson , Andriana Karuk , Tilo Kircher , Bonnie Klimes-Dougan , Sheri-Michelle Koopowitz , Thomas Lancaster , Ramona Leenings , Meng Li , David E. J. Linden , Frank P. MacMaster , David M. A. Mehler , Susanne Meinert , Elisa Melloni , Bryon A. Mueller , Benson Mwangi , Igor Nenadić , Amar Ojha , Yasumasa Okamoto , Mardien L. Oudega , Brenda W. J. H. Penninx , Sara Poletti , Edith Pomarol-Clotet , Maria J. Portella , Elena Pozzi , Joaquim Radua , Elena Rodríguez-Cano , Matthew D. Sacchet , Raymond Salvador , Anouk Schrantee , Kang Sim , Jair C. Soares , Aleix Solanes , Dan J. Stein , Frederike Stein , Aleks Stolicyn , Sophia I. Thomopoulos , Yara J. Toenders , Aslihan Uyar-Demir , Eduard Vieta , Yolanda Vives-Gilabert , Henry Völzke , Martin Walter , Heather C. Whalley , Sarah Whittle , Nils Winter , Katharina Wittfeld , Margaret J. Wright , Mon-Ju Wu , Tony T. Yang , Carlos Zarate , Dick J. Veltman , Lianne Schmaal , Paul M. Thompson

Major Depressive Disorder is one of the leading causes of disability worldwide, yet its diagnosis still depends largely on subjective clinical assessments. Integrating Artificial Intelligence (AI) holds promise for developing objective,…

Artificial Intelligence · Computer Science 2026-05-01 Dorsa Macky Aleagha , Payam Zohari , Mostafa Haghir Chehreghani

This paper explores the use of multi-view features and their discriminative transforms in a convolutional deep neural network (CNN) architecture for a continuous large vocabulary speech recognition task. Mel-filterbank energies and…

Computation and Language · Computer Science 2018-02-19 Vikramjit Mitra , Wen Wang , Chris Bartels , Horacio Franco , Dimitra Vergyri

Automatic assessment of dysarthric speech is essential for sustained treatments and rehabilitation. However, obtaining atypical speech is challenging, often leading to data scarcity issues. To tackle the problem, we propose a novel…

Computation and Language · Computer Science 2023-05-01 Eun Jung Yeo , Kwanghee Choi , Sunhee Kim , Minhwa Chung

As speech-interfaces are getting richer and widespread, speech emotion recognition promises more attractive applications. In the continuous emotion recognition (CER) problem, tracking changes across affective states is an important and…

Sound · Computer Science 2021-10-11 Berkay Kopru , Engin Erzin

This article proposes a robust brain-inspired audio feature extractor (RBA-FE) model for depression diagnosis, using an improved hierarchical network architecture. Most deep learning models achieve state-of-the-art performance for…

Sound · Computer Science 2025-06-10 Yu-Xuan Wu , Ziyan Huang , Bin Hu , Zhi-Hong Guan