English

Data-driven Approach to Differentiating between Depression and Dementia from Noisy Speech and Language Data

Computation and Language 2022-10-10 v1 Machine Learning

Abstract

A significant number of studies apply acoustic and linguistic characteristics of human speech as prominent markers of dementia and depression. However, studies on discriminating depression from dementia are rare. Co-morbid depression is frequent in dementia and these clinical conditions share many overlapping symptoms, but the ability to distinguish between depression and dementia is essential as depression is often curable. In this work, we investigate the ability of clustering approaches in distinguishing between depression and dementia from human speech. We introduce a novel aggregated dataset, which combines narrative speech data from multiple conditions, i.e., Alzheimer's disease, mild cognitive impairment, healthy control, and depression. We compare linear and non-linear clustering approaches and show that non-linear clustering techniques distinguish better between distinct disease clusters. Our interpretability analysis shows that the main differentiating symptoms between dementia and depression are acoustic abnormality, repetitiveness (or circularity) of speech, word finding difficulty, coherence impairment, and differences in lexical complexity and richness.

Keywords

Cite

@article{arxiv.2210.03303,
  title  = {Data-driven Approach to Differentiating between Depression and Dementia from Noisy Speech and Language Data},
  author = {Malikeh Ehghaghi and Frank Rudzicz and Jekaterina Novikova},
  journal= {arXiv preprint arXiv:2210.03303},
  year   = {2022}
}

Comments

W-NUT at COLING 2022

R2 v1 2026-06-28T02:58:35.569Z