English
Related papers

Related papers: Uncovering Voice Misuse Using Symbolic Mismatch

200 papers

Voice disorders significantly impact patient quality of life, yet non-invasive automated diagnosis remains under-explored due to both the scarcity of pathological voice data, and the variability in recording sources. This work introduces…

Digital biomarkers for depression have largely relied on static acoustic descriptors, pooled summary statistics, or conventional machine learning representations. Such approaches may miss nonlinear temporal organization embedded in…

Sound · Computer Science 2026-04-30 Himadri S Samanta

We present a series of two studies conducted to understand user's affective states during voice-based human-machine interactions. Emphasis is placed on the cases of communication errors or failures. In particular, we are interested in…

Human-Computer Interaction · Computer Science 2022-07-19 Sujeong Kim , Abhinav Garlapati , Jonah Lubin , Amir Tamrakar , Ajay Divakaran

Noise pollution is one of the topmost quality of life issues for urban residents in the United States. Continued exposure to high levels of noise has proven effects on health, including acute effects such as sleep disruption, and long-term…

How much audio is needed to fully observe a multilingual ASR model's learned sub-token inventory across languages, and does data disparity in multilingual pre-training affect how these tokens are utilized during inference? We address this…

Computation and Language · Computer Science 2025-10-28 Siyu Liang , Nicolas Ballier , Gina-Anne Levow , Richard Wright

The performance of speaker verification systems is adversely affected by speaker aging. However, due to challenges in data collection, particularly the lack of sustained and large-scale longitudinal data for individuals, research on speaker…

Sound · Computer Science 2025-05-28 Zhiqi Ai , Meixuan Bao , Zhiyong Chen , Zhi Yang , Xinnuo Li , Shugong Xu

Speech sound disorder (SSD) refers to a type of developmental disorder in young children who encounter persistent difficulties in producing certain speech sounds at the expected age. Consonant errors are the major indicator of SSD in…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-17 Si-Ioi Ng , Cymie Wing-Yee Ng , Jingyu Li , Tan Lee

In the U.S., approximately 15-17% of children 2-8 years of age are estimated to have at least one diagnosed mental, behavioral or developmental disorder. However, such disorders often go undiagnosed, and the ability to evaluate and treat…

Audio and Speech Processing · Electrical Eng. & Systems 2022-03-30 Jialu Li , Mark Hasegawa-Johnson , Nancy L. McElwain

The phenomenon of sound symbolism, the non-arbitrary mapping between word sounds and meanings, has long been demonstrated through anecdotal experiments like Bouba Kiki, but rarely tested at scale. We present the first computational…

Computation and Language · Computer Science 2025-12-16 Anika Sharma , Tianyi Niu , Emma Wrenn , Shashank Srivastava

The DIarization and Speech Processing for LAnguage understanding in Conversational Environments - Medical (DISPLACE-M) challenge introduces a conversational AI benchmark for understanding goal-oriented, real-world medical dialogues. The…

In this study we developed an automated system that evaluates speech and language features from audio recordings of neuropsychological examinations of 92 subjects in the Framingham Heart Study. A total of 265 features were used in an…

Artificial Intelligence · Computer Science 2017-10-23 Tuka Alhanai , Rhoda Au , James Glass

We present a multilinear statistical model of the human tongue that captures anatomical and tongue pose related shape variations separately. The model is derived from 3D magnetic resonance imaging data of 11 speakers sustaining speech…

Computer Vision and Pattern Recognition · Computer Science 2018-04-18 Alexander Hewer , Stefanie Wuhrer , Ingmar Steiner , Korin Richmond

Malicious actors may seek to use different voice-spoofing attacks to fool ASV systems and even use them for spreading misinformation. Various countermeasures have been proposed to detect these spoofing attacks. Due to the extensive work…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-23 Awais Khan , Khalid Mahmood Malik , James Ryan , Mikul Saravanan

In this article, we introduce a novel problem of audio-visual autism behavior recognition, which includes social behavior recognition, an essential aspect previously omitted in AI-assisted autism screening research. We define the task at…

Languages have long been described according to their perceived rhythmic attributes. The associated typologies are of interest in psycholinguistics as they partly predict newborns' abilities to discriminate between languages and provide…

Audio and Speech Processing · Electrical Eng. & Systems 2024-01-29 François Deloche , Laurent Bonnasse-Gahot , Judit Gervain

Cognitive behavioural therapy is widely used to help patients understand and manage psychological distress. It is often delivered through spoken conversation, where therapists attend not only to what patients say, but also to how they say…

In this paper, we study the associations between human faces and voices. Audiovisual integration, specifically the integration of facial and vocal information is a well-researched area in neuroscience. It is shown that the overlapping…

Computer Vision and Pattern Recognition · Computer Science 2018-11-05 Changil Kim , Hijung Valentina Shin , Tae-Hyun Oh , Alexandre Kaspar , Mohamed Elgharib , Wojciech Matusik

Detecting duplicate patient participation in clinical trials is a major challenge because repeated patients can undermine the credibility and accuracy of the trial's findings and result in significant health and financial risks. Developing…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-23 Malikeh Ehghaghi , Marija Stanojevic , Ali Akram , Jekaterina Novikova

Depression is a mental disorder and can cause a variety of symptoms, including psychological, physical, and social. Speech has been proved an objective marker for the early recognition of depression. For this reason, many studies have been…

Machine Learning · Computer Science 2026-05-12 Loukas Ilias , Dimitris Askounis

Human activity and environment produces sounds such as, at home, the noise produced by water, cough, or television. These sounds can be used to determine the activity in the environment. The objective is to monitor a person's activity or…

Artificial Intelligence · Computer Science 2013-11-11 Serge Smidtas , Magalie Peyrot
‹ Prev 1 8 9 10 Next ›