English
Related papers

Related papers: Design and development a children's speech databas…

200 papers

This paper presents the "Ethiopian" system for the SLT 2021 Children Speech Recognition Challenge. Various data processing and augmentation techniques are proposed to tackle children's speech recognition problem, especially the lack of the…

Sound · Computer Science 2020-11-10 Guoguo Chen , Xingyu Na , Yongqing Wang , Zhiyong Yan , Junbo Zhang , Sifan Ma , Yujun Wang

Speech-comprehension difficulties are common among older people. Standard speech tests do not fully capture such difficulties because the tests poorly resemble the context-rich, story-like nature of ongoing conversation and are typically…

Computation and Language · Computer Science 2025-03-04 Björn Herrmann

Computational modeling of naturalistic conversations in clinical applications has seen growing interest in the past decade. An important use-case involves child-adult interactions within the autism diagnosis and intervention domain. In this…

Audio and Speech Processing · Electrical Eng. & Systems 2019-10-30 Nithin Rao Koluguri , Manoj Kumar , So Hyun Kim , Catherine Lord , Shrikanth Narayanan

We present an experimental dataset, Basic Dataset for Sorani Kurdish Automatic Speech Recognition (BD-4SK-ASR), which we used in the first attempt in developing an automatic speech recognition for Sorani Kurdish. The objective of the…

Computation and Language · Computer Science 2019-12-03 Akam Qader , Hossein Hassani

SpeechBrain is an open-source and all-in-one speech toolkit. It is designed to facilitate the research and development of neural speech processing technologies by being simple, flexible, user-friendly, and well-documented. This paper…

Spontaneous speech in the form of conversations, meetings, voice-mail, interviews, oral history, etc. is one of the most ubiquitous forms of human communication. Search engines providing access to such speech collections have the potential…

Human-Computer Interaction · Computer Science 2013-12-19 Donna Vakharia , Rachel Gibbs

Silent speech interfaces have been recently proposed as a way to enable communication when the acoustic signal is not available. This introduces the need to build visual speech recognition systems for silent and whispered speech. However,…

Computer Vision and Pattern Recognition · Computer Science 2018-02-20 Stavros Petridis , Jie Shen , Doruk Cetin , Maja Pantic

The goal of the BabyLM is to stimulate new research connections between cognitive modeling and language model pretraining. We invite contributions in this vein to the BabyLM Workshop, which will also include the 4th iteration of the BabyLM…

Speech foundation models have shown strong transferability across a wide range of speech applications. However, their robustness to age-related domain shift in speaker diarization remains underexplored. In this work, we present a…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-20 Anfeng Xu , Tiantian Feng , Shrikanth Narayanan

Modelling the process that a listener actuates in deriving the words intended by a speaker requires setting a hypothesis on how lexical items are stored in memory. This work aims at developing a system that imitates humans when identifying…

Audio and Speech Processing · Electrical Eng. & Systems 2021-07-07 Maria-Gabriella Di Benedetto , Stefanie Shattuck-Hufnagel , Jeung-Yoon Choi , Luca De Nardis , Javier Arango , Ian Chan , Alec DeCaprio

Emphasizing problem formulation in AI literacy activities with children is vital, yet we lack empirical studies on their structure and affordances. We propose that participatory design involving teachable machines facilitates problem…

Human-Computer Interaction · Computer Science 2024-03-01 Utkarsh Dwivedi , Salma Elsayed-Ali , Elizabeth Bonsignore , Hernisa Kacorri

In this paper, we identify challenges in children's current information retrieval process, and propose conversational robots as an opportunity to ease this process in a responsible way. Tools children currently use in this process, such as…

Information Retrieval · Computer Science 2021-06-16 T. Beelen , E. Velner , R. Ordelman , K. P. Truong , V. Evers , T. Huibers

Recent efforts in Spoken Dialogue Modeling aim to synthesize spoken dialogue without the need for direct transcription, thereby preserving the wealth of non-textual information inherent in speech. However, this approach faces a challenge…

Computation and Language · Computer Science 2024-07-03 Yu-Kuan Fu , Cheng-Kuang Lee , Hsiu-Hsuan Wang , Hung-yi Lee

In this paper, we present a database of emotional speech intended to be open-sourced and used for synthesis and generation purpose. It contains data for male and female actors in English and a male actor in French. The database covers 5…

Computation and Language · Computer Science 2018-06-26 Adaeze Adigwe , Noé Tits , Kevin El Haddad , Sarah Ostadabbas , Thierry Dutoit

We present the design of an online social skills development interface for teenagers with autism spectrum disorder (ASD). The interface is intended to enable private conversation practice anywhere, anytime using a web-browser. Users…

Reliable transcription of child-adult conversations in clinical settings is crucial for diagnosing developmental disorders like Autism. Recent advances in deep learning and availability of large scale transcribed data has led to development…

Audio and Speech Processing · Electrical Eng. & Systems 2025-08-15 Aditya Ashvin , Rimita Lahiri , Aditya Kommineni , Somer Bishop , Catherine Lord , Sudarsana Reddy Kadiri , Shrikanth Narayanan

Large Language Models (LLMs), predominantly trained on adult conversational data, face significant challenges when generating authentic, child-like dialogue for specialized applications. We present a comparative study evaluating five…

Computation and Language · Computer Science 2025-10-29 Syed Zohaib Hassan , Pål Halvorsen , Miriam S. Johnson , Pierre Lison

Here we study polysemy as a potential learning bias in vocabulary learning in children. Words of low polysemy could be preferred as they reduce the disambiguation effort for the listener. However, such preference could be a side-effect of…

Computation and Language · Computer Science 2020-09-24 Bernardino Casas , Neus Català , Ramon Ferrer-i-Cancho , Antoni Hernández-Fernández , Jaume Baixeries

Labeled audio data is insufficient to build satisfying speech recognition systems for most of the languages in the world. There have been some zero-resource methods trying to perform phoneme or word-level speech recognition without labeled…

Computation and Language · Computer Science 2025-01-14 Haoyu Wang , Wei-Qiang Zhang , Hongbin Suo , Yulong Wan

It has been suggested in developmental psychology literature that the communication of affect between mothers and their infants correlates with the socioemotional and cognitive development of infants. In this study, we obtained day-long…

Audio and Speech Processing · Electrical Eng. & Systems 2019-12-13 Xuewen Yao , Dong He , Tiancheng Jing , Kaya de Barbaro