Related papers: Comparison of aerosol emissions during specific sp…
During the phase of landing, an important aircraft-noise source emanates from the interaction of the landing-gear wake with the deployed flap. In the present work we cast this problem in an academic framework, by studying a simplified…
Speaker identification in noisy audio recordings, specifically those from collaborative learning environments, can be extremely challenging. There is a need to identify individual students talking in small groups from other students talking…
The velocity relaxation of an impulsively forced spherical particle in a fluid confined by two parallel plane walls is studied using a direct numerical simulation approach. During the relaxation process, the momentum of the particle is…
Which phonemes convey more speaker traits is a long-standing question, and various perception experiments were conducted with human subjects. For speaker recognition, studies were conducted with the conventional statistical models and the…
The COVID-19 pandemic has heightened the urgency to understand and prevent pathogen transmission, specifically regarding infectious airborne particles. Extensive studies validate the understanding of larger (droplets) and smaller (aerosols)…
Some speech recognition tasks, such as automatic speech recognition (ASR), are approaching or have reached human performance in many reported metrics. Yet, they continue to struggle in complex, real-world, situations, such as with distanced…
Because the performance of speech separation is excellent for speech in which two speakers completely overlap, research attention has been shifted to dealing with more realistic scenarios. However, domain mismatch between training/test…
Background: Pulmonary auscultation is a common tool for diagnosing various respiratory diseases. Previous studies have documented many details of pulmonary sounds in humans. However, information on sound generation and pressure loss inside…
Our aural experience plays an integral role in the perception and memory of the events in our lives. Some of the sounds we encounter throughout the day stay lodged in our minds more easily than others; these, in turn, may serve as powerful…
The process of human speech production involves coordinated respiratory action to elicit acoustic speech signals. Typically, speech is produced when air is forced from the lungs and is modulated by the vocal tract, where such actions are…
This study explores prosodic production in latent aphasia, a mild form of aphasia associated with left-hemisphere brain damage (e.g. stroke). Unlike prior research on moderate to severe aphasia, we investigated latent aphasia, which can…
Speaker diarization accuracy can be affected by both acoustics and conversation characteristics. Determining the cause of diarization errors is difficult because speaker voice acoustics and conversation structure co-vary, and the…
The COVID-19 pandemic has resulted in more than 125 million infections and more than 2.7 million casualties. In this paper, we attempt to classify covid vs non-covid cough sounds using signal processing and deep learning methods. Air…
Human speech production encompasses physiological processes that naturally react to physic stress. Stress caused by physical activity (PA), e.g., running, may lead to significant changes in a person's speech. The major changes are related…
Understanding the lip movement and inferring the speech from it is notoriously difficult for the common person. The task of accurate lip-reading gets help from various cues of the speaker and its contextual or environmental setting. Every…
The research is aimed at creating a method for rheological testing of viscoelastic fluids, the droplets of which, when stretched, form thinning filaments, i.e. exhibit the property of spinning. The typical example of such fluid is oral…
A noise source model, consisting of a pulse sequence at random times with memory, is presented. By varying the memory we can obtain variable randomness of the stochastic process. The delay time between pulses, i. e. the noise memory,…
The average predictability (aka informativity) of a word in context has been shown to condition word duration (Seyfarth, 2014). All else being equal, words that tend to occur in more predictable environments are shorter than words that tend…
In this study, the flow field around face masks was visualized and evaluated using computational fluid dynamics. The protective efficiency of face masks suppressing droplet infection owing to differences in the shape, medium, and doubling…
Acoustics-to-word models are end-to-end speech recognizers that use words as targets without relying on pronunciation dictionaries or graphemes. These models are notoriously difficult to train due to the lack of linguistic knowledge. It is…