Related papers: A study of vowel nasalization using instantaneous …
Pre-aspiration is defined as the period of glottal friction occurring in sequences of vocalic/consonantal sonorants and phonetically voiceless obstruents. We propose two machine learning methods for automatic measurement of pre-aspiration…
Using molecular dynamics simulation, we study acoustic resonance in low-temperature glass by applying a small periodic shear at a boundary wall. Shear wave resonance occurs as the frequency $\omega$ approaches $\omega_\ell= \pi…
Vocal tract configurations play a vital role in generating distinguishable speech sounds, by modulating the airflow and creating different resonant cavities in speech production. They contain abundant information that can be utilized to…
We analyzed the auditory-perceptual space across a substantial portion of the human vocal range (220-1046 Hz) using multidimensional scaling analysis of cochlea-scaled spectra from 250-ms vowel segments, initially studied in Friedrichs et…
It is increasingly considered that human speech perception and production both rely on articulatory representations. In this paper, we investigate whether this type of representation could improve the performances of a deep generative model…
Stochastic resonance is a general phenomenon usually observed in one-dimensional, amplitude modulated, bistable systems.We show experimentally the emergence of phase stochastic resonance in the bidimensional response of a forced…
Voice Onset Time (VOT), a key measurement of speech for basic research and applied medical studies, is the time between the onset of a stop burst and the onset of voicing. When the voicing onset precedes burst onset the VOT is negative; if…
Audio-Visual Speech Recognition (AVSR) seeks to model, and thereby exploit, the dynamic relationship between a human voice and the corresponding mouth movements. A recently proposed multimodal fusion strategy, AV Align, based on…
We study the spectrum and entanglement of phonons produced by temporal changes in homogeneous one-dimensional atomic condensates. To characterize the experimentally accessible changes, we first consider the dynamics of the condensate when…
An inversion of the speech polarity may have a dramatic detrimental effect on the performance of various techniques of speech processing. An automatic method for determining the speech polarity (which is dependent upon the recording setup)…
This thesis investigates acoustic properties of the vocal tract. Starting from a historical background (to name a few: Galen, Ibn Sina/Avicenna, Mersenne, Hooke, Euler, Kempelen, Abbe Mical, Kratzenstein, Wheatstone, Helmholtz, Riesz,…
Speaker diarization systems are challenged by a trade-off between the temporal resolution and the fidelity of the speaker representation. By obtaining a superior temporal resolution with an enhanced accuracy, a multi-scale approach is a way…
The Generation and propagation of the human voice is studied in two-dimensions using a full-body domain, using direct numerical simulation. The fluid/air in the vocal tract is modeled as a compressible and viscous fluid interacting with the…
Konkani is a highly nasalised language which makes it unique among Indo-Aryan languages. This work investigates the acoustic-phonetic properties of Konkani oral and nasal vowels. For this study, speech samples from six speakers (3 male and…
Airflow through the nasal cavity exhibits a wide variety of fluid dynamicsbehaviour due to the intricacy of the nasal geometry. The flow is naturallyunsteady and perhaps turbulent, despite CFD in the literature that assumesa steady laminar…
Coupling, synchronization, and non-linear dynamics of resonator modes are omnipresent in nature and highly relevant for a multitude of applications ranging from lasers to Josephson arrays and spin torque oscillators. Nanomechanical…
Any-to-any voice conversion technologies convert the vocal timbre of an utterance to any speaker even unseen during training. Although there have been several state-of-the-art any-to-any voice conversion models, they were all based on clean…
We use a multifractal formalism to study the effect of stochastic resonance in a noisy bistable system driven by various input signals. To characterize the response of a stochastic bistable system we introduce a new measure based on the…
This paper reports the phenomenon of resonance weakening and streaming onset in two phase acoustofluidics by performing numerical simulations of a capillary droplet suspended in a microfluidic chamber. The simulations show that depending on…
Stochastic resonance is a well established phenomenon, which proves relevant for a wide range of applications, of broad trans-disciplinary breath. Consider a one dimensional bistable stochastic system, characterized by a deterministic…