Related papers: A study of vowel nasalization using instantaneous …
A body of literature has developed concerning "cloaking by anomalous localized resonance". The mathematical heart of the matter involves the behavior of a divergence-form elliptic equation in the plane, $\nabla\cdot (a(x)\nabla u(x)) =…
The effect of gravity and proper acceleration on the frequency spectrum of an optical resonator - both rigid or deformable - is considered in the framework of general relativity. The optical resonator is modeled either as a rod of matter…
The effect of electron-phonon coupling on the current noise in a molecular junction is investigated within a simple model. The model comprises a 1-level bridge representing a molecular level that connects between two free electron…
Converting input symbols to output audio in TTS requires modelling the durations of speech sounds. Leading non-autoregressive (NAR) TTS models treat duration modelling as a regression problem. The same utterance is then spoken with…
Excitons in twisted bilayers of transition metal dichalcogenides have strongly modified dispersion relations due to the formation of periodic moir\'e potentials. The strong coupling to a light field in an optical cavity leads to the…
Stochastic resonance is a counter-intuitive concept[1,2], ; the addition of noise to a noisy system induces coherent amplification of its response. First suggested as a mechanism for the cyclic recurrence of ice ages, stochastic resonance…
The phase time coupling effect on NMR relaxation is investigated based on coupled and uncoupled phase diffusion. The results indicate that phase and time coupling could significantly impact the NMR relaxation time. The spectral density term…
Diphthong vowels exhibit a degree of inherent dynamic change, the extent of which can vary synchronically and diachronically, such that diphthong vowels can become monophthongs and vice versa. Modelling this type of change requires defining…
Audio captioning is the task of automatically creating a textual description for the contents of a general audio signal. Typical audio captioning methods rely on deep neural networks (DNNs), where the target of the DNN is to map the input…
In this paper, we propose several methods that incorporate vocal tract length (VTL) warped features for spoken keyword spotting (KWS). The first method, VTL-independent KWS, involves training a single deep neural network (DNN) that utilizes…
Stochastic resonance is a phenomenon where a noise of appropriate intensity enhances the input signal strength. In this work, by employing the recently developed convex optimization methods in the context of dynamical systems and stochastic…
An overexpanded jet in a truncated ideally contoured nozzle is found to feature a tonal behavior. The flow field is investigated to understand its origin and show how it modifies side-load properties. The temporal and spatial organization…
Velopharyngeal dysfunction (VPD) is characterized by inadequate velopharyngeal closure during speech and often causes hypernasality and reduced intelligibility. Although speech-based machine learning models can perform well under…
A state-of-the-art 1D acoustic synthesizer has been previously developed, and coupled to speaker-specific biomechanical models of oropharynx in ArtiSynth. As expected, the formant frequencies of the synthesized vowel sounds were shown to be…
Repeated reading (RR) helps learners, who have little to no experience with reading fluently to gain confidence, speed and process words automatically. The benefits of repeated readings include helping all learners with fact recall, aiding…
The search for fractionalization in quantum spin liquids largely relies on their decoupling with the environment. However, the spin-lattice interaction is inevitable in a real setting. While the Majorana fermion evades a strong decay due to…
The great majority of current voice technology applications relies on acoustic features characterizing the vocal tract response, such as the widely used MFCC of LPC parameters. Nonetheless, the airflow passing through the vocal folds, and…
This manuscript is about further elucidating the concept of noising. The concept of noising first appeared in \cite{CVPR14}, in the context of curvature estimation and vertex localization on planar shapes. There are indications that noising…
The introduction of audio latent diffusion models possessing the ability to generate realistic sound clips on demand from a text description has the potential to revolutionize how we work with audio. In this work, we make an initial attempt…
Acoustic perturbations in a parallel relativistic flow of an inviscid fluid are considered. The general expression for the frequency of the sound waves in a uniformly (with zero shear) moving medium is derived. It is shown that relativity…