Related papers: A study of vowel nasalization using instantaneous …
Recent research has shown that language and the socio-cognitive phenomena associated with it can be aptly modeled and visualized through networks of linguistic entities. However, most of the existing works on linguistic networks focus only…
Long-living coupled transverse and longitudinal phonon modes are explored in dense and regular arrangements of flat microfluidic droplets. The collective oscillations are driven by hydrodynamic interactions between the confined droplets and…
All Italian consonants affected by gemination, that is affricates, fricatives, liquids, nasals, and stops, were analyzed within a project named GEMMA that lasted over a span of about 25 years. Results of the analysis on stops, as published…
Automatic pronunciation evaluation plays an important role in pronunciation training and second language education. This field draws heavily on concepts from automatic speech recognition (ASR) to quantify how close the pronunciation of…
Encouraged by the success of deep neural networks on a variety of visual tasks, much theoretical and experimental work has been aimed at understanding and interpreting how vision networks operate. Meanwhile, deep neural networks have also…
A deep neural network (DNN)-based model has been developed to predict non-parametric distributions of durations of phonemes in specified phonetic contexts and used to explore which factors influence durations most. Major factors in US…
Resonance trapping appears in open many-particle quantum systems at high level density when the coupling to the continuum of decay channels reaches a critical strength. Here a reorganization of the system takes place and a separation of…
Non-verbal Vocalizations (NVs), such as laughter and sighs, are vital for conveying emotion and intention in human speech, yet most existing speech systems neglect them, which severely compromises communicative richness and emotional…
Spectro-temporal dynamics of consonant-vowel (CV) transition regions are considered to provide robust cues related to articulation. In this work, we propose an objective measure of precise articulation, dubbed the objective articulation…
When two systems are coupled, the driver system can function as an external forcing over the driven or response system. Also, an external forcing can independently perturb the driven system, leading us to examine the interplay between the…
Pitch and Formant frequencies are important features in speech processing applications. The period of the vocal cord's output for vowels is known as the pitch or the fundamental frequency, and formant frequencies are essentially resonance…
Modelling of early language acquisition aims to understand how infants bootstrap their language skills. The modelling encompasses properties of the input data used for training the models, the cognitive hypotheses and their algorithmic…
We introduced a measurement procedure for the involuntary response of voice fundamental-frequency to frequency modulated auditory stimulation. This involuntary response plays an essential role in voice fundamental frequency control while…
Speaker Diarization is the problem of separating speakers in an audio. There could be any number of speakers and final result should state when speaker starts and ends. In this project, we analyze given audio file with 2 channels and 2…
Frequency locking in forced oscillatory systems typically occurs in 'V'-shaped domains in the plane spanned by the forcing frequency and amplitude, the so-called Arnol'd tongues. Here, we show that if the medium is spatially extended and…
Frequency locking to an external forcing frequency is a {well} known phenomenon. In the auditory system, it results in a localized traveling wave, the shape of which is essential for efficient discrimination between incoming frequencies. An…
The dynamical backaction from a periodically driven optical or microwave cavity can reduce the damping of a mechanical resonator, leading to parametric instability accompanied by self-sustained oscillations. Fundamentally, the driving…
The transduction process that occurs in the inner ear of the auditory system is a complex mechanism which requires a non-linear dynamical description. In addition to this, the stochastic phenomena that naturally arise in the inner ear…
The various speech sounds of a language are obtained by varying the shape and position of the articulators surrounding the vocal tract. Analyzing their variations is crucial for understanding speech production, diagnosing speech disorders…
Vocal fold (VF) motion is fundamental to voice production and diagnosis in speech and health sciences. The motion is a consequence of air flow interacting with elastic vocal fold structures. Motivated by existing lumped mass models and…