Related papers: Modal locking between vocal fold and vocal tract o…
Speech requires programming the sequence of vocal gestures that produce the sounds of words. Here we explored the timing of this program by asking our participants to pronounce, as quickly as possible, a sequence of…
At present emotion extraction from speech is a very important issue due to its diverse applications. Hence, it becomes absolutely necessary to obtain models that take into consideration the speaking styles of a person, vocal tract…
Our objective is an audio-visual model for separating a single speaker from a mixture of sounds such as other speakers and background noise. Moreover, we wish to hear the speaker even when the visual cues are temporarily absent due to…
Nasalization of vowels is a phenomenon where oral and nasal tracts participate simultaneously for the production of speech. Acoustic coupling of oral and nasal tracts results in a complex production system, which is subjected to a…
In this work, we study the acoustic forces acting on particles due to sound scattering at the interface with an elastic substrate. Utilizing the Green's function formalism, we predict that excitation of leaking Rayleigh wave results in…
In Rapela (2016) we reported traveling waves (TWs) on electrocorticographic (ECoG) recordings from an epileptic subject over speech processing brain regions, while the subject rhythmically produced consonant-vowel syllables (CVSs). In…
As a first step towards a complete computational model of speech learning involving perception-production loops, we investigate the forward mapping between pseudo-motor commands and articulatory trajectories. Two phonological feature sets,…
According to the acoustic fluidization hypothesis, elastic waves at a characteristic frequency form inside seismic faults even in the absence of an external perturbation. These waves are able to generate a normal stress which contrasts the…
While there has been significant progress towards modelling coherence in written discourse, the work in modelling spoken discourse coherence has been quite limited. Unlike the coherence in text, coherence in spoken discourse is also…
We propose an explanation for the onset of oscillations seen in numerical simulations of dense, inclined flows of inelastic, frictional spheres. It is based on a phase transition between disordered and ordered collisional states that may be…
Fluctuations and noise may alter the behavior of dynamical systems considerably. For example, oscillations may be sustained by demographic fluctuations in biological systems where a stable fixed point is found in the absence of noise. We…
Collective oscillation of cells in a population has been reported under diverse biological contexts and with vastly different molecular constructs. Could there be common principles similar to those that govern spontaneous oscillation in…
Many hearables contain an in-ear microphone, which may be used to capture the own voice of its user in noisy environments. Since the in-ear microphone mostly records body-conducted speech due to ear canal occlusion, it suffers from…
Human beings have developed fantastic abilities to integrate information from various sensory sources exploring their inherent complementarity. Perceptual capabilities are therefore heightened, enabling, for instance, the well-known…
While voice-based AI systems have achieved remarkable generative capabilities, their interactions often feel conversationally broken. This paper examines the interactional friction that emerges in modular Speech-to-Speech…
Synchronization of self-sustained oscillators under fixed-frequency and amplitude forcing is well understood, but how time-varying forcing mangles phase locking has been much less explored. Theory predicts that slow, deterministic…
The inverse problem of determining the cross-sectional area of a human vocal tract during the utterance of a vowel is considered in terms of the data consisting of the absolute value of sound pressure at the lips. If the upper lip is curved…
It is proposed that the theory of dynamical systems offers appropriate tools to model many phonological aspects of both speech production and perception. A dynamic account of speech rhythm is shown to be useful for description of both…
The transduction process that occurs in the inner ear of the auditory system is a complex mechanism which requires a non-linear dynamical description. In addition to this, the stochastic phenomena that naturally arise in the inner ear…
This work unveils the enigmatic link between phonemes and facial features. Traditional studies on voice-face correlations typically involve using a long period of voice input, including generating face images from voices and reconstructing…