Related papers: A study of vowel nasalization using instantaneous …
Recent years have seen an explosion in the availability of Voice User Interfaces. However, user surveys suggest that there are issues with respect to usability, and it has been hypothesised that contemporary voice-enabled systems are…
We report a statistical analysis about people agglomeration soundscape. Specifically, we investigate the normalized sound amplitudes and intensities that emerge from people collective meetings. Our findings support the existence of…
We adopt the concept of the correlation matrix to study correlations among sequences of time-extended events occuring repeatedly at consecutive time-intervals. As an application we analyse the magnetoencephalography recordings obtained from…
We study the control of noise-induced spatio-temporal current density patterns in a semiconductor nanostructure (double barrier resonant tunnelling diode) by multiple time-delayed feedback. We find much more pronounced resonant features of…
While speaking at different rates, articulators (like tongue, lips) tend to move differently and the enunciations are also of different durations. In the past, affine transformation and DNN have been used to transform articulatory movements…
Stochastic resonance (SR) is a coherence enhancement effect due to noise that occurs in periodically-driven nonlinear dynamical systems. A very broad range of physical and biological systems present this effect such as climate change,…
Despite significant advances in ASR, the specific acoustic cues models rely on remain unclear. Prior studies have examined such cues on a limited set of phonemes and outdated models. In this work, we apply a feature attribution technique to…
Speaker clustering is an essential step in conventional speaker diarization systems and is typically addressed as an audio-only speech processing task. The language used by the participants in a conversation, however, carries additional…
The dynamics of an ensemble of bistable elements with global time-delayed coupling under the influence of noise is studied analytically and numerically. Depending on the noise level the system undergoes ordering transitions and demonstrates…
Various hand-crafted features representations of bio-signals rely primarily on the amplitude or power of the signal in specific frequency bands. The phase component is often discarded as it is more sample specific, and thus more sensitive…
In this study, we propose the global context guided channel and time-frequency transformations to model the long-range, non-local time-frequency dependencies and channel variances in speaker representations. We use the global context…
Vocal wow is a slow 0.2 - 3 Hz modulation of the voice that may be distinguished from the 4 - 7 Hz modulation of tremor or vibrato. We use a simple model of laryngeal muscle activation, mediated by time-delayed auditory feedback, to show…
Our environment is filled with rich and dynamic acoustic information. When we walk into a cathedral, the reverberations as much as appearance inform us of the sanctuary's wide open space. Similarly, as an object moves around us, we expect…
Positive feedback and cooperativity in the regulation of gene expression are generally considered to be necessary for obtaining bistable expression states. Recently, a novel mechanism of bistability termed emergent bistability has been…
The resonant interaction of electrically excited travelling surface acoustic waves and magnetization has been hitherto probed through the acoustic component. In this work it is investigated using time-resolved magneto-optical detection of…
Verbal communication transmits information across diverse linguistic levels, with neural synchronization (NS) between speakers and listeners emerging as a putative mechanism underlying successful exchange. However, the specific speech…
Most studies on speaker verification systems focus on long-duration utterances, which are composed of sufficient phonetic information. However, the performances of these systems are known to degrade when short-duration utterances are…
As circuits continue to miniaturize, noise has become a significant obstacle to performance optimization. Stochastic resonance in logic circuits offers an innovative approach to harness noise constructively; however, current implementations…
The resonant acousto-optic effect is studied both analytically and numerically in the terahertz range where the transverse-optical (TO) phonons play the role of a mediator which strongly couples the ultrasound and light fields. A…
The sound of our speech is influenced by the places we come from. Great Britain contains a wide variety of distinctive accents which are of interest to linguistics. In particular, the "a" vowel in words like "class" is pronounced…