English
Related papers

Related papers: An acoustic glottal source for vocal tract physica…

200 papers

Optical vibration sensing enables recovering the scene sound directly from the surface vibration of nearby objects, turning everyday objects into ``visual microphones''. However, most prior methods had focused on capturing the vibrations of…

Computer Vision and Pattern Recognition · Computer Science 2026-04-30 Shai Bagon , Matan Kichler , Mark Sheinin

Granular materials are inherently heterogeneous, leading to challenges in formulating accurate models of sound propagation. In order to quantify acoustic responses in space and time, we perform experiments in a photoelastic granular…

Materials Science · Physics 2015-05-19 Eli T. Owens , Karen E. Daniels

We present a soft corrugated tube sensor designed to estimate strain in each half segment. When air flows through the tube, the internal corrugated cavities induce pressure oscillations that excite the tube's standing wave resonance mode,…

Robotics · Computer Science 2026-04-23 Michael Chun , Ananya Nukala , Tae Myung Huh

A Prompt-based Text-To-Speech model allows a user to control different aspects of speech, such as speaking rate and perceived gender, through natural language instruction. Although user-friendly, such approaches are on one hand constrained:…

Computation and Language · Computer Science 2025-07-14 Atli Sigurgeirsson , Simon King

Hypernasality is a common characteristic symptom across many motor-speech disorders. For voiced sounds, hypernasality introduces an additional resonance in the lower frequencies and, for unvoiced sounds, there is reduced articulatory…

Audio and Speech Processing · Electrical Eng. & Systems 2020-09-14 Michael Saxon , Ayush Tripathi , Yishan Jiao , Julie Liss , Visar Berisha

Human categorization of sound seems predominantly based on sound source properties. To estimate these source properties we propose a novel sound analysis method, which separates sound into different sonic textures: tones, pulses, and…

Sound · Computer Science 2017-05-16 Ronald A. J. van Elburg , Tjeerd C. Andringa

Prosody conveys rich emotional and semantic information of the speech signal as well as individual idiosyncrasies. We propose a stand-alone model that maps text-to-prosodic features such as F0 and energy and can be used in downstream tasks…

Audio and Speech Processing · Electrical Eng. & Systems 2025-08-14 Eray Eren , Qingju Liu , Hyeongwoo Kim , Pablo Garrido , Abeer Alwan

Realistic mathematical modeling of voice production has been recently boosted by applications to different fields like bioprosthetics, quality speech synthesis and pathological diagnosis. In this work, we revisit a two-mass model of the…

Neurons and Cognition · Quantitative Biology 2013-12-12 María Florencia Assaneo , Marcos A. Trevisan

We show how to cope with the acoustic identification of poroelastic materials when the specimen is in the form of a cylinder. We apply our formulation, based on the Biot model, approximated by the equivalent elastic solid model, to a long…

Classical Physics · Physics 2007-05-23 Zine Fellah , Jean-Philippe Groby , Erick Ogam , Thierry Scotti , Armand Wirgin

Many hearables contain an in-ear microphone, which may be used to capture the own voice of its user. However, due to the hearable occluding the ear canal, the in-ear microphone mostly records body-conducted speech, typically suffering from…

Audio and Speech Processing · Electrical Eng. & Systems 2024-09-09 Mattes Ohlenbusch , Christian Rollwage , Simon Doclo

While rendering and animation of photorealistic 3D human body models have matured and reached an impressive quality over the past years, modeling the spatial audio associated with such full body models has been largely ignored so far. In…

Sound · Computer Science 2024-07-23 Chao Huang , Dejan Markovic , Chenliang Xu , Alexander Richard

There has been fascinating work on creating artistic transformations of images by Gatys. This was revolutionary in how we can in some sense alter the 'style' of an image while generally preserving its 'content'. In our work, we present a…

Sound · Computer Science 2024-12-24 Prateek Verma , Julius O. Smith

We address the problem of reconstructing articulatory movements, given audio and/or phonetic labels. The scarce availability of multi-speaker articulatory data makes it difficult to learn a reconstruction that generalizes to new speakers…

Computation and Language · Computer Science 2023-09-13 Rosanna Turrisi , Raffaele Tavarone , Leonardo Badino

Most of the prior studies in the spatial \ac{DoA} domain focus on a single modality. However, humans use auditory and visual senses to detect the presence of sound sources. With this motivation, we propose to use neural networks with audio…

Sound · Computer Science 2021-05-14 Xinyuan Qian , Maulik Madhavi , Zexu Pan , Jiadong Wang , Haizhou Li

Ducted flow devices for a range of purposes, such as air-moving fans, are routinely characterised experimentally to understand their acoustic performance as part of the continuing trend for quiet, high efficiency design. The International…

Fluid Dynamics · Physics 2014-04-28 Timothy J. Newman , Anurag Agarwal , Ann P. Dowling , Ludovic Desvard

Cough is a protective reflex conveying information on the state of the respiratory system. Cough assessment has been limited so far to subjective measurement tools or uncomfortable (i.e., non-wearable) cough monitors. This limits the…

Audio and Speech Processing · Electrical Eng. & Systems 2024-12-04 Jesús Monge-Alvarez , Carlos Hoyos-Barceló , Luis M. San-José-Revuelta , Pablo Casaseca-de-la-Higuera

Voice conversion (VC) techniques aim to modify speaker identity of an utterance while preserving the underlying linguistic information. Most VC approaches ignore modeling of the speaking style (e.g. emotion and emphasis), which may contain…

Audio and Speech Processing · Electrical Eng. & Systems 2020-05-20 Songxiang Liu , Yuewen Cao , Shiyin Kang , Na Hu , Xunying Liu , Dan Su , Dong Yu , Helen Meng

We propose the use of a self-oscillating dynamical system --the pre-Galileian clock equation-- for modeling the laryngeal tone. The parameters are shown to be the minimal control needed for generating the prosody of the human speech. Based…

Biological Physics · Physics 2009-11-10 Roberto D'Autilia

This paper presents an active noise control experiment designed to validate a real-time control strategy for reduction of the noise scattered from a three-dimensional body. The control algorithm relies on estimating the scattered noise by…

Classical Physics · Physics 2016-08-16 Emmanuel Friot , Régine Guillermin , Muriel Winninger

Speech foundation models have demonstrated exceptional capabilities in speech-related tasks. Nevertheless, these models often struggle with non-verbal audio data, such as vocalizations, baby crying, etc., which are critical for various…

Audio and Speech Processing · Electrical Eng. & Systems 2025-02-25 Alkis Koudounas , Moreno La Quatra , Marco Sabato Siniscalchi , Elena Baralis