中文
相关论文

相关论文: Modal locking between vocal fold and vocal tract o…

200 篇论文

This work addresses friction-induced modal interactions in jointed structures, and their effects on the passive mitigation of vibrations by means of friction damping. Under the condition of (nearly) commensurable natural frequencies, the…

斑图形成与孤子 · 物理学 2021-01-12 Malte Krack , Lawrence A. Bergman , Alexander F. Vakakis

We investigate bubble deformations in an homogeneous and isotropic turbulent flow by means of direct numerical simulations of a single bubble in turbulence. We examine interface deformations by decomposing the local radius into the…

流体动力学 · 物理学 2024-07-24 Aliénor Rivière , Kamel Abahri , Stéphane Perrard

The complex behavior of many natural and engineered systems emerges from the interaction of a small number of effective degrees of freedom. Discovering the physical basis of the interactions between these degrees of freedom directly from…

混沌动力学 · 物理学 2026-03-18 Annie Z. Xia , Melody X. Lim , Jason Z. Kim , Bryan VanSaders , Heinrich Jaeger

Text does not fully specify the spoken form, so text-to-speech models must be able to learn from speech data that vary in ways not explained by the corresponding text. One way to reduce the amount of unexplained variation in training data…

Dialogue models falter in noisy, multi-speaker environments, often producing irrelevant responses and awkward turn-taking. We present AV-Dialog, the first multimodal dialog framework that uses both audio and visual cues to track the target…

计算与语言 · 计算机科学 2025-11-17 Tuochao Chen , Bandhav Veluri , Hongyu Gong , Shyamnath Gollakota

Vocal feedback (e.g., `mhm', `yeah', `okay') is an important component of spoken dialogue and is crucial to ensuring common ground in conversational systems. The exact meaning of such feedback is conveyed through both lexical and prosodic…

计算与语言 · 计算机科学 2025-05-20 Livia Qian , Carol Figueroa , Gabriel Skantze

The ubiquitous phenomenon of synchronization is inherently characteristic of dynamical dissipative non-linear systems. In particular, synchronization has been theoretically and experimentally demonstrated for exciton-polariton condensates…

This paper proposes ESTVocoder, a novel excitation-spectral-transformed neural vocoder within the framework of source-filter theory. The ESTVocoder transforms the amplitude and phase spectra of the excitation into the corresponding speech…

声音 · 计算机科学 2024-11-19 Xiao-Hang Jiang , Hui-Peng Du , Yang Ai , Ye-Xin Lu , Zhen-Hua Ling

Cross-lingual alignment in pretrained language models enables knowledge transfer across languages. Similar alignment has been reported in Whisper-style speech encoders, based on spoken translation retrieval using representational…

计算与语言 · 计算机科学 2026-04-07 Ryan Soh-Eun Shim , Domenico De Cristofaro , Chengzhi Martin Hu , Alessandro Vietti , Barbara Plank

The introduction of audio latent diffusion models possessing the ability to generate realistic sound clips on demand from a text description has the potential to revolutionize how we work with audio. In this work, we make an initial attempt…

音频与语音处理 · 电气工程与系统科学 2023-10-17 Dimitrios Bralios , Gordon Wichern , François G. Germain , Zexu Pan , Sameer Khurana , Chiori Hori , Jonathan Le Roux

Experimental results and their interpretations are presented on the nonlinear acoustic effects of multiple scattered elastic waves in unconsolidated granular media. Short wave packets with a central frequency higher than the so-called…

经典物理 · 物理学 2008-12-18 Vincent Tournat , Vitalyi Gusev

We present a physics-informed voiced backend renderer for singing-voice synthesis. Given synthetic single-channel audio and a fund-amental--frequency trajectory, we train a time-domain Webster model as a physics-informed neural network to…

声音 · 计算机科学 2026-03-03 Minhui Lu , Joshua D. Reiss

Speech foundation models have demonstrated exceptional capabilities in speech-related tasks. Nevertheless, these models often struggle with non-verbal audio data, such as vocalizations, baby crying, etc., which are critical for various…

音频与语音处理 · 电气工程与系统科学 2025-02-25 Alkis Koudounas , Moreno La Quatra , Marco Sabato Siniscalchi , Elena Baralis

We introduce a neural auto-encoder that transforms the musical dynamic in recordings of singing voice via changes in voice level. Since most recordings of singing voice are not annotated with voice level we propose a means to estimate the…

音频与语音处理 · 电气工程与系统科学 2023-10-06 Frederik Bous , Axel Roebel

A sound synthesis model for woodwind instruments is developed using modal decomposition of the input impedance, accounting for viscothermal losses as well as localized nonlinear losses at the end of the resonator. To extend the definition…

经典物理 · 物理学 2024-01-12 N Szwarcberg , T Colinot , C Vergez , M Jousserand

Despite important progress, conversational systems often generate dialogues that sound unnatural to humans. We conjecture that the reason lies in their different training and testing conditions: agents are trained in a controlled "lab"…

计算与语言 · 计算机科学 2021-04-01 Alberto Testoni , Raffaella Bernardi

The effect of electron-phonon coupling on the current noise in a molecular junction is investigated within a simple model. The model comprises a 1-level bridge representing a molecular level that connects between two free electron…

介观与纳米尺度物理 · 物理学 2007-05-23 Michael Galperin , Abraham Nitzan , Mark A. Ratner

Current speech production systems predominantly rely on large transformer models that operate as black boxes, providing little interpretability or grounding in the physical mechanisms of human speech. We address this limitation by proposing…

音频与语音处理 · 电气工程与系统科学 2025-10-08 Akshay Anand , Chenxu Guo , Cheol Jun Cho , Jiachen Lian , Gopala Anumanchipalli

Assessment of voice signals has long been performed with the assumption of periodicity as this facilitates analysis. Near periodicity of normal voice signals makes short-time harmonic modeling an appealing choice to extract vocal feature…

音频与语音处理 · 电气工程与系统科学 2022-02-10 Takeshi Ikuma , Andrew J. McWhorter , Lacey Adkins , Melda Kunduk

We use linear stability analysis and hybrid lattice Boltzmann simulations to study the dynamical behaviour of an active nematic confined in a channel made of viscoelastic material. We find that the quiescent, ordered active nematic is…

软凝聚态物质 · 物理学 2023-12-20 Francesco Mori , Saraswat Bhattacharyya , Julia M. Yeomans , Sumesh P. Thampi