中文
相关论文

相关论文: How far are vowel formants from computed vocal tra…

200 篇论文

Although many previous studies have carried out multimodal learning with real-time MRI data that captures the audio-visual kinematics of the vocal tract during speech, these studies have been limited by their reliance on multi-speaker…

The goal of this work is to recover articulatory information from the speech signal by acoustic-to-articulatory inversion. One of the main difficulties with inversion is that the problem is underdetermined and inversion methods generally…

计算与语言 · 计算机科学 2007-05-23 Blaise Potard , Yves Laprie

Acoustic-to-articulatory inversion (AAI) methods estimate articulatory movements from the acoustic speech signal, which can be useful in several tasks such as speech recognition, synthesis, talking heads and language tutoring. Most earlier…

音频与语音处理 · 电气工程与系统科学 2020-08-06 Tamás Gábor Csapó

Accurate segmentation of the vocal tract from magnetic resonance imaging (MRI) data is essential for various voice and speech applications. Manual segmentation is time intensive and susceptible to errors. This study aimed to evaluate the…

计算机视觉与模式识别 · 计算机科学 2025-03-03 Subin Erattakulangara , Karthika Kelat , Katie Burnham , Rachel Balbi , Sarah E. Gerard , David Meyer , Sajan Goud Lingala

The Vocal Joystick Vowel Corpus, by Washington University, was used to study monophthongs pronounced by native English speakers. The objective of this study was to quantitatively measure the extent at which speech recognition methods can…

计算与语言 · 计算机科学 2017-02-24 Keith Y. Patarroyo , Vladimir Vargas-Calderón

The articulatory-acoustic relationship is many-to-one and non linear and this is a great limitation for studying speech production. A simplification is proposed to set a bijection between the vowel space (f1, f2) and the parametric space of…

声音 · 计算机科学 2021-11-03 Frédéric Berthommier

Imprecise vowel articulation can be observed in people with Parkinson's disease (PD). Acoustic features measuring vowel articulation have been demonstrated to be effective indicators of PD in its assessment. Standard clinical vowel…

音频与语音处理 · 电气工程与系统科学 2021-08-18 Yuanyuan Liu , Nelly Penttilä , Tiina Ihalainen , Juulia Lintula , Rachel Convey , Okko Räsänen

We introduce and analyze a virtual element method (VEM) for the Helmholtz problem with approximating spaces made of products of low order VEM functions and plane waves. We restrict ourselves to the 2D Helmholtz equation with impedance…

数值分析 · 数学 2015-05-20 Ilaria Perugia , Paola Pietra , Alessandro Russo

Real-time Magnetic Resonance Imaging (rtMRI) visualizes vocal tract action, offering a comprehensive window into speech articulation. However, its signals are high dimensional and noisy, hindering interpretation. We investigate compact…

图像与视频处理 · 电气工程与系统科学 2026-01-30 Jay Park , Hong Nguyen , Sean Foley , Jihwan Lee , Yoonjeong Lee , Dani Byrd , Shrikanth Narayanan

In this paper, we propose a novel approach for accurate detection of the vowel onset points (VOPs). VOP is the instant at which the vowel begins in the speech signal. Precise identification of VOPs is important for various speech…

音频与语音处理 · 电气工程与系统科学 2019-08-26 Kumud Tripathi , K. Sreenivasa Rao

Confining sound is of significant importance for the manipulation and routing acoustic waves. We propose a Helmholtz resonator (HR) based subwavelength sound channel formed at the interface of two metamaterials, for this purpose. The…

应用物理 · 物理学 2021-03-31 Yun Zhou , Prabhakar R. Bandaru , Daniel F. Sievenpiper

Source separation for music is the task of isolating contributions, or stems, from different instruments recorded individually and arranged together to form a song. Such components include voice, bass, drums and any other…

声音 · 计算机科学 2021-04-29 Alexandre Défossez , Nicolas Usunier , Léon Bottou , Francis Bach

Sonorant sounds are characterized by regions with prominent formant structure, high energy and high degree of periodicity. In this work, the vocal-tract system, excitation source and suprasegmental features derived from the speech signal…

声音 · 计算机科学 2021-07-02 Bidisha Sharma , S. R. Mahadeva Prasanna

In this paper, we combine Hidden Markov Models (HMMs) with i-vector extractors to address the problem of text-dependent speaker recognition with random digit strings. We employ digit-specific HMMs to segment the utterances into digits, to…

音频与语音处理 · 电气工程与系统科学 2019-07-16 Nooshin Maghsoodi , Hossein Sameti , Hossein Zeinali , Themos~Stafylakis

Speaker embeddings achieve promising results on many speaker verification tasks. Phonetic information, as an important component of speech, is rarely considered in the extraction of speaker embeddings. In this paper, we introduce phonetic…

声音 · 计算机科学 2018-06-15 Yi Liu , Liang He , Jia Liu , Michael T. Johnson

Acoustic articulatory inversion is a major processing challenge, with a wide range of applications from speech synthesis to feedback systems for language learning and rehabilitation. In recent years, deep learning methods have been applied…

音频与语音处理 · 电气工程与系统科学 2026-03-13 Sofiane Azzouz , Pierre-André Vuissoz , Yves Laprie

Under general assumptions, the numbers of semiclassical resonances is known to be bounded from above by a negative power of $h$ which is given by the fractal dimension of the trapped set. In this paper we provide examples of operators with…

偏微分方程分析 · 数学 2025-12-04 Jean-Francois Bony , Setsuro Fujiie , Thierry Ramond , Maher Zerzeri

Using a known speaker-intrinsic normalization procedure, formant data are scaled by the reciprocal of the geometric mean of the first three formant frequencies. This reduces the influence of the talker but results in a distorted vowel…

声音 · 计算机科学 2016-12-13 T. V. Ananthapadmanabha , A. G. Ramakrishnan

Articulatory features can provide interpretable and flexible controls for the synthesis of human vocalizations by allowing the user to directly modify parameters like vocal strain or lip position. To make this manipulation through…

声音 · 计算机科学 2023-07-11 David Südholt , Mateo Cámara , Zhiyuan Xu , Joshua D. Reiss

The inverse problem of determining the cross-sectional area of a human vocal tract during the utterance of a vowel is considered. The frequency-dependent boundary condition at the lips is expressed in terms of the acoustic impedance of a…

数学物理 · 物理学 2021-03-16 Tuncay Aktosun , Paul Sacks , Xiao-Chuan Xu