English
Related papers

Related papers: Inferring Pitch from Coarse Spectral Features

200 papers

Articulatory acoustic inversion aims to reconstruct the complete geometry of the vocal tract from the speech signal. In this paper, we present a comparative study of several levels of phonetic segmentation accuracy, together with a…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-13 Sofiane Azzouz , Pierre-André Vuissoz , Yves Laprie

Emotions lie on a continuum, but current models treat emotions as a finite valued discrete variable. This representation does not capture the diversity in the expression of emotion. To better represent emotions we propose the use of natural…

Sound · Computer Science 2023-12-08 Hira Dhamyal , Benjamin Elizalde , Soham Deshmukh , Huaming Wang , Bhiksha Raj , Rita Singh

In this paper, we consider the effect of a bandwidth extension of narrow-band speech signals (0.3-3.4 kHz) to 0.3-8 kHz on speaker verification. Using covariance matrix based verification systems together with detection error trade-off…

Sound · Computer Science 2022-04-06 Marcos Faundez-Zanuy , Mattias Nilsson , W. Bastiaan Kleijn

We explore the use of the bispectrum for understanding quasiperiodic oscillations. The bispectrum is a statistic which probes the relations between the relative phases of the Fourier spectrum at different frequencies. The use of the…

High Energy Astrophysical Phenomena · Physics 2015-05-27 Thomas J. Maccarone , Philip Uttley , Michiel van der Klis , Rudy Wijnands , Paolo S. Coppi

We consider a scenario in which qubit-like probes are used to sense an external field that linearly affects their energy splitting. Following the frequency estimation approach in which one optimizes the state and sensing time of the probes…

The goal of this paper is to provide a theory linear regression based entirely on approximations. It will be argued that the standard linear regression model based theory whether frequentist or Bayesian has failed and that this failure is…

Methodology · Statistics 2024-02-16 Laurie Davies

The phase of an optical field inside a linear amplifier is widely known to diffuse with a diffusion coefficient that is inversely proportional to the photon number. The same process occurs in lasers which limits its intrinsic linewidth and…

Quantum Physics · Physics 2019-11-06 A. Chia , M. Hajdusek , R. Fazio , L. C. Kwek , V. Vedral

Speech signals are complex intermingling of various informative factors, and this information blending makes decoding any of the individual factors extremely difficult. A natural idea is to factorize each speech frame into independent…

Sound · Computer Science 2017-06-27 Dong Wang , Lantian Li , Ying Shi , Yixiang Chen , Zhiyuan Tang

Resonances are common in wave physics and their full and rigorous characterization is crucial to correctly tailor the response of a system in both time and frequency domains. However, they have been conventionally described by the quality…

Optics · Physics 2025-04-09 Isam Ben Soltane , Nicolas Bonod

In a hybrid speech model, both voiced and unvoiced components can coexist in a segment. Often, the voiced speech is regarded as the deterministic component, and the unvoiced speech and additive noise are the stochastic components.…

Audio and Speech Processing · Electrical Eng. & Systems 2021-05-05 Alfredo Esquivel Jaramillo , Jesper Kjær Nielsen , Mads Græsbøll Christensen

We use tensor analysis techniques for high-dimensional data to gain insight into pitch curves, which play an important role in linguistics research. In particular, we propose that demeaned phonetics pitch curve data can be modeled as having…

Methodology · Statistics 2018-08-17 Michael Hornstein , Shuheng Zhou , Kerby Shedden

In this work, we propose a new mathematical vocoder algorithm(modified spectral inversion) that generates a waveform from acoustic features without phase estimation. The main benefit of using our proposed method is that it excludes the…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-17 Hyun Gon Ryu , Jeong-Hoon Kim , Simon See

This paper provides a computational analysis of poetry reading audio signals at a large scale to unveil the musicality within professionally-read poems. Although the acoustic characteristics of other types of spoken language have been…

Sound · Computer Science 2024-04-02 Kahyun Choi , Minje Kim

Searching for a weak signal at an unknown frequency is a canonical task in experiments probing fundamental physics such as gravitational-wave observatories and ultra-light dark matter haloscopes. These state-of-the-art sensors are limited…

This paper proposes a speech rhythm-based method for speaker embeddings to model phoneme duration using a few utterances by the target speaker. Speech rhythm is one of the essential factors among speaker characteristics, along with acoustic…

Sound · Computer Science 2024-02-13 Kenichi Fujita , Atsushi Ando , Yusuke Ijima

When it comes to authentication in speaker verification systems, not all utterances are created equal. It is essential to estimate the quality of test utterances in order to account for varying acoustic conditions. In addition to the…

Audio and Speech Processing · Electrical Eng. & Systems 2024-07-12 Nicholas Klein , Ganesh Sivaraman , Elie Khoury

Some glottal analysis approaches based upon linear prediction or complex cepstrum approaches have been proved to be effective to estimate glottal source from real speech utterances. We propose a new approach employing both an all-pole…

Sound · Computer Science 2016-12-16 Yiqiao Chen , John N. Gowdy

The extraction of signals from noise is a common problem in all areas of science and engineering. A particularly useful version is that of forecasting: determining a causal filter that estimates a future value of a hidden process from past…

Optimization and Control · Mathematics 2026-02-02 Serhii Kryhin , Tatiana Mouzykantskii , Vivishek Sudhir

Although recent works on neural vocoder have improved the quality of synthesized audio, there still exists a gap between generated and ground-truth audio in frequency space. This difference leads to spectral artifacts such as hissing noise…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-15 Ji-Hoon Kim , Sang-Hoon Lee , Ji-Hyun Lee , Seong-Whan Lee

The properties of the normal distribution under linear transformation, as well the easy way to compute the covariance matrix of marginals and conditionals, offer a unique opportunity to get an insight about several aspects of uncertainties…

Data Analysis, Statistics and Probability · Physics 2018-02-12 Giulio D'Agostini