English
Related papers

Related papers: Unsupervised Harmonic Parameter Estimation Using D…

200 papers

In this paper, we address the problem of pitch estimation using Self Supervised Learning (SSL). The SSL paradigm we use is equivariance to pitch transposition, which enables our model to accurately perform pitch estimation on monophonic…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-28 Alain Riou , Stefan Lattner , Gaëtan Hadjeres , Geoffroy Peeters

Speech enhancement in hearing aids remains a difficult task in nonstationary acoustic environments, mainly because current signal processing algorithms rely on fixed, manually tuned parameters that cannot adapt in situ to different users or…

Use of an autoencoder (AE) as a normal model is a state-of-the-art technique for unsupervised-anomaly detection in sounds (ADS). The AE is trained to minimize the sample mean of the anomaly score of normal sounds in a mini-batch. One…

Audio and Speech Processing · Electrical Eng. & Systems 2019-07-22 Yuma Koizumi , Shoichiro Saito , Masataka Yamaguchi , Shin Murata , Noboru Harada

Despite the growing interest in unsupervised learning, extracting meaningful knowledge from unlabelled audio remains an open challenge. To take a step in this direction, we recently proposed a problem-agnostic speech encoder (PASE), that…

Audio and Speech Processing · Electrical Eng. & Systems 2020-04-21 Mirco Ravanelli , Jianyuan Zhong , Santiago Pascual , Pawel Swietojanski , Joao Monteiro , Jan Trmal , Yoshua Bengio

The fragility of quantum systems makes them ideally suited for sensing applications at the nanoscale. However, interpreting the output signal of a qubit-based sensor is generally complicated by background clutter due to out-of-band spectral…

Quantum Physics · Physics 2020-08-19 Virginia Frey , Leigh M. Norris , Lorenza Viola , Michael J. Biercuk

We propose a method using a long short-term memory (LSTM) network to estimate the noise power spectral density (PSD) of single-channel audio signals represented in the short time Fourier transform (STFT) domain. An LSTM network common to…

Signal Processing · Electrical Eng. & Systems 2020-11-11 Xiaofei Li , Simon Leglaive , Laurent Girin , Radu Horaud

This study introduces a novel approach for estimating plane-wave coefficients in sound field reconstruction, specifically addressing challenges posed by error-in-variable phase perturbations. Such systematic errors typically arise from…

Signal Processing · Electrical Eng. & Systems 2026-02-06 Yuyang Liu , Johan Karlsson , Filip Elvander

Dimensionality reduction is critical for deploying dense retrieval systems at scale, yet mainstream post-hoc methods face a fundamental trade-off: principal component analysis (PCA) preserves dominant variance but underutilizes…

Information Retrieval · Computer Science 2026-04-20 Yongkang Li , Panagiotis Eustratiadis , Evangelos Kanoulas

Resource allocation and transceivers in wireless networks are usually designed by solving optimization problems subject to specific constraints, which can be formulated as variable or functional optimization. If the objective and constraint…

Machine Learning · Computer Science 2020-01-06 Dong Liu , Chengjian Sun , Chenyang Yang , Lajos Hanzo

This work addresses the mismatch problem between the distribution of training data (source) and testing data (target), in the challenging context of dysarthric speech recognition. We focus on Speaker Adaptation (SA) in command speech…

Computation and Language · Computer Science 2023-09-13 Rosanna Turrisi , Leonardo Badino

We introduce a framework for subspace methods which approximate the spectra of self-adjoint, unbounded operators in a local region. Using the projection-valued measure, we derive integrated spectral inequalities that also apply to unbounded…

Numerical Analysis · Mathematics 2026-01-06 Timothy Stroschein

Perceptual metrics are traditionally used to evaluate the quality of natural signals, such as images and audio. They are designed to mimic the perceptual behaviour of human observers and usually reflect structures found in natural signals.…

Sound · Computer Science 2023-12-07 Tashi Namgyal , Alexander Hepburn , Raul Santos-Rodriguez , Valero Laparra , Jesus Malo

We investigate unpaired image inverse problems, a challenging setting where only independent, non-paired sets of noisy measurements and clean target signals are available for training. We propose a novel inverse problem solver based on…

Machine Learning · Computer Science 2026-05-21 Donggyu Lee , Taekyung Lee , Jaewoong Choi

Physiological signals, such as the electrocardiogram and the phonocardiogram are very often corrupted by noisy sources. Usually, artificial intelligent algorithms analyze the signal regardless of its quality. On the other hand, physicians…

Signal Processing · Electrical Eng. & Systems 2023-04-25 Jorge Oliveira , Margarida Carvalho , Diogo Marcelo Nogueira , Miguel Coimbra

In this paper, we investigate the usage of autoencoders in modeling textual data. Traditional autoencoders suffer from at least two aspects: scalability with the high dimensionality of vocabulary size and dealing with task-irrelevant words.…

Machine Learning · Computer Science 2015-12-15 Shuangfei Zhai , Zhongfei Zhang

Anomaly detection relies on designing a score to determine whether a particular event is uncharacteristic of a given background distribution. One way to define a score is to use autoencoders, which rely on the ability to reconstruct certain…

High Energy Physics - Phenomenology · Physics 2022-03-30 Katherine Fraser , Samuel Homiller , Rashmish K. Mishra , Bryan Ostdiek , Matthew D. Schwartz

Denoising is omnipresent in image processing. It is usually addressed with algorithms relying on a set of hyperparameters that control the quality of the recovered image. Manual tuning of those parameters can be a daunting task, which calls…

Image and Video Processing · Electrical Eng. & Systems 2024-01-19 Arthur Floquet , Sayantan Dutta , Emmanuel Soubies , Duong Hung Pham , Denis Kouame

Sound processing in the human auditory system is complex and highly non-linear, whereas hearing aids (HAs) still rely on simplified descriptions of auditory processing or hearing loss to restore hearing. Even though standard HA…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-21 Fotios Drakopoulos , Sarah Verhulst

In neural-based audio feature extraction, ensuring that representations capture disentangled information is crucial for model interpretability. However, existing disentanglement methods often rely on assumptions that are highly dependent on…

Sound · Computer Science 2025-10-07 Benoit Ginies , Xiaoyu Bie , Olivier Fercoq , Gaël Richard

Growing research demonstrates that synthetic failure modes imply poor generalization. We compare commonly used audio-to-audio losses on a synthetic benchmark, measuring the pitch distance between two stationary sinusoids. The results are…

Sound · Computer Science 2020-12-11 Joseph Turian , Max Henry