English
Related papers

Related papers: SELEBI: Percussion-aware Time Stretching via Selec…

200 papers

Recent research has delved into speech enhancement (SE) approaches that leverage audio embeddings from pre-trained models, diverging from time-frequency masking or signal prediction techniques. This paper introduces an efficient and…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-16 Xingwei Sun , Heinrich Dinkel , Yadong Niu , Linzhang Wang , Junbo Zhang , Jian Luan

Methods are proposed for modifying the reverberation characteristics of sound fields in rooms by employing a loudspeaker with adjustable directivity, realized with a compact spherical loudspeaker array (SLA). These methods are based on…

Audio and Speech Processing · Electrical Eng. & Systems 2024-01-09 Hai Morgenstern , Boaz Rafaely

Non-uniform time stepping in acoustic propagation models can be used to preserve accuracy or reduce computational cost for an acoustic simulation with a wave front propagating through a domain with both heterogeneous and homogenous regions,…

Numerical Analysis · Mathematics 2025-07-11 Matthew J. King , B. E. Treeby , B. T. Cox

Time-domain speech enhancement (SE) has recently been intensively investigated. Among recent works, DEMUCS introduces multi-resolution STFT loss to enhance performance. However, some resolutions used for STFT contain non-stationary signals,…

Sound · Computer Science 2023-03-28 Hao Shi , Masato Mimura , Longbiao Wang , Jianwu Dang , Tatsuya Kawahara

Temporal segmentation of untrimmed videos and photo-streams is currently an active area of research in computer vision and image processing. This paper proposes a new approach to improve the temporal segmentation of photo-streams. The…

Image and Video Processing · Electrical Eng. & Systems 2019-06-28 Mariella Dimiccoli , Herwig Wendt

Differentiable digital signal processing (DDSP) techniques, including methods for audio synthesis, have gained attention in recent years and lend themselves to interpretability in the parameter space. However, current differentiable…

In dynamic acoustic environments with time-varying interferers, effective beamforming requires identifying stationary regions over time. The Capon beamformer, a whitened matched filter constrained to maintain unity gain in the desired…

Signal Processing · Electrical Eng. & Systems 2026-05-26 Manan Mittal , Ryan M. Corey , Diego Cuji , John R. Buck , Andrew C. Singer

In this paper, a novel architecture for speaker recognition is proposed by cascading speech enhancement and speaker processing. Its aim is to improve speaker recognition performance when speech signals are corrupted by noise. Instead of…

Computation and Language · Computer Science 2020-05-25 Yanpei Shi , Qiang Huang , Thomas Hain

Speaker extraction aims to extract the target speech signal from a multi-talker environment given a target speaker's reference speech. We recently proposed a time-domain solution, SpEx, that avoids the phase estimation in frequency-domain…

Audio and Speech Processing · Electrical Eng. & Systems 2020-08-19 Meng Ge , Chenglin Xu , Longbiao Wang , Eng Siong Chng , Jianwu Dang , Haizhou Li

Domain adaptation methods aim to bridge the gap between datasets by enabling knowledge transfer across domains, reducing the need for additional expert annotations. However, many approaches struggle with reliability in the target domain, an…

Image and Video Processing · Electrical Eng. & Systems 2026-05-14 Arnaud Judge , Nicolas Duchateau , Thierry Judge , Roman A. Sandler , Joseph Z. Sokol , Christian Desrosiers , Olivier Bernard , Pierre-Marc Jodoin

Speaker verification (SV) utilizing features obtained from models pre-trained via self-supervised learning has recently demonstrated impressive performances. However, these pre-trained models (PTMs) usually have a temporal resolution of 20…

Audio and Speech Processing · Electrical Eng. & Systems 2026-01-28 Jisoo Myoung , Sangwook Han , Kihyuk Kim , Jong Won Shin

Recent progress in diffusion-based audio generation and restoration has substantially improved performance across heterogeneous conditioning regimes, including text-conditioned audio generation and audio-conditioned super-resolution.…

Sound · Computer Science 2026-05-07 Xuanhao Zhang , Chang Li

Operating power amplifiers (PAs) at lower input back-off (IBO) levels is an effective way to improve PA efficiency, but often introduces severe nonlinear distortion that degrades transmission performance. Amplitude-phase-time block…

Signal Processing · Electrical Eng. & Systems 2026-05-01 Meidong Xia , Min Fan , Wei Xu , Haiming Wang , Xiaohu You

Single-channel speech separation in time domain and frequency domain has been widely studied for voice-driven applications over the past few years. Most of previous works assume known number of speakers in advance, however, which is not…

Audio and Speech Processing · Electrical Eng. & Systems 2020-04-02 Yiming Xiao , Haijian Zhang

Semilinear hyperbolic stochastic partial differential equations (SPDEs) find widespread applications in the natural and engineering sciences. However, the traditional Gaussian setting may prove too restrictive, as phenomena in mathematical…

Numerical Analysis · Mathematics 2023-07-04 Andrea Barth , Andreas Stein

Current quantum computers suffer from non-stationary noise channels with high error rates, which undermines their reliability and reproducibility. We propose a Bayesian inference-based adaptive algorithm that can learn and mitigate quantum…

Quantum Physics · Physics 2023-08-30 Samudra Dasgupta , Arshag Danageozian , Travis S. Humble

We propose a novel method to reconstruct the spatio-temporal amplitude and phase of the electric field of ultrashort laser pulses using spatially-resolved spectral interferometry. This method is based on a fiber-optic coupler interferometer…

This paper presents two single channel speech dereverberation methods to enhance the quality of speech signals that have been recorded in an enclosed space. For both methods, the room acoustics are modeled using a nonnegative approximation…

Sound · Computer Science 2017-09-19 Nasser Mohammadiha , Simon Doclo

Traditional sound diffusers are quasi-random phase gratings attached to reflecting surfaces whose purpose is to augment the spatiotemporal incoherence of the acoustic field scattered from reflective surfaces. This configuration allows one…

Applied Physics · Physics 2022-11-23 Janghoon Kang , Michael R. Haberman

Monitoring structures of elastic materials for defect detection by means of ultrasound waves (Structural Health Monitoring, SHM) demands for an efficient computation of parameters which characterize their mechanical behavior.…

Numerical Analysis · Mathematics 2019-11-13 Rebecca Klein , Thomas Schuster , Anne Wald
‹ Prev 1 3 4 5 6 7 10 Next ›