中文
相关论文

相关论文: A Perceptually Motivated Filter Bank with Perfect …

200 篇论文

This paper introduces an explainable DNN-based beamformer with a postfilter (ExNet-BF+PF) for multichannel signal processing. Our approach combines the U-Net network with a beamformer structure to address this problem. The method involves a…

音频与语音处理 · 电气工程与系统科学 2024-11-19 Adi Cohen , Daniel Wong , Jung-Suk Lee , Sharon Gannot

Many super-resolution (SR) models are optimized for high performance only and therefore lack efficiency due to large model complexity. As large models are often not practical in real-world applications, we investigate and propose novel loss…

图像与视频处理 · 电气工程与系统科学 2021-06-03 Dario Fuoli , Luc Van Gool , Radu Timofte

Data is said to follow the transform (or analysis) sparsity model if it becomes sparse when acted on by a linear operator called a sparsifying transform. Several algorithms have been designed to learn such a transform directly from data,…

机器学习 · 统计学 2019-01-30 Luke Pfister , Yoram Bresler

Reconstructing a 3D sound field from sparse microphone measurements is a fundamental yet ill-posed problem, which we address through Acoustic Transfer Function (ATF) magnitude estimation. ATF magnitude encapsulates key perceptual and…

音频与语音处理 · 电气工程与系统科学 2026-05-12 Ege Erdem , Shoichi Koyama , Tomohiko Nakamura , Orchisama Das , Zoran Cvetković

While much of modern speech and audio processing relies on deep neural networks trained using fixed audio representations, recent studies suggest great potential in acoustic frontends learnt jointly with a backend. In this study, we focus…

音频与语音处理 · 电气工程与系统科学 2023-02-21 Mark Anderson , Tomi Kinnunen , Naomi Harte

This review chapter aims to strengthen the link between frame theory and signal processing tasks in psychoacoustics. On the one side, the basic concepts of frame theory are presented and some proofs are provided to explain those concepts in…

声音 · 计算机科学 2020-09-11 Peter Balazs , Nicki Holighaus , Thibaud Necciari , Diana Stoeva

Frame is the corner stone for designing decomposition and reconstruction operations in signal processing. Famous frames include wavelets, curvelets,and Gabor. A celebrated result indicates that if a synthesis frame is chosen for…

最优化与控制 · 数学 2017-04-10 Wen-Liang Hwang

Similarity-based collaborative filtering (CF) models have long demonstrated strong offline performance and conceptual simplicity. However, their scalability is limited by the quadratic cost of maintaining dense item-item similarity…

信息检索 · 计算机科学 2026-01-27 Domenico de Gioia , Claudio Pomo , Ludovico Boratto , Tommaso Di Noia

We present the Fourier-Invertible Neural Encoder (FINE), a compact and interpretable architecture for dimension reduction in translation-equivariant datasets. FINE integrates reversible filters and monotonic activation functions with a…

机器学习 · 计算机科学 2025-12-02 Anqiao Ouyang , Hongyi Ke , Qi Wang

To address the limitations of existing Generative Fixed-Filter Active Noise Control (GFANC) methods, which rely on filter decomposition and recombination and require supervised learning with labeled data, this paper proposes a…

音频与语音处理 · 电气工程与系统科学 2026-05-04 Ziyi Yang , Zhengding Luo , Yisong Zou , Boxiang Wang , Qirui Huang , Woon-Seng Gan

Musicians and audio engineers sculpt and transform their sounds by connecting multiple processors, forming an audio processing graph. However, most deep-learning methods overlook this real-world practice and assume fixed graph settings. To…

声音 · 计算机科学 2023-05-09 Sungho Lee , Jaehyun Park , Seungryeol Paik , Kyogu Lee

Tremendous progress has been made in sequential processing with the recent advances in recurrent neural networks. However, recurrent architectures face the challenge of exploding/vanishing gradients during training, and require significant…

神经与进化计算 · 计算机科学 2021-10-13 Bing Han , Cheng Wang , Kaushik Roy

Despite the recent popularity of deep generative state space models, few comparisons have been made between network architectures and the inference steps of the Bayesian filtering framework -- with most models simultaneously approximating…

机器学习 · 统计学 2020-09-29 Bryan Lim , Stefan Zohren , Stephen Roberts

We present filters with rational exponents in order to provide a continuum of filter behavior not classically achievable. We discuss their stability, the flexibility they afford, and various representations useful for analysis, design and…

信号处理 · 电气工程与系统科学 2025-04-01 Samiya A Alkhairy

In this paper, a novel decomposition method for non-stationary and nonlinear signals is proposed. This method is inspired by the adaptive wavelet filter bank of the empirical wavelet transform (EWT) and Fourier intrinsic band functions…

信号处理 · 电气工程与系统科学 2019-12-03 Wei Zhou , Zhongren Feng , Xiongjiang Wang , Hao Lv

Compressed sensing is triggering a major evolution in signal acquisition. It consists in sampling a sparse signal at low rate and later using computational power for its exact reconstruction, so that only the necessary information is…

统计力学 · 物理学 2012-06-07 Florent Krzakala , Marc Mézard , François Sausset , Yifan Sun , Lenka Zdeborová

We propose Re-parameterized Refocusing Convolution (RefConv) as a replacement for regular convolutional layers, which is a plug-and-play module to improve the performance without any inference costs. Specifically, given a pre-trained model,…

计算机视觉与模式识别 · 计算机科学 2023-10-17 Zhicheng Cai , Xiaohan Ding , Qiu Shen , Xun Cao

In recent years, filterbank learning has become an increasingly popular strategy for various audio-related machine learning tasks. This is partly due to its ability to discover task-specific audio characteristics which can be leveraged in…

音频与语音处理 · 电气工程与系统科学 2022-11-14 Frank Cwitkowitz , Mojtaba Heydari , Zhiyao Duan

We introduce a novel method for designing attenuation filters in digital audio reverberation systems based on Feedback Delay Networks (FDNs). Our approach uses Second Order Sections (SOS) of Infinite Impulse Response (IIR) filters arranged…

声音 · 计算机科学 2025-11-26 Ilias Ibnyahya , Joshua D. Reiss

Achieving high-fidelity audio compression while preserving perceptual quality across diverse content remains a key challenge in Neural Audio Coding (NAC). We introduce MUFFIN, a fully convolutional Neural Psychoacoustic Coding (NPC)…

声音 · 计算机科学 2025-05-13 Dianwen Ng , Kun Zhou , Yi-Wen Chao , Zhiwei Xiong , Bin Ma , Eng Siong Chng
‹ 上一页 1 8 9 10 下一页 ›