English
Related papers

Related papers: Geometrically-Motivated Primary-Ambient Decomposit…

200 papers

Conventional approaches to sound localization and separation are based on microphone arrays in artificial systems. Inspired by the selective perception of human auditory system, we design a multi-source listening system which can separate…

Sound · Computer Science 2019-11-11 Xuecong Sun , Han Jia , Zhe Zhang , Yuzhen Yang , Zhaoyong Sun , Jun Yang

One-dimensional signal decomposition is a well-established and widely used technique across various scientific fields. It serves as a highly valuable pre-processing step for data analysis. While traditional decomposition techniques often…

Machine Learning · Computer Science 2025-06-09 Samuele Salti , Andrea Pinto , Alessandro Lanza , Serena Morigi

Motivated by recent developments in perturbative calculations of the nonlinear evolution of large-scale structure, we present an iterative algorithm to reconstruct the initial conditions in a given volume starting from the dark matter…

Cosmology and Nongalactic Astrophysics · Physics 2020-12-01 Marcel Schmittfull , Tobias Baldauf , Matias Zaldarriaga

Outdoor scene reconstruction remains challenging due to the stark contrast between well-textured, nearby regions and distant backgrounds dominated by low detail, uneven illumination, and sky effects. We introduce a two-stage Gaussian…

Graphics · Computer Science 2025-10-13 Deborah Pintani , Ariel Caputo , Noah Lewis , Marc Stamminger , Fabio Pellacini , Andrea Giachetti

This work presents a new cyclic architecture that extracts high-frequency patterns from images and re-insert them as geometric features. This procedure allows us to enhance the resolution of low-cost depth sensors capturing fine details on…

Computer Vision and Pattern Recognition · Computer Science 2022-03-08 Liran Azaria , Dan Raviv

In the stereo-to-multichannel upmixing problem for music, one of the main tasks is to set the directionality of the instrument sources in the multichannel rendering results. In this paper, we propose a modified variational autoencoder model…

Audio and Speech Processing · Electrical Eng. & Systems 2022-03-24 Haici Yang , Sanna Wager , Spencer Russell , Mike Luo , Minje Kim , Wontak Kim

Can uncorrelated surrounding sound sources be used to generate extended diffuse sound fields? By definition, targets are a constant sound pressure level, a vanishing average sound intensity, uncorrelated sound waves arriving isotropically…

Audio and Speech Processing · Electrical Eng. & Systems 2024-02-22 Franz Zotter , Stefan Riedel , Lukas Gölles , Matthias Frank

The research reported in this paper addresses the fundamental task of separation of locally moving or deforming image areas from a static or globally moving background. It builds on the latest developments in the field of robust principal…

Computer Vision and Pattern Recognition · Computer Science 2016-03-21 Salehe Erfanian Ebadi , Valia Guerra Ones , Ebroul Izquierdo

Self-supervised audio-visual source separation leverages natural correlations between audio and vision modalities to separate mixed audio signals. In this work, we first systematically analyse the performance of existing multimodal fusion…

Multimedia · Computer Science 2025-10-10 Han Hu , Dongheng Lin , Qiming Huang , Yuqi Hou , Hyung Jin Chang , Jianbo Jiao

In this work, we present a two-stage method for speaker extraction under reverberant and noisy conditions. Given a reference signal of the desired speaker, the clean, but the still reverberant, desired speaker is first extracted from the…

Sound · Computer Science 2023-03-14 Aviad Eisenberg , Sharon Gannot , Shlomo E. Chazan

This letter introduces an innovative method to enhance the quality of audio time stretching by precisely decomposing a sound into sines, transients, and noise and by improving the processing of the latter component. While there are…

Audio and Speech Processing · Electrical Eng. & Systems 2023-12-25 Eloi Moliner , Leonardo Fierro , Alec Wright , Matti Hämäläinen , Vesa Välimäki

Image foreground extraction is a classical problem in image processing and vision, with a large range of applications. In this dissertation, we focus on the extraction of text and graphics in mixed-content images, and design novel…

Computer Vision and Pattern Recognition · Computer Science 2018-04-10 Shervin Minaee

In this paper, we present a novel multi-channel speech extraction system to simultaneously extract multiple clean individual sources from a mixture in noisy and reverberant environments. The proposed method is built on an improved…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-17 Jisi Zhang , Catalin Zorila , Rama Doddipatla , Jon Barker

Beamforming is a signal processing technique. It has been studied in many areas such as radar, sonar, seismology and wireless communications, to name but a few. It can be used for a myriad of purposes, such as detecting the presence of a…

Other Computer Science · Computer Science 2012-12-27 Hidri Adel , Meddeb Souad , Abdulqadir Alaqeeli , Amiri Hamid

The end-to-end approaches for single-channel target speech extraction have attracted widespread attention. However, the studies for end-to-end multi-channel target speech extraction are still relatively limited. In this work, we propose two…

Audio and Speech Processing · Electrical Eng. & Systems 2020-10-23 Jiangyu Han , Xinyuan Zhou , Yanhua Long , Yijie Li

Acquired images for medical and other purposes can be affected by noise from both the equipment used in the capturing or the environment. This can have adverse effect on the information therein. Thus, the need to restore the image to its…

Image and Video Processing · Electrical Eng. & Systems 2024-01-19 E. G. Onyedinma , I. E. Onyenwe

This paper presents novel approaches for efficient feature extraction using environmental sound magnitude spectrogram. We propose approach based on the visual domain. This approach included three methods. The first method is based on…

Computer Vision and Pattern Recognition · Computer Science 2012-09-27 Sameh Souli , Zied Lachiri

Channel state information (CSI) is critical for multi-input multi-output (MIMO) orthogonal frequency division multiplexing (OFDM) system. Pilot-based channel estimation methods suffer from high pilot overhead and low channel acquisition…

Signal Processing · Electrical Eng. & Systems 2026-05-11 Hongning Ruan , Zhaoyang Zhang , Zirui Chen , Ziqing Xing , Zhaohui Yang

This paper is motivated by the problem of integrating multiple sources of measurements. We consider two multiple-input-multiple-output (MIMO) channels, a primary channel and a secondary channel, with dependent input signals. The primary…

Information Theory · Computer Science 2013-06-25 Yuan Wang , Haonan Wang , Louis Scharf

A spatial active noise control (ANC) method based on the individual kernel interpolation of primary and secondary sound fields is proposed. Spatial ANC is aimed at cancelling unwanted primary noise within a continuous region by using…

Audio and Speech Processing · Electrical Eng. & Systems 2022-02-11 Kazuyuki Arikawa , Shoichi Koyama , Hiroshi Saruwatari