English
Related papers

Related papers: Time-Domain Wideband Image Source Method for Spher…

200 papers

We propose a novel Neural Steering technique that adapts the target area of a spatial-aware multi-microphone sound source separation algorithm during inference without the necessity of retraining the deep neural network (DNN). To achieve…

Audio and Speech Processing · Electrical Eng. & Systems 2024-10-23 Martin Strauss , Wolfgang Mack , María Luis Valero , Okan Köpüklü

Spherical microphone arrays (SMAs) and spherical loudspeaker arrays (SLAs) facilitate the study of room acoustics due to the three-dimensional analysis they provide. More recently, systems that combine both arrays, referred to as…

Audio and Speech Processing · Electrical Eng. & Systems 2024-01-09 Hai Morgenstern , Boaz Rafaely , Markus Noisternig

Sum-frequency generation (SFG) spectroscopy provides a versatile method for the investigation of non-centrosymmetric media and interfaces. Here, using tunable picosecond infrared (IR) pulses from a free-electron laser, the nonlinear optical…

Materials Science · Physics 2023-08-01 Riko Kiessling , Martin Wolf , Alexander Paarmann

This paper describes a spatial-aware speaker diarization system for the multi-channel multi-party meeting. The diarization system obtains direction information of speaker by microphone array. Speaker spatial embedding is generated by…

Audio and Speech Processing · Electrical Eng. & Systems 2022-09-27 Jie Wang , Yuji Liu , Binling Wang , Yiming Zhi , Song Li , Shipeng Xia , Jiayang Zhang , Feng Tong , Lin Li , Qingyang Hong

This contribution introduces a dataset of 7th-order Ambisonic Room Impulse Responses (HOA-RIRs), created using the Image Source Method. By employing higher-order Ambisonics, our dataset enables precise spatial audio reproduction, a critical…

Sound · Computer Science 2025-06-02 Shivam Saini , Jürgen Peissig

We propose a novel diffusion model, termed the non-identical diffusion model, and investigate its application to wireless orthogonal frequency division multiplexing (OFDM) channel generation. Unlike the standard diffusion model that uses a…

Signal Processing · Electrical Eng. & Systems 2026-01-30 Yuzhi Yang , Omar Alhussein , Mérouane Debbah

Generating the motion of orchestral conductors from a given piece of symphony music is a challenging task since it requires a model to learn semantic music features and capture the underlying distribution of real conducting motion. Prior…

Audio and Speech Processing · Electrical Eng. & Systems 2023-11-14 Zhuoran Zhao , Jinbin Bai , Delong Chen , Debang Wang , Yubo Pan

In this paper, we present HOMULA-RIR, a dataset of room impulse responses (RIRs) acquired using both higher-order microphones (HOMs) and a uniform linear array (ULA), in order to model a remote attendance teleconferencing scenario.…

Audio and Speech Processing · Electrical Eng. & Systems 2024-02-22 Federico Miotello , Paolo Ostan , Mirco Pezzoli , Luca Comanducci , Alberto Bernardini , Fabio Antonacci , Augusto Sarti

This study explores the use of text-prompted MRI image generation with the Stable Diffusion (SD) model to address challenges in acquiring real MRI datasets, such as high costs, limited rare case samples, and privacy concerns. The SD model,…

Image and Video Processing · Electrical Eng. & Systems 2025-05-30 Xinxian Fan , Mengye Lyu

Deep Brain Stimulation (DBS) is effective in treating neurological disorders but involves invasive surgery. Non-invasive DBS aims to overcome surgical risks by means of externally applied electromagnetic fields. In this work, we present a…

Medical Physics · Physics 2026-02-24 Mika Söderström , Melker Carlsson , Patrik Nicolausson , Mariana Dalarsson

A new phenomenological approach to spatially-resolved research of nonlinear (NL) microwave properties of operating thin-film superconducting resonators is proposed. The approach is based on frequency and spatial singularity of Laser…

Superconductivity · Physics 2010-08-24 Alexander P. Zhuravel , Steven M. Anlage , Alexey V. Ustinov

Automatic speech recognition (ASR) on multi-talker recordings is challenging. Current methods using 3D spatial data from multi-channel audio and visual cues focus mainly on direct waves from the target speaker, overlooking reflection wave…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-13 Yiwen Shao , Shi-Xiong Zhang , Dong Yu

We present a novel algorithm for high resolution coherent imaging of sound sources in random scattering media using time resolved measurements of the acoustic pressure at an array of receivers. The sound waves travel a long distance between…

Numerical Analysis · Mathematics 2017-12-15 Liliana Borcea , Ilker Kocyigit

We present a novel approach to improve the performance of learning-based speech dereverberation using accurate synthetic datasets. Our approach is designed to recover the reverb-free signal from a reverberant speech signal. We show that…

Audio and Speech Processing · Electrical Eng. & Systems 2022-12-13 Rohith Aralikatti , Zhenyu Tang , Dinesh Manocha

Multiple-input multiple-output (MIMO) array based millimeter-wave (MMW) imaging has a tangible prospect in applications of concealed weapons detection. A near-field imaging algorithm based on wavenumber domain processing is proposed for a…

Signal Processing · Electrical Eng. & Systems 2021-01-25 Shiyong Li , Shuoguang Wang , Moeness G. Amin , Guoqiang Zhao

An inverse source reconstruction (ISR) based 3-D near-field (NF) passive radar microwave imaging method utilizing modulated signals is presented. The modulated signals from a non-cooperative transmitter are scattered by the targets of…

Image and Video Processing · Electrical Eng. & Systems 2026-05-06 Quanfeng Wang , Alexander H. Paulus , Thomas F. Eibert

This paper proposes a neural network based speech separation method using spatially distributed microphones. Unlike with traditional microphone array settings, neither the number of microphones nor their spatial arrangement is known in…

Audio and Speech Processing · Electrical Eng. & Systems 2020-05-01 Dongmei Wang , Zhuo Chen , Takuya Yoshioka

Room Impulse Responses estimation is a fundamental problem in spatial audio processing and speech enhancement. In this paper, we build upon our previously introduced diffusion-based inpainting framework for Room Impulse Response…

Sound · Computer Science 2026-03-31 Sagi Della Torre , Mirco Pezzoli , Fabio Antonacci , Sharon Gannot

We propose a mesh-based neural network (MESH2IR) to generate acoustic impulse responses (IRs) for indoor 3D scenes represented using a mesh. The IRs are used to create a high-quality sound experience in interactive applications and audio…

Sound · Computer Science 2022-07-13 Anton Ratnarajah , Zhenyu Tang , Rohith Chandrashekar Aralikatti , Dinesh Manocha

The image-source method is widely applied to compute room impulse responses (RIRs) of shoebox rooms with arbitrary absorption. However, with increasing RIR lengths, the number of image sources grows rapidly, leading to slow computation. In…

Audio and Speech Processing · Electrical Eng. & Systems 2023-10-12 Sebastian J. Schlecht , Karolina Prawda , Rudolf Rabenstein , Maximilian Schäfer