English
Related papers

Related papers: Deep Audio Zooming: Beamwidth-Controllable Neural …

200 papers

The goal of Audio-Visual Segmentation (AVS) is to localize and segment the sounding source objects from video frames. Research on AVS suffers from data scarcity due to the high cost of fine-grained manual annotations. Recent works attempt…

Computer Vision and Pattern Recognition · Computer Science 2025-05-30 Kyungbok Lee , You Zhang , Zhiyao Duan

Domain generalization aims to train models on multiple source domains so that they can generalize well to unseen target domains. Among many domain generalization methods, Fourier-transform-based domain generalization methods have gained…

Image and Video Processing · Electrical Eng. & Systems 2023-12-14 Hongyi Pan , Bin Wang , Zheyuan Zhang , Xin Zhu , Debesh Jha , Ahmet Enis Cetin , Concetto Spampinato , Ulas Bagci

Sonography techniques use multiple transducer elements for tissue visualization. Signals detected at each element are sampled prior to digital beamforming. The sampling rates required to perform high resolution digital beamforming are…

Information Theory · Computer Science 2013-07-25 Tanya Chernyakova , Yonina C. Eldar

Recent advances in Visual Anomaly Detection (VAD) have introduced sophisticated algorithms leveraging embeddings generated by pre-trained feature extractors. Inspired by these developments, we investigate the adaptation of such algorithms…

Accurately estimating and simulating the physical properties of objects from real-world sound recordings is of great practical importance in the fields of vision, graphics, and robotics. However, the progress in these directions has been…

Sound · Computer Science 2024-09-23 Xutong Jin , Chenxi Xu , Ruohan Gao , Jiajun Wu , Guoping Wang , Sheng Li

Frequency offset modulation (FOM) is proposed as a new concept to provide both high energy efficiency and high spectral efficiency for communications. In the FOM system, an array of transmitters (TXs) is deployed and only one TX is…

Information Theory · Computer Science 2016-12-22 Xihua Zou , Wei Pan , Ge Yu , Bin Luo , Lianshan Yan

The images and sounds that we perceive undergo subtle but geometrically consistent changes as we rotate our heads. In this paper, we use these cues to solve a problem we call Sound Localization from Motion (SLfM): jointly estimating camera…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Ziyang Chen , Shengyi Qian , Andrew Owens

Differential microphone arrays offer a promising solution for far-field acoustic signal acquisition due to their high spatial directivity and compact array structure. A key challenge lies in designing differential beamformers that are…

Audio and Speech Processing · Electrical Eng. & Systems 2026-02-27 Tiantian Xiong , Yongyi Deng , Kunlong Zhao , Jilu Jin , Xueqin Luo , Gongping Huang , Jingdong Chen , Jacob Benesty

Deep, high-resolution imaging is essential for unraveling biological complexity and advancing medical diagnostics, yet scattering fundamentally limits optical methods. Among the most promising approaches, photoacoustic imaging achieves…

The frequency dependent beampatterns of an active sonar projector filters the acoustic signal that is transmitted into the medium, also known as the transmit waveform. This filtering encodes information about the target's bearing relative…

Signal Processing · Electrical Eng. & Systems 2021-07-28 David A. Hague , Matthew D. Tidwell

An direction of development in the extraction of features from audio signals is based on processing raw samples in the time domain. Such an approach appears to be effective, especially in the era of neural networks. An example is SincNet.…

Sound · Computer Science 2026-04-22 Waldek Maciejko

Ambisonics is a spatial audio format describing a sound field. First-order Ambisonics (FOA) is a popular format comprising only four channels. This limited channel count comes at the expense of spatial accuracy. Ideally one would be able to…

Audio and Speech Processing · Electrical Eng. & Systems 2025-08-04 Ismael Nawfal , Symeon Delikaris Manias , Mehrez Souden , Juha Merimaa , Joshua Atkins , Elisabeth McMullin , Shadi Pirhosseinloo , Daniel Phillips

Binaural beamforming algorithms for head-mounted assistive listening devices are crucial to improve speech quality and speech intelligibility in noisy environments, while maintaining the spatial impression of the acoustic scene. While the…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-22 N. Gößling , D. Marquardt , I. Merks , T. Zhang , S. Doclo

Deep neural networks have shown promise for music audio signal processing applications, often surpassing prior approaches, particularly as end-to-end models in the waveform domain. Yet results to date have tended to be constrained by low…

Audio and Speech Processing · Electrical Eng. & Systems 2020-06-11 William Mitchell , Scott H. Hawley

The spatial covariance matrix has been considered to be significant for beamformers. Standing upon the intersection of traditional beamformers and deep neural networks, we propose a causal neural beamformer paradigm called Embedding and…

Sound · Computer Science 2021-09-03 Andong Li , Wenzhe Liu , Chengshi Zheng , Xiaodong Li

The phenomenon of the displacement of the position of the pressure, intensity and acoustic radiation force maxima along the axis of focused acoustic beams under increasing driving amplitudes (nonlinear focal shift) is studied for the case…

Classical Physics · Physics 2017-09-05 Noé Jiménez , Francisco Camarena , Nuria González-Salido

A general method for compressing the modulation time-bandwidth product of analog signals is introduced and experimentally demonstrated. As one of its applications, this physics-based signal grooming performs feature-selective stretch,…

Optics · Physics 2015-06-16 Mohammad H. Asghari , Bahram Jalali

Acoustic source localization has been applied in different fields, such as aeronautics and ocean science, generally using multiple microphones array data to reconstruct the source location. However, the model-based beamforming methods fail…

Sound · Computer Science 2022-04-01 Guanxing Zhou , Hao Liang , Xinghao Ding , Yue Huang , Xiaotong Tu , Saqlain Abbas

Recently, many deep learning based beamformers have been proposed for multi-channel speech separation. Nevertheless, most of them rely on extra cues known in advance, such as speaker feature, face image or directional information. In this…

Audio and Speech Processing · Electrical Eng. & Systems 2022-12-08 Yanjie Fu , Haoran Yin , Meng Ge , Longbiao Wang , Gaoyan Zhang , Jianwu Dang , Chengyun Deng , Fei Wang

Unmanned air vehicles often produce significant noise from their propulsion systems. Using this broadband signal as "acoustic illumination" for an auxiliary sensing system could make vehicles more robust at a minimal cost. We present an…

Robotics · Computer Science 2023-04-18 Alisha Sharma , Jason Geder , Joseph Lingevitch , Theodore Martin , Daniel Lofaro , Donald Sofge
‹ Prev 1 4 5 6 7 8 10 Next ›