English
Related papers

Related papers: Multispectral representation of Distributed Acoust…

200 papers

Active speaker detection requires a solid integration of multi-modal cues. While individual modalities can approximate a solution, accurate predictions can only be achieved by explicitly fusing the audio and visual features and modeling…

Computer Vision and Pattern Recognition · Computer Science 2021-10-06 Juan León-Alcázar , Fabian Caba Heilbron , Ali Thabet , Bernard Ghanem

Speech enhancement and speech separation are two related tasks, whose purpose is to extract either one or more target speech signals, respectively, from a mixture of sounds generated by several sources. Traditionally, these tasks have been…

Audio and Speech Processing · Electrical Eng. & Systems 2021-03-16 Daniel Michelsanti , Zheng-Hua Tan , Shi-Xiong Zhang , Yong Xu , Meng Yu , Dong Yu , Jesper Jensen

Nowadays, long distance optical fibre transmission systems use polarization diversity multiplexed signals to enhance transmission performance. Distributed acoustic sensors (DAS) use the same propagation medium ie. single mode optical fibre,…

Signal Processing · Electrical Eng. & Systems 2020-07-03 Sterenn Guerrier , Christian Dorize , Elie Awwad , Jérémie Renaudier

The increasing demand for reliable connectivity in industrial environments necessitates effective spectrum utilization strategies, especially in the context of shared spectrum bands. However, the dynamic spectrum-sharing mechanisms often…

Systems and Control · Electrical Eng. & Systems 2025-04-03 Sicheng Liu , Qun Wang , Zhuwei Qin , Weishan Zhang , Jingyi Wang , Xiang Ma

A pooling mechanism is essential for mean opinion score (MOS) prediction, facilitating the transformation of variable-length audio features into a concise fixed-size representation that effectively encodes speech quality. Existing pooling…

Sound · Computer Science 2025-09-01 Cheng-Yeh Yang , Kuan-Tang Huang , Chien-Chun Wang , Hung-Shin Lee , Hsin-Min Wang , Berlin Chen

In this paper, we study data-aided sensing (DAS) for distributed detection in wireless sensor networks (WSNs) when sensors' measurements are correlated. In particular, we derive a node selection criterion based on the J-divergence in DAS…

Information Theory · Computer Science 2020-11-18 Jinho Choi

Multi-source data classification is a critical yet challenging task for remote sensing image interpretation. Existing methods lack adaptability to diverse land cover types when modeling frequency domain features. To this end, we propose a…

Image and Video Processing · Electrical Eng. & Systems 2025-07-08 Yikang Zhao , Feng Gao , Xuepeng Jin , Junyu Dong , Qian Du

Speech signals encode emotional, linguistic, and pathological information within a shared acoustic channel; however, disentanglement is typically assessed indirectly through downstream task performance. We introduce an information-theoretic…

Sound · Computer Science 2026-02-25 Bipasha Kashyap , Björn W. Schuller , Pubudu N. Pathirana

Purpose: We previously established an open-access lung sound database, HF_Lung_V1, and developed deep learning models for inhalation, exhalation, continuous adventitious sound (CAS), and discontinuous adventitious sound (DAS) detection. The…

Animals hear and vocalize across frequency ranges that differ substantially from humans, often extending into the ultrasonic domain. Yet most computational bioacoustics systems rely on audio models pre-trained at 16 kHz, restricting their…

Long-range Rayleigh-based Distributed Acoustic Sensing (DAS) systems are often limited in their sensitivity and bandwidth. The former limitation is a result of the low backscattered power and poor \textit{dynamic-strain to optical-phase}…

Signal Processing · Electrical Eng. & Systems 2020-09-17 Nadav Arbel , Lihi Shiloh , Nadav Levanon , Avishay Eyal

Audio-visual segmentation (AVS) aims to segment the sounding objects in video frames. Although great progress has been witnessed, we experimentally reveal that current methods reach marginal performance gain within the use of the unlabeled…

Computer Vision and Pattern Recognition · Computer Science 2024-03-19 Jinxiang Liu , Yikun Liu , Fei Zhang , Chen Ju , Ya Zhang , Yanfeng Wang

Adaptive sampling results in dramatic improvements in the recovery of sparse signals in white Gaussian noise. A sequential adaptive sampling-and-refinement procedure called Distilled Sensing (DS) is proposed and analyzed. DS is a form of…

Statistics Theory · Mathematics 2010-05-31 Jarvis Haupt , Rui Castro , Robert Nowak

Sound-guided object segmentation has drawn considerable attention for its potential to enhance multimodal perception. Previous methods primarily focus on developing advanced architectures to facilitate effective audio-visual interactions,…

Sound · Computer Science 2025-03-18 Chen Liu , Liying Yang , Peike Li , Dadong Wang , Lincheng Li , Xin Yu

The process of analyzing audio signals in search of cetacean vocalizations is in many cases a very arduous task, requiring many complex computations, a plethora of digital processing techniques and the scrutinization of an audio signal with…

Sound · Computer Science 2022-03-22 Jacques van Wyk , Jaco Versfeld , Johan du Preez

Speaker diarization systems are challenged by a trade-off between the temporal resolution and the fidelity of the speaker representation. By obtaining a superior temporal resolution with an enhanced accuracy, a multi-scale approach is a way…

Audio and Speech Processing · Electrical Eng. & Systems 2022-03-31 Tae Jin Park , Nithin Rao Koluguri , Jagadeesh Balam , Boris Ginsburg

Wireless sensor networks consist of sensor nodes that are physically distributed over different locations. Spatial filtering procedures exploit the spatial correlation across these sensor signals to fuse them into a filtered signal…

Signal Processing · Electrical Eng. & Systems 2022-11-04 Cem Ates Musluoglu , Alexander Bertrand

Automated monitoring of marine mammals in the St. Lawrence Estuary faces extreme challenges: calls span low-frequency moans to ultrasonic clicks, often overlap, and are embedded in variable anthropogenic and environmental noise. We…

Audio and Speech Processing · Electrical Eng. & Systems 2025-11-03 Amine Razig , Youssef Soulaymani , Loubna Benabbou , Pierre Cauchy

Automatic Speech Recognition (ASR) systems must be robust to the myriad types of noises present in real-world environments including environmental noise, room impulse response, special effects as well as attacks by malicious actors…

Sound · Computer Science 2024-09-26 Muhammad A. Shah , Bhiksha Raj

Deformable medical image registration is a crucial aspect of medical image analysis. In recent years, researchers have begun leveraging auxiliary tasks (such as supervised segmentation) to provide anatomical structure information for the…

Computer Vision and Pattern Recognition · Computer Science 2024-10-01 Hongchao Zhou , Shunbo Hu
‹ Prev 1 4 5 6 7 8 10 Next ›