中文
相关论文

相关论文: Histogram Layer Time Delay Neural Networks for Pas…

200 篇论文

In artificial-intelligence-aided signal processing, existing deep learning models often exhibit a black-box structure, and their validity and comprehensibility remain elusive. The integration of topological methods, despite its relatively…

机器学习 · 计算机科学 2023-11-28 Pingyao Feng , Siheng Yi , Qingrui Qu , Zhiwang Yu , Yifei Zhu

Time delay estimation (TDE) plays a key role in acoustic echo cancellation (AEC) using adaptive filter method. Considerable residual echo will be left if estimation error arises. Here, in this paper, we proposed an adaptive filter bank…

声音 · 计算机科学 2025-02-11 Lu Ma

In passive radar, a network of distributed sensors exploit signals from so-called Illuminators-of-Opportunity to detect and localize targets. We consider the case where the IO signal is available at each receiver node through a reference…

信号处理 · 电气工程与系统科学 2026-05-19 Mats Viberg , Daniele Gerosa , Tomas McKelvey , Patrik Dammert , Thomas Eriksson

This paper investigates different methods and various neural network architectures applicable in the time series classification domain. The data is obtained from a fleet of gas sensors that measure and track quantities such as oxygen and…

机器学习 · 计算机科学 2023-07-06 Mohamed Abouelnaga , Julien Vitay , Aida Farahani

Target identification of ship-radiated noise is a crucial area in underwater target recognition. However, there is currently a lack of multi-target ship datasets that accurately represent real-world underwater acoustic conditions. To…

音频与语音处理 · 电气工程与系统科学 2024-06-10 Xiaoyang Du , Feng Hong

Deep convolutional neural networks generally perform well in underwater object recognition tasks on both optical and sonar images. Many such methods require hundreds, if not thousands, of images per class to generalize well to unseen…

计算机视觉与模式识别 · 计算机科学 2020-10-27 Mateusz Ochal , Jose Vazquez , Yvan Petillot , Sen Wang

In low-visibility marine environments characterized by turbidity and darkness, acoustic cameras serve as visual sensors capable of generating high-resolution 2D sonar images. However, acoustic camera images are interfered with by complex…

计算机视觉与模式识别 · 计算机科学 2024-06-06 Xiaoteng Zhou , Katsunori Mizuno , Yilong Zhang

It is a challenging problem to detect and recognize targets on complex large-scene Synthetic Aperture Radar (SAR) images. Recently developed deep learning algorithms can automatically learn the intrinsic features of SAR images, but still…

计算机视觉与模式识别 · 计算机科学 2022-01-25 Siyan Li , Yue Xiao , Yuhang Zhang , Lei Chu , Robert C. Qiu

Continuous Sign Language Recognition (CSLR) is a crucial task for understanding the languages of deaf communities. Contemporary keypoint-based approaches typically rely on spatio-temporal encoding, where spatial interactions among keypoints…

计算机视觉与模式识别 · 计算机科学 2026-03-18 Suvajit Patra , Soumitra Samanta

Neural network based architectures used for sound recognition are usually adapted from other application domains such as image recognition, which may not harness the time-frequency representation of a signal. The ConditionaL Neural Networks…

声音 · 计算机科学 2019-04-30 Fady Medhat , David Chesmore , John Robinson

Abandoned oil and gas wells pose significant environmental risks due to the potential leakage of hydrocarbons, brine and chemical pollutants. Detecting such leaks remains extremely challenging due to the weak acoustic emission and high…

应用物理 · 物理学 2025-11-12 Guanlin Zhu , Zechun Deng , Jiaxin Shen , Junchi Yang

In this paper we present a novel algorithm and efficient data structure for anomaly detection based on temporal data. Time-series data are represented by a sequence of symbolic time intervals, describing increasing and decreasing trends, in…

数据结构与算法 · 计算机科学 2019-11-05 Roni Mateless , Michael Segal , Robert Moskovitch

Anomalous audio in speech recordings is often caused by speaker voice distortion, external noise, or even electric interferences. These obstacles have become a serious problem in some fields, such as high-quality music mixing and speech…

音频与语音处理 · 电气工程与系统科学 2021-02-11 Qiang Huang , Thomas Hain

Scene recognition, particularly for aerial and underwater images, often suffers from various types of degradation, such as blurring or overexposure. Previous works that focus on convolutional neural networks have been shown to be able to…

计算机视觉与模式识别 · 计算机科学 2024-10-22 Jianqi Zhang , Mengxuan Wang , Jingyao Wang , Lingyu Si , Changwen Zheng , Fanjiang Xu

We consider Heterogeneous Transfer Learning (HTL) from a source to a new target domain for high-dimensional regression with differing feature sets. Most homogeneous TL methods assume that target and source domains share the same feature…

机器学习 · 统计学 2025-12-02 Jae Ho Chang , Massimiliano Russo , Subhadeep Paul

Millions of hearing impaired people around the world routinely use some variants of sign languages to communicate, thus the automatic translation of a sign language is meaningful and important. Currently, there are two sub-problems in Sign…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Jie Huang , Wengang Zhou , Qilin Zhang , Houqiang Li , Weiping Li

A two space dimensional active nonlinear nonlocal cochlear model is formulated in the time domain to capture nonlinear hearing effects such as compression, multi-tone suppression and difference tones. The micromechanics of the basilar…

定量方法 · 定量生物学 2010-07-07 M. Drew LaMar , J. Xin , Y. Qi

We aim to investigate advancing the state of the art of detection, classification and localization (DCL) in the field of bioacoustics. The two primary goals are to develop transferable technologies for detection and classification in: (1)…

分布式、并行与集群计算 · 计算机科学 2016-05-06 Peter J. Dugan , Christopher W. Clark , Yann André LeCun , Sofie M. Van Parijs

Large audio language models are increasingly used for complex audio understanding tasks, but they struggle with temporal tasks that require precise temporal grounding, such as word alignment and speaker diarization. The standard approach,…

机器学习 · 计算机科学 2026-02-12 Joesph An , Phillip Keung , Jiaqi Wang , Orevaoghene Ahia , Noah A. Smith

In this paper, we present a deep neural network (DNN)-based acoustic scene classification framework. Two hierarchical learning methods are proposed to improve the DNN baseline performance by incorporating the hierarchical taxonomy…

声音 · 计算机科学 2016-08-16 Yong Xu , Qiang Huang , Wenwu Wang , Mark D. Plumbley