English
Related papers

Related papers: Computational Auditory Periphery Models: the Retur…

200 papers

Single-channel speech enhancement approaches do not always improve automatic recognition rates in the presence of noise, because they can introduce distortions unhelpful for recognition. Following a trend towards end-to-end training of…

Sound · Computer Science 2021-12-14 Peter Plantinga , Deblin Bagchi , Eric Fosler-Lussier

A soundscape is composed of three types of sound: biophony (sounds made by animals), geophony (natural abiotic sounds) and anthropophony (sounds made by humans). A key research question in the field of soundscape ecology is how these…

Spectrum analysis systems in online water quality testing are designed to detect types and concentrations of pollutants and enable regulatory agencies to respond promptly to pollution incidents. However, spectral data-based testing devices…

Machine Learning · Computer Science 2023-08-15 Haiwen Du , Zheng Ju , Yu An , Honghui Du , Dongjie Zhu , Zhaoshuo Tian , Aonghus Lawlor , Ruihai Dong

Perch 2.0 is a supervised bioacoustics foundation model pretrained on 14,597 species, including birds, mammals, amphibians, and insects, and has state-of-the-art performance on multiple benchmarks. Given that Perch 2.0 includes almost no…

Machine Learning · Computer Science 2025-12-04 Andrea Burns , Lauren Harrell , Bart van Merriënboer , Vincent Dumoulin , Jenny Hamer , Tom Denton

In this work, we propose a novel variational Bayesian adaptive learning approach for cross-domain knowledge transfer to address acoustic mismatches between training and testing conditions, such as recording devices and environmental noise.…

Audio and Speech Processing · Electrical Eng. & Systems 2025-01-28 Hu Hu , Sabato Marco Siniscalchi , Chao-Han Huck Yang , Chin-Hui Lee

This paper introduces a novel framework integrating nonlinear acoustic computing and reinforcement learning to enhance advanced human-robot interaction under complex noise and reverberation. Leveraging physically informed wave equations…

Robotics · Computer Science 2025-05-07 Xiaoliang Chen , Xin Yu , Le Chang , Yunhe Huang , Jiashuai He , Shibo Zhang , Jin Li , Likai Lin , Ziyu Zeng , Xianling Tu , Shuyu Zhang

Cochlear wavenumber and impedance are mechanistic variables that encode information regarding how the cochlea works - specifically wave propagation and Organ of Corti dynamics. These mechanistic variables underlie interesting features of…

Quantitative Methods · Quantitative Biology 2024-07-02 Samiya A Alkhairy

There is a significant need for precise and reliable forecasting of the far-field noise emanating from shipping vessels. Conventional full-order models based on the Navier-Stokes equations are unsuitable, and sophisticated model reduction…

Machine Learning · Computer Science 2024-04-15 Indu Kant Deo , Akash Venkateshwaran , Rajeev K. Jaiman

Acoustic data provide scientific and engineering insights in fields ranging from bioacoustics and communications to ocean and earth sciences. In this review, we survey recent advances and the transformative potential of machine learning…

Sound · Computer Science 2025-07-08 Ryan A. McCarthy , You Zhang , Samuel A. Verburg , William F. Jenkins , Peter Gerstoft

We develop an analytic model of the mammalian cochlea. We use a mixed physical-phenomenological approach by utilizing existing work on the physics of classical box-representations of the cochlea, and behavior of recent data-derived…

Tissues and Organs · Quantitative Biology 2021-08-20 Samiya A Alkhairy , Christopher A Shera

Machine learning techniques are an active area of research for speech enhancement for hearing aids, with one particular focus on improving the intelligibility of a noisy speech signal. Recent work has shown that feature encodings from…

Sound · Computer Science 2024-07-19 Robert Sutherland , George Close , Thomas Hain , Stefan Goetze , Jon Barker

Most sounds of interest consist of complex, time-dependent admixtures of tones of diverse frequencies and variable amplitudes. To detect and process these signals, the ear employs a highly nonlinear, adaptive, real-time spectral analyzer:…

Neurons and Cognition · Quantitative Biology 2014-08-12 T. Reichenbach , A. J. Hudspeth

Temporal coding is one approach to representing information in spiking neural networks. An example of its application is the location of sounds by barn owls that requires especially precise temporal coding. Dependent upon the azimuthal…

Neurons and Cognition · Quantitative Biology 2014-01-24 Thomas Pfeil , Anne-Christine Scherzer , Johannes Schemmel , Karlheinz Meier

Current transfer learning methods for high-dimensional linear regression assume feature alignment across domains, restricting their applicability to semantically matched features. In many real-world scenarios, however, distinct features in…

Methodology · Statistics 2025-12-29 Jiancheng Jiang , Xuejun Jiang , Hongxia Jin

This computer science master thesis aims at modelling the nonlinearities of a loudspeaker. A piecewise linear approximation is initially explored and then we present a nonlinear Volterra model to simulate the behavior of the system. The…

Sound · Computer Science 2017-03-02 Alessandro Loriga

Aligning acoustic and linguistic representations is a central challenge to bridge the pre-trained models in knowledge transfer for automatic speech recognition (ASR). This alignment is inherently structured and asymmetric: while multiple…

Computation and Language · Computer Science 2026-03-06 Xugang Lu , Peng Shen , Hisashi Kawai

Visual-to-auditory sensory substitution devices can assist the blind in sensing the visual environment by translating the visual information into a sound pattern. To improve the translation quality, the task performances of the blind are…

Computer Vision and Pattern Recognition · Computer Science 2019-04-22 Di Hu , Dong Wang , Xuelong Li , Feiping Nie , Qi Wang

We propose WHISPER-GPT: A generative large language model (LLM) for speech and music that allows us to work with continuous audio representations and discrete tokens simultaneously as part of a single architecture. There has been a huge…

Sound · Computer Science 2024-12-20 Prateek Verma

In their everyday life, the speech recognition performance of human listeners is influenced by diverse factors, such as the acoustic environment, the talker and listener positions, possibly impaired hearing, and optional hearing devices.…

Audio and Speech Processing · Electrical Eng. & Systems 2021-04-02 Marc René Schädler

Deep Neural Networks (DNN) have been successful in en- hancing noisy speech signals. Enhancement is achieved by learning a nonlinear mapping function from the features of the corrupted speech signal to that of the reference clean speech…

Machine Learning · Computer Science 2016-06-16 Zhenzhou Wu , Sunil Sivadas , Yong Kiam Tan , Ma Bin , Rick Siow Mong Goh
‹ Prev 1 8 9 10 Next ›