English
Related papers

Related papers: SUBARU: A Practical Approach to Power Saving in He…

200 papers

On-device directional hearing requires audio source separation from a given direction while achieving stringent human-imperceptible latency requirements. While neural nets can achieve significantly better performance than traditional…

Sound · Computer Science 2021-12-14 Anran Wang , Maruchi Kim , Hao Zhang , Shyamnath Gollakota

Due to the lack of target speech annotations in real-recorded far-field conversational datasets, speech enhancement (SE) models are typically trained on simulated data. However, the trained models often perform poorly in real-world…

Sound · Computer Science 2025-06-24 Longjie Luo , Lin Li , Qingyang Hong

As technology grows, higher frequency signals are required to be processed in various applications. In order to digitize such signals, conventional analog to digital convertors are facing implementation challenges due to the higher sampling…

Information Theory · Computer Science 2014-11-27 Amir Zandieh , Alireza Zareian , Masoumeh Azghani , Farokh Marvasti

Noise robustness is essential for deploying automatic speech recognition (ASR) systems in real-world environments. One way to reduce the effect of noise interference is to employ a preprocessing module that conducts speech enhancement, and…

Beamforming is a powerful tool designed to enhance speech signals from the direction of a target source. Computing the beamforming filter requires estimating spatial covariance matrices (SCMs) of the source and noise signals. Time-frequency…

Audio and Speech Processing · Electrical Eng. & Systems 2022-05-10 Tsubasa Ochiai , Marc Delcroix , Tomohiro Nakatani , Shoko Araki

Cognitive Radio requires efficient and reliable spectrum sensing of wideband signals. In order to cope with the sampling rate bottleneck, new sampling methods have been proposed that sample below the Nyquist rate. However, such techniques…

Information Theory · Computer Science 2017-04-26 Deborah Cohen , Yonina C. Eldar

Outdoor acoustic events detection is an exciting research field but challenged by the need for complex algorithms and deep learning techniques, typically requiring many computational, memory, and energy resources. This challenge discourages…

Audio and Speech Processing · Electrical Eng. & Systems 2020-01-30 Gianmarco Cerutti , Rahul Prasad , Alessio Brutti , Elisabetta Farella

Speaker Recognition and Speaker Identification are challenging tasks with essential applications such as automation, authentication, and security. Deep learning approaches like SincNet and AM-SincNet presented great results on these tasks.…

Sound · Computer Science 2020-10-20 João Antônio Chagas Nunes , David Macêdo , Cleber Zanchettin

In this paper, a binaural beamforming algorithm for hearing aid applications is introduced.The beamforming algorithm is designed to be robust to some error in the estimate of the target speaker direction. The algorithm has two main…

Audio and Speech Processing · Electrical Eng. & Systems 2019-11-21 Hala As'ad , Martin Bouchard , Homayoun Kamkar-Parsi

While self-supervised learning (SSL) has revolutionized audio representation, the excessive parameterization and quadratic computational cost of standard Transformers limit their deployment on resource-constrained devices. To address this…

Sound · Computer Science 2026-03-30 Harunori Kawano , Takeshi Sasaki

Compressed Sensing MRI reconstructs images of the body's internal anatomy from undersampled measurements, thereby reducing scan time. Recently, deep learning has shown great potential for reconstructing high-fidelity images from highly…

Image and Video Processing · Electrical Eng. & Systems 2025-04-07 Armeet Singh Jatyani , Jiayun Wang , Aditi Chandrashekar , Zihui Wu , Miguel Liu-Schiaffini , Bahareh Tolooshams , Anima Anandkumar

Sound event detection (SED) is a hot topic in consumer and smart city applications. Existing approaches based on Deep Neural Networks are very effective, but highly demanding in terms of memory, power, and throughput when targeting…

Machine Learning · Computer Science 2021-01-13 Gianmarco Cerutti , Renzo Andri , Lukas Cavigelli , Michele Magno , Elisabetta Farella , Luca Benini

With recent research advancements, deep learning models are becoming attractive and powerful choices for speech enhancement in real-time applications. While state-of-the-art models can achieve outstanding results in terms of speech quality…

Audio and Speech Processing · Electrical Eng. & Systems 2021-05-20 Sebastian Braun , Hannes Gamper , Chandan K. A. Reddy , Ivan Tashev

This paper introduces a novel Soft Acoustic Curvature (SAC) sensor. SAC incorporates integrated audio components and features an acoustic channel within a flexible structure. A reference acoustic wave, generated by a speaker at one end of…

Sound · Computer Science 2024-09-30 Mohammad Sheikh Sofla , Hanita Golshanian , Vishnu Rajendran S , Amir Ghalamzan E

Particle beam microscopy (PBM) performs nanoscale imaging by pixelwise capture of scalar values representing noisy measurements of the response from secondary electrons (SEs) integrated over a dwell time. Extended to metrology, goals…

Data Analysis, Statistics and Probability · Physics 2023-08-15 Akshay Agarwal , Minxu Peng , Vivek K. Goyal

We present a method to maintain the subjective perception of volume of audio signals and, at the same time, reduce their absolute peak value. We focus on achieving this without compromising the perceived audio quality. This is specially…

Signal Processing · Electrical Eng. & Systems 2022-02-17 A. Jeannerot , N. de Koeijer , P. Martínez-Nuevo , M. B. Møller , J. Dyreby , P. Prandoni

Self-supervised language models are very effective at predicting high-level cortical responses during language comprehension. However, the best current models of lower-level auditory processing in the human brain rely on either…

Computation and Language · Computer Science 2022-05-31 Aditya R. Vaidya , Shailee Jain , Alexander G. Huth

Dysarthric speech reconstruction (DSR) aims to transform dysarthric speech into normal speech by improving the intelligibility and naturalness. This is a challenging task especially for patients with severe dysarthria and speaking in…

Sound · Computer Science 2024-02-01 Xueyuan Chen , Yuejiao Wang , Xixin Wu , Disong Wang , Zhiyong Wu , Xunying Liu , Helen Meng

We present a sub-Nyquist analog-to-digital converter of wideband inputs. Our circuit realizes the recently proposed modulated wideband converter, which is a flexible platform for sampling signals according to their actual bandwidth…

Cellular Automata and Lattice Gases · Physics 2009-12-15 Moshe Mishali , Yonina C. Eldar , Oleg Dounaevsky , Eli Shoshan

Automatic Speech Recognition (ASR) systems are pivotal in transcribing speech into text, yet the errors they introduce can significantly degrade the performance of downstream tasks like summarization. This issue is particularly pronounced…

‹ Prev 1 8 9 10 Next ›