English
Related papers

Related papers: ASE: Practical Acoustic Speed Estimation Beyond Do…

200 papers

In recent decades, running has become an increasingly popular pastime activity due to its accessibility, ease of practice, and anticipated health benefits. However, the risk of running-related injuries is substantial for runners of…

Sound · Computer Science 2025-04-11 Philipp Wagner , Andreas Triantafyllopoulos , Alexander Gebhard , Björn Schuller

Recent studies shows that the orthogonal time frequency space (OTFS) waveform is a promising candidate for future communication. To meet users' potential demand for Integrated Sensing and Communication (ISAC) applications in 6G, the usage…

Information Theory · Computer Science 2025-10-01 Dazhuo Wang , Yonghong Zeng , Yuhong Wang , Francois Chin , Yugang Ma , Sumei Sun

Automatic speech recognition (ASR) systems, increasingly prevalent in education, healthcare, employment, and mobile technology, face significant challenges in inclusivity, particularly for the 80 million-strong global community of people…

Computation and Language · Computer Science 2024-05-13 Dena Mujtaba , Nihar R. Mahapatra , Megan Arney , J. Scott Yaruss , Hope Gerlach-Houck , Caryn Herring , Jia Bin

This paper investigates the performance of joint time delay and Doppler-stretch estimation with the random stepp ed-frequency (RSF) signal. Applying the ambiguity function (AF) to implement the estimation, we derive the compact expressions…

Information Theory · Computer Science 2016-05-20 Tong Zhao , Zheng Nan , Tianyao Huang

Compared with automatic speech recognition (ASR), the human auditory system is more adept at handling noise-adverse situations, including environmental noise and channel distortion. To mimic this adeptness, auditory models have been widely…

Computation and Language · Computer Science 2016-09-16 Peng Dai , Xue Teng , Frank Rudzicz , Ing Yann Soon

Speech disfluencies, such as filled pauses or repetitions, are disruptions in the typical flow of speech. Stuttering is a speech disorder characterized by a high rate of disfluencies, but all individuals speak with some disfluencies and the…

Audio and Speech Processing · Electrical Eng. & Systems 2023-11-03 Amrit Romana , Kazuhito Koishida , Emily Mower Provost

Face recognition systems have become increasingly vulnerable to security threats in recent years, prompting the use of Face Anti-spoofing (FAS) to protect against various types of attacks, such as phone unlocking, face payment, and…

Computer Vision and Pattern Recognition · Computer Science 2023-09-19 Mouxiao Huang

The Future wireless communication systems face the challenging task of simultaneously providing high quality of service (QoS) and broadband data transmission, while also minimizing power consumption, latency, and system complexity. Although…

Information Theory · Computer Science 2024-10-03 Amina Darghouthi , Abdelhakim Khlifi , Hmaied Shaiek , Fatma Ben Salah , Belgacem Chibani

The promising application of femtosecond laser filamentation in atmospheric remote sensing brings imperative demand for diagnosing the spatiotemporal dynamics of filamentation. Acoustic emission (AE) during filamentation opens a door to…

Optics · Physics 2023-07-12 Binpeng Shang , Nan Zhang , Pengfei Qi , Shishi Tao , Lie Lin , Weiwei Liu

Audio-Visual Target Speaker Extraction (AVTSE) aims to isolate a target speaker's voice in a multi-speaker environment with visual cues as auxiliary. Most of the existing AVTSE methods encode visual and audio features simultaneously,…

Sound · Computer Science 2025-11-13 Zixuan Li , Xueliang Zhang , Lei Miao , Zhipeng Yan , Ying Sun , Chong Zhu

Accurate and efficient auscultation-based diagnostics are vital for early disease detection, especially in resource-limited settings where specialized clinical expertise is scarce. Traditional auscultation, which heavily depends on…

Sound · Computer Science 2025-03-26 Pingjie Wang , Liudan Zhao , Zihan Zhao , Miao He , Xin Sun , Ya Zhang , Kun Sun , Yanfeng Wang , Yu Wang

The goal of speech enhancement (SE) is to eliminate the background interference from the noisy speech signal. Generative models such as diffusion models (DM) have been applied to the task of SE because of better generalization in unseen…

Sound · Computer Science 2023-09-06 Wen Wang , Dongchao Yang , Qichen Ye , Bowen Cao , Yuexian Zou

Active headrests can reduce low-frequency noise around ears based on active noise control (ANC) system. Both the control system using fixed control filters and the remote microphone-based adaptive control system provide good noise reduction…

Computer Vision and Pattern Recognition · Computer Science 2024-01-22 Yuteng Liu , Haowen Li , Haishan Zou , Jing Lu , Zhibin Lin

We explore on various attention methods on frequency and channel dimensions for sound event detection (SED) in order to enhance performance with minimal increase in computational cost while leveraging domain knowledge to address the…

Sound · Computer Science 2023-08-30 Hyeonuk Nam , Seong-Hu Kim , Deokki Min , Yong-Hwa Park

Adversarial examples (AEs) are crafted by adding human-imperceptible perturbations to inputs such that a machine-learning based classifier incorrectly labels them. They have become a severe threat to the trustworthiness of machine learning.…

Sound · Computer Science 2019-12-05 Qiang Zeng , Jianhai Su , Chenglong Fu , Golam Kayas , Lannan Luo

Inefficient driving behaviors, such as overly conservative yielding, remain a key obstacle to deployment of autonomous vehicles (AVs). Instantaneous driving efficiency metrics are crucial for self-driving decision-making because they affect…

Robotics · Computer Science 2026-04-28 Xiaohua Zhao , Zhaowei Huang , Chen Chen , Haiyi Yang

Optical fibers have long been employed as sensors in a wide range of commercial systems. Distributed Acoustic Sensing (DAS) extends this concept by enabling the detection and localization of acoustic sources along the fiber, using…

Signal Processing · Electrical Eng. & Systems 2025-09-25 Knut H. Grythe , Jan Erik Håkegård

Conventional ultrasound (US) imaging employs the delay and sum (DAS) receive beamforming with dynamic receive focus for image reconstruction due to its simplicity and robustness. However, the DAS beamforming follows a geometrical method of…

Image and Video Processing · Electrical Eng. & Systems 2023-04-20 M. S. Asif , Gayathri Malamal , A. N. Madhavanunni , Vikram Melapudi , V Rahul , Abhijit Patil , Rajesh Langoju , Mahesh Raveendranatha Panicker

This paper considers an affine frequency division multiplexing (AFDM)-based integrated sensing and communications (ISAC) system, where the AFDM waveform is used to simultaneously carry communications information and sense targets. To…

Signal Processing · Electrical Eng. & Systems 2022-08-30 Yuanhan Ni , Zulin Wang , Peng Yuan , Qin Huang

Sound event detection and sound event localization requires different features from audio input signals. While sound event detection mainly relies on time-frequency patterns to distinguish different event classes, sound event localization…

Audio and Speech Processing · Electrical Eng. & Systems 2019-11-27 T. N. T. Nguyen , D. L. Jones , R. Ranjan , S. Jayabalan , W. S. Gan