中文
相关论文

相关论文: Relative Acoustic Features for Distance Estimation…

200 篇论文

Speaker identification typically involves three stages. First, a front-end speaker embedding model is trained to embed utterance and speaker profiles. Second, a scoring function is applied between a runtime utterance and each speaker…

音频与语音处理 · 电气工程与系统科学 2022-02-22 Zhenning Tan , Yuguang Yang , Eunjung Han , Andreas Stolcke

The materials of surfaces in a room play an important room in shaping the auditory experience within them. Different materials absorb energy at different levels. The level of absorption also varies across frequencies. This paper…

音频与语音处理 · 电气工程与系统科学 2019-10-29 Constantinos Papayiannis , Christine Evers , Patrick A. Naylor

Simultaneous use of high-end wearable wireless devices like smart glasses is challenging in a dense indoor environment due to the high nature of interference. In this scenario, the millimeter wave (mmWave) band offers promising potential…

信息论 · 计算机科学 2016-06-14 Kiran Venugopal , Robert W. Heath

Accurately estimating sound source positions is crucial for robot audition. However, existing sound source localization methods typically rely on a microphone array with at least two spatially preconfigured microphones. This requirement…

机器人学 · 计算机科学 2025-06-23 Jiang Wang , Runwu Shi , Benjamin Yen , He Kong , Kazuhiro Nakadai

We introduce a novel algorithm for online estimation of acoustic impulse responses (AIRs) which allows for fast convergence by exploiting prior knowledge about the fundamental structure of AIRs. The proposed method assumes that the…

音频与语音处理 · 电气工程与系统科学 2021-05-10 Thomas Haubner , Andreas Brendel , Walter Kellermann

Spatial hearing, the brain's ability to use auditory cues to identify the origin of sounds, is crucial for everyday listening. While simplified paradigms have advanced the understanding of spatial hearing, their lack of ecological validity…

音频与语音处理 · 电气工程与系统科学 2025-10-15 Fulvio Missoni , Katarina Poole , Lorenzo Picinali , Andrea Canessa

The Relative Transfer Matrix (ReTM), recently introduced as a generalization of the relative transfer function for multiple receivers and sources, shows promising performance when applied to speech enhancement and speaker separation in…

音频与语音处理 · 电气工程与系统科学 2025-10-23 Wageesha N. Manamperi , Thushara D. Abhayapala

Based on the analysis of existing acoustic methods and instruments, a prototype of an automated instrument has been developed to perform joint measurements in situ of two parameters: sound speed and ultrasound attenuation. The device is…

信号处理 · 电气工程与系统科学 2021-09-21 Aleksandr N. Grekov , Nikolay A. Grekov , Evgeniy Sychov , K. A. Kuzmin

In this paper, we present an acoustic localization system for multiple devices. In contrast to systems which localise a device relative to one or several anchor points, we focus on the joint localisation of several devices relative to each…

网络与互联网体系结构 · 计算机科学 2015-12-17 Seyed-Mohsen Moosavi-Dezfooli , Yvonne-Anne Pignolet , Dacfey Dzung

The study and application of signal detection techniques based on cross-correlation method for acoustic transient signals in noisy and reverberant environments are presented. These techniques are shown to provide high signal to noise ratio,…

仪器与探测器 · 物理学 2015-02-19 S. Adrián-Martínez , M. Ardid , M. Bou-Cabo , I. Felis , C. Llorens , J. A. Martínez-Mora , M. Saldaña

We present AdVerb, a novel audio-visual dereverberation framework that uses visual cues in addition to the reverberant sound to estimate clean audio. Although audio-only dereverberation is a well-studied problem, our approach incorporates…

计算机视觉与模式识别 · 计算机科学 2023-08-25 Sanjoy Chowdhury , Sreyan Ghosh , Subhrajyoti Dasgupta , Anton Ratnarajah , Utkarsh Tyagi , Dinesh Manocha

In this paper, we propose a model to perform speech dereverberation by estimating its spectral magnitude from the reverberant counterpart. Our models are capable of extracting features that take into account both short and long-term…

声音 · 计算机科学 2017-11-20 Joao Felipe Santos , Tiago H. Falk

This paper addresses the challenging scenario for the distant-talking control of a music playback device, a common portable speaker with four small loudspeakers in close proximity to one microphone. The user controls the device through…

声音 · 计算机科学 2014-05-07 Ramin Pichevar , Jason Wung , Daniele Giacobello , Joshua Atkins

The reverberation time is one of the most important parameters used to characterize the acoustic property of an enclosure. In real-world scenarios, it is much more convenient to estimate the reverberation time blindly from recorded speech…

声音 · 计算机科学 2021-12-10 Kaitong Zheng , Chengshi Zheng , Jinqiu Sang , Yulong Zhang , Xiaodong Li

Dereverberation is often performed directly on the reverberant audio signal, without knowledge of the acoustic environment. Reverberation time, T60, however, is an essential acoustic factor that reflects how reverberation may impact a…

音频与语音处理 · 电气工程与系统科学 2023-02-13 Yuying Li , Yuchen Liu , Donald S. Williamson

Reverberation is present in our workplaces, our homes, concert halls and theatres. This paper investigates how deep learning can use the effect of reverberation on speech to classify a recording in terms of the room in which it was…

音频与语音处理 · 电气工程与系统科学 2020-11-03 Constantinos Papayiannis , Christine Evers , Patrick A. Naylor

This research paper presents a novel audio fingerprinting system for Automatic Content Recognition (ACR). By using signal processing techniques and statistical transformations, our proposed method generates compact fingerprints of audio…

声音 · 计算机科学 2023-05-18 Anoubhav Agarwaal , Prabhat Kanaujia , Sartaki Sinha Roy , Susmita Ghose

Respiratory rate (RR) is a key vital sign for clinical assessment and mental well-being, yet it is rarely monitored in everyday life due to the lack of unobtrusive sensing technologies. In-ear audio sensing is promising due to its high…

声音 · 计算机科学 2026-02-04 Michael Küttner , Valeria Zitz , Supraja Ramesh , Michael Beigl , Tobias Röddiger

Augmented reality devices have the potential to enhance human perception and enable other assistive functionalities in complex conversational environments. Effectively capturing the audio-visual context necessary for understanding these…

计算机视觉与模式识别 · 计算机科学 2022-01-07 Hao Jiang , Calvin Murdock , Vamsi Krishna Ithapu

Knowing the geometry of a space is desirable for many applications, e.g. sound source localization, sound field reproduction or auralization. In circumstances where only acoustic signals can be obtained, estimating the geometry of a room is…

声音 · 计算机科学 2019-07-03 Linh Nguyen , Jaime Valls Miro , Xiaojun Qiu