中文
相关论文

相关论文: HRTF measurement for accurate sound localization c…

200 篇论文

We propose a transfer learning framework for sound source reconstruction in Near-field Acoustic Holography (NAH), which adapts a well-trained data-driven model from one type of sound source to another using a physics-informed procedure. The…

音频与语音处理 · 电气工程与系统科学 2025-07-16 Xinmeng Luan , Mirco Pezzoli , Fabio Antonacci , Augusto Sarti

This paper presents a sound source localization strategy that relies on a microphone array embedded in an unmanned ground vehicle and an asynchronous close-talking microphone near the operator. A signal coarse alignment strategy is combined…

机器人学 · 计算机科学 2025-07-30 Victor Liu , Timothy Du , Jordy Sehn , Jack Collier , François Grondin

Individualized head-related impulse responses (HRIRs) enable binaural rendering, but dense per-listener measurements are costly. We address HRIR spatial up-sampling from sparse per-listener measurements: given a few measured HRIRs for a…

音频与语音处理 · 电气工程与系统科学 2026-03-31 Shaoheng Xu , Chunyi Sun , Jihui Zhang , Amy Bastine , Prasanga N. Samarasinghe , Thushara D. Abhayapala , Hongdong Li

We present the generalized iterative residual fitting (IRF) for the computation of the spherical harmonic transform (SHT) of band-limited signals on the sphere. The proposed method is based on the partitioning of the subspace of…

信息论 · 计算机科学 2017-09-11 Usama Elahi , Zubair Khalid , Rodney A. Kennedy , Jason D. McEwen

Sound event localization and detection (SELD) consists of two subtasks, which are sound event detection and direction-of-arrival estimation. While sound event detection mainly relies on time-frequency patterns to distinguish different sound…

音频与语音处理 · 电气工程与系统科学 2022-06-07 Thi Ngoc Tho Nguyen , Karn N. Watcharasupat , Ngoc Khanh Nguyen , Douglas L. Jones , Woon-Seng Gan

Sound event localization aims at estimating the positions of sound sources in the environment with respect to an acoustic receiver (e.g. a microphone array). Recent advances in this domain most prominently focused on utilizing deep…

This study introduces a novel real-time betatron tune measurement algorithm, utilizing Schottky signals and an FPGA-based backend architecture, specifically designed for rapidly ramping synchrotrons, with particular application to the…

加速器物理 · 物理学 2025-06-11 Peihan Sun , Manzhou Zhang , Renxian Yuan , Deming Li , Jian Dong , Ying Shi

Diffusion models have demonstrated remarkable success in generative tasks, including audio super-resolution (SR). In many applications like movie post-production and album mastering, substantial computational budgets are available for…

声音 · 计算机科学 2025-08-05 Yizhu Jin , Zhen Ye , Zeyue Tian , Haohe Liu , Qiuqiang Kong , Yike Guo , Wei Xue

End-to-end neural TTS training has shown improved performance in speech style transfer. However, the improvement is still limited by the training data in both target styles and speakers. Inadequate style transfer performance occurs when the…

声音 · 计算机科学 2021-06-21 Xiaochun An , Frank K. Soong , Lei Xie

Acoustic reflector localization is an important issue in audio signal processing, with direct applications in spatial audio, scene reconstruction, and source separation. Several methods have recently been proposed to estimate the 3D…

声音 · 计算机科学 2017-01-06 Luca Remaggi , Philip J. B. Jackson , Philip Coleman , Wenwu Wang

Fast Fourier Transform (FFT) relies on the HRV frequency-domain analysis techniques. It requires re-sampling of the inherently unevenly sampled heartbeat time-series (RR tachogram) to produce an evenly sampled time series of the heartbeat.…

医学物理 · 物理学 2022-08-04 Amin Gasmi

Most existing methods in binaural sound source localization rely on some kind of aggregation of phase-and level-difference cues in the time-frequency plane. While different ag-gregation schemes exist, they are often heuristic and suffer in…

声音 · 计算机科学 2016-10-03 Antoine Deleforge , Florence Forbes

Audio production style transfer is the task of processing an input to impart stylistic elements from a reference recording. Existing approaches often train a neural network to estimate control parameters for a set of audio effects. However,…

This contribution introduces a dataset of 7th-order Ambisonic Room Impulse Responses (HOA-RIRs), created using the Image Source Method. By employing higher-order Ambisonics, our dataset enables precise spatial audio reproduction, a critical…

声音 · 计算机科学 2025-06-02 Shivam Saini , Jürgen Peissig

Automatic evaluation of ST systems is typically performed by comparing translation hypotheses with one or more reference translations. While effective to some extent, this approach inherits the limitation of reference-based evaluation that…

计算与语言 · 计算机科学 2026-04-09 Mauro Cettolo , Marco Gaido , Matteo Negri , Sara Papi , Luisa Bentivogli

Terahertz (THz) wireless communication has emerged as a promising solution for future data center interconnects; however, accurate channel characterization and system-level performance evaluation in complex indoor environments remain…

信息论 · 计算机科学 2026-03-26 Mingjie Zhu , Yejian Lyu , Ziming Yu , Chong Han

Orientation and mobility (O&M) instruction for blind and low-vision learners is effective but difficult to standardize and repeat at scale due to the reliance on instructor availability, physical mock-ups, and variable real-world outdoor…

人机交互 · 计算机科学 2026-02-24 Daniel A. Muñoz

A target recognition framework relying on near-field integrated sensing and communication (ISAC) systems is proposed. By exploiting the distance-dependent spatial signatures provided by the near-field spherical wavefront, high-accuracy…

信号处理 · 电气工程与系统科学 2026-03-17 Zongyao Zhao , Zhaolin Wang , Lincong Han , Jing Jin , Kaibin Huang

Transcranial focused ultrasound stimulation (TUS) holds promise for non-invasive neural modulation in treating neurological disorders. Most clinically relevant targets are deep within the brain (near or at its geometric center), surrounded…

Semi-supervised heterogeneous domain adaptation (SHDA) addresses learning across domains with distinct feature representations and distributions, where source samples are labeled while most target samples are unlabeled, with only a small…

机器学习 · 计算机科学 2025-02-20 Yuan Yao , Xiaopu Zhang , Yu Zhang , Jian Jin , Qiang Yang