中文
相关论文

相关论文: On Time Delay Interpolation for Improved Acoustic …

200 篇论文

Sound source localization (SSL) demonstrates remarkable results in controlled settings but struggles in real-world deployment due to dual imbalance challenges: intra-task imbalance arising from long-tailed direction-of-arrival (DoA)…

声音 · 计算机科学 2026-01-27 Zexia Fan , Yu Chen , Qiquan Zhang , Kainan Chen , Xinyuan Qian

Stereo matching provides depth estimation from binocular images for downstream applications. These applications mostly take video streams as input and require temporally consistent depth maps. However, existing methods mainly focus on the…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Jiaxi Zeng , Chengtang Yao , Yuwei Wu , Yunde Jia

We introduce a novel direct calibration algorithm to address phase delay, gain, and offset mismatches in Analog-to-Digital Converter (ADC) time interleaving systems. These mismatches, common in high-speed data acquisition, degrade system…

天体物理仪器与方法 · 物理学 2025-11-27 Chi-kwan Chan , Hina Suzuki , David Forbes , Andrew Thomas West , Arash Roshanineshat , Daniel P. Marrone , Amy Lowitz

Spatial and temporal delays in a wireless multi-antenna system, paired with an orthogonal frequency division multiplexing (OFDM) waveform, can be utilized to estimate the Angle of Arrival (AoA) and Time of Arrival (ToA) of scatterers in the…

信号处理 · 电气工程与系统科学 2025-03-25 Chandrashekhar Rai , Debarati Sen

Acoustic local positioning systems (ALPSs) are an interesting alternative for indoor positioning due to certain advantages over other approaches, including their relatively high accuracy, low cost, and room-level signal propagation.…

Speaker identification typically involves three stages. First, a front-end speaker embedding model is trained to embed utterance and speaker profiles. Second, a scoring function is applied between a runtime utterance and each speaker…

音频与语音处理 · 电气工程与系统科学 2022-02-22 Zhenning Tan , Yuguang Yang , Eunjung Han , Andreas Stolcke

Dispersion scan is a self-referenced measurement technique for ultrashort pulses. Similar to frequency-resolved optical gating, the dispersion scan technique records the dependence of nonlinearly generated spectra as a function of a…

仪器与探测器 · 物理学 2018-02-14 Esmerando Escoto , Tamas Nagy , Ayhan Tajalli , Günter Steinmeyer

In this study, we propose a dense frequency-time attentive network (DeFT-AN) for multichannel speech enhancement. DeFT-AN is a mask estimation network that predicts a complex spectral masking pattern for suppressing the noise and…

音频与语音处理 · 电气工程与系统科学 2023-03-07 Dongheon Lee , Jung-Woo Choi

The success of deep learning-based speaker verification systems is largely attributed to access to large-scale and diverse speaker identity data. However, collecting data from more identities is expensive, challenging, and often limited by…

音频与语音处理 · 电气工程与系统科学 2025-08-27 Tianchi Liu , Ruijie Tao , Qiongqiong Wang , Yidi Jiang , Hardik B. Sailor , Ke Zhang , Jingru Lin , Haizhou Li

Distributed antenna arrays have been proposed for many applications ranging from space-based observatories to automated vehicles. Achieving good performance in distributed antenna systems requires stringent synchronization at the wavelength…

信号处理 · 电气工程与系统科学 2022-12-26 Jason M. Merlo , Serge R. Mghabghab , Jeffrey A. Nanzer

Multiple wireless sensing tasks, e.g., radar detection for driver safety, involve estimating the "channel" or relationship between signal transmitted and received. In this work, we focus on a certain channel model known as the delay-doppler…

信息论 · 计算机科学 2020-11-24 Alisha Zachariah

Computational time reversal imaging can be used to locate the position of multiple scatterers in a known background medium. Here, we discuss a sparse approximation method for computational time-reversal imaging. The method is formulated…

其他定量生物学 · 定量生物学 2009-04-23 M. Andrecut

In this paper, we propose an effective sound event detection (SED) method based on the audio spectrogram transformer (AST) model, pretrained on the large-scale AudioSet for audio tagging (AT) task, termed AST-SED. Pretrained AST models have…

音频与语音处理 · 电气工程与系统科学 2023-03-08 Kang Li , Yan Song , Li-Rong Dai , Ian McLoughlin , Xin Fang , Lin Liu

Different machines can exhibit diverse frequency patterns in their emitted sound. This feature has been recently explored in anomaly sound detection and reached state-of-the-art performance. However, existing methods rely on the manual or…

声音 · 计算机科学 2023-09-07 Hejing Zhang , Jian Guan , Qiaoxi Zhu , Feiyang Xiao , Youde Liu

The Wigner-Smith (WS) time delay matrix relates a lossless system's scattering matrix to its frequency derivative. First proposed in the realm of quantum mechanics to characterize time delays experienced by particles during a collision,…

计算工程、金融与科学 · 计算机科学 2023-05-17 Utkarsh R. Patel , Yiqian Mao , Eric Michielssen

Audio embeddings enable large scale comparisons of the similarity of audio files for applications such as search and recommendation. Due to the subjectivity of audio similarity, it can be desirable to design systems that answer not only…

Time delay interferometry (TDI) is a key technique employed in gravitational wave (GW) space missions to mitigate laser frequency noise by combining multiple laser links and establishing an equivalent equal arm interferometry. The null…

广义相对论与量子宇宙学 · 物理学 2024-09-10 Gang Wang

This article presents a new approach for the wireless clock synchronization of Decawave ultra-wideband transceivers based on the time difference of arrival. The presented techniques combine the time-of-arrival and time-difference-of-arrival…

信号处理 · 电气工程与系统科学 2019-03-05 Juri Sidorenko , Volker Schatz , Norbert Scherer-Negenborn , Michael Arens , Urs Hugentobler

Conventional approaches to sound localization and separation are based on microphone arrays in artificial systems. Inspired by the selective perception of human auditory system, we design a multi-source listening system which can separate…

声音 · 计算机科学 2019-11-11 Xuecong Sun , Han Jia , Zhe Zhang , Yuzhen Yang , Zhaoyong Sun , Jun Yang

Binaural target sound extraction (TSE) aims to extract a desired sound from a binaural mixture of arbitrary sounds while preserving the spatial cues of the desired sound. Indeed, for many applications, the target sound signal and its…