English
Related papers

Related papers: HRTF measurement for accurate sound localization c…

200 papers

It has been shown that light speckle fluctuations provide a means for noninvasive measurements of cerebral blood flow index (CBFi). While conventional Diffuse Correlation Spectroscopy (DCS) provides marginal brain sensitivity for CBFi in…

Sound source localisation is used in many consumer devices, to isolate audio from individual speakers and reject noise. Localization is frequently accomplished by ``beamforming'', which combines phase-shifted audio streams to increase power…

Sound · Computer Science 2025-02-13 Saeid Haghighatshoar , Dylan R Muir

Speech is a cost-effective and non-intrusive data source for identifying acute and chronic heart failure (HF). However, there is a lack of research on whether Chinese syllables contain HF-related information, as observed in other…

Audio and Speech Processing · Electrical Eng. & Systems 2025-08-22 Yue Pan , Liwei Liu , Changxin Li , Xinyao Wang , Yili Xia , Hanyue Zhang , Ming Chu

In many multi-microphone algorithms for noise reduction, an estimate of the relative transfer function (RTF) vector of the target speaker is required. The state-of-the-art covariance whitening (CW) method estimates the RTF vector as the…

Audio and Speech Processing · Electrical Eng. & Systems 2023-10-30 Wiebke Middelberg , Henri Gode , Simon Doclo

The modulation transfer function (MTF) represents the frequency domain response of imaging modalities. Here, we report a method for estimating the MTF from sample images. Test images were generated from a number of images, including those…

Image and Video Processing · Electrical Eng. & Systems 2017-12-05 Rino Saiga , Akihisa Takeuchi , Kentaro Uesugi , Yasuko Terada , Yoshio Suzuki , Ryuta Mizutani

Diffusion models have become a leading paradigm for image super-resolution (SR), but existing methods struggle to guarantee both the high-frequency perceptual quality and the low-frequency structural fidelity of generated images. Although…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Hexin Zhang , Dong Li , Jie Huang , Bingzhou Wang , Xueyang Fu , Zhengjun Zha

Deep learning-based techniques for automatic dysarthric speech detection have recently attracted interest in the research community. State-of-the-art techniques typically learn neurotypical and dysarthric discriminative representations by…

Audio and Speech Processing · Electrical Eng. & Systems 2021-10-04 Ina Kodrasi

Sound source localization (SSL) is a critical technology for determining the position of sound sources in complex environments. However, existing methods face challenges such as high computational costs and precise calibration requirements,…

Sound · Computer Science 2025-05-28 Yiyuan Yang , Shitong Xu , Niki Trigoni , Andrew Markham

This study aims to advance hardware-level computations for travel-time tomography applications in which the wavelength is close to the diameter of the information that has to be recovered. Such can be the case, for example, in the imaging…

Instrumentation and Methods for Astrophysics · Physics 2017-05-10 Mika Takala , Timo D. Hämäläinen , Sampsa Pursiainen

Recent advances in remote heart rate measurement, motivated by data-driven approaches, have notably enhanced accuracy. However, these improvements primarily focus on recovering the rPPG signal, overlooking the implicit challenges of…

Computer Vision and Pattern Recognition · Computer Science 2024-03-12 Joaquim Comas , Adria Ruiz , Federico Sukno

Future space observatories achieve detection of gravitational waves by interferometric measurements of a carrier phase, allowing to determine relative distance changes, in combination with an absolute distance measurement based on the…

Instrumentation and Methods for Astrophysics · Physics 2024-01-15 Philipp Euringer , Gerald Hechenblaikner , Francis Soualle , Walter Fichter

Convolutional neural networks (CNN) are widely used for speech emotion recognition (SER). In such cases, the short time fourier transform (STFT) spectrogram is the most popular choice for representing speech, which is fed as input to the…

Audio and Speech Processing · Electrical Eng. & Systems 2019-08-09 Shruti Gupta , Md. Shah Fahad , Akshay Deepak

In this work, we incorporated acoustically derived source features, aperiodicity, periodicity and pitch as additional targets to an acoustic-to-articulatory speech inversion (SI) system. We also propose a Temporal Convolution based SI…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-01 Yashish M. Siriwardena , Carol Espy-Wilson

Sounds, especially music, contain various harmonic components scattered in the frequency dimension. It is difficult for normal convolutional neural networks to observe these overtones. This paper introduces a multiple rates dilated causal…

Sound · Computer Science 2022-06-22 Weixing Wei , Peilin Li , Yi Yu , Wei Li

Cough is a primary symptom of most respiratory diseases, and changes in cough characteristics provide valuable information for diagnosing respiratory diseases. The characterization of cough sounds still lacks concrete evidence, which makes…

Sound · Computer Science 2023-08-08 Naveenkumar Vodnala , Pratap Reddy Lankireddy , Padmasai Yarlagadda

Developing and selecting hearing aids is a time consuming process which is simplified by using objective models. Previously, the framework for auditory discrimination experiments (FADE) accurately simulated benefits of hearing aid…

Audio and Speech Processing · Electrical Eng. & Systems 2021-02-12 David Hülsmeier , Marc René Schädler , Birger Kollmeier

Many hearing-impaired listeners struggle to localize sounds due to poor availability of binaural cues. Listeners with a cochlear implant and a contralateral hearing aid -- so-called bimodal listeners -- are amongst the worst performers, as…

Audio and Speech Processing · Electrical Eng. & Systems 2018-03-15 Benjamin Dieudonné , Tom Francart

This paper addresses the problem of multichannel online dereverberation. The proposed method is carried out in the short-time Fourier transform (STFT) domain, and for each frequency band independently. In the STFT domain, the time-domain…

Sound · Computer Science 2020-11-10 Xiaofei Li , Laurent Girin , Sharon Gannot , Radu Horaud

The binaural minimum-variance distortionless-response (BMVDR) beamformer is a well-known noise reduction algorithm that can be steered using the relative transfer function (RTF) vector of the desired speech source. Exploiting the…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-22 Nico Gößling , Wiebke Middelberg , Simon Doclo

Accurate direct measurements of far-field thermal infrared emission become increasingly important because conventional methods, relying on indirect assessments, such as reflectance/transmittance, are inaccurate or even unfeasible to…

Instrumentation and Detectors · Physics 2023-03-14 Xiu Liu , Hakan Salihoglu , Xiao Luo , Hyeong Seok Yun , Lin Jing , Bowen Yu , Sheng Shen
‹ Prev 1 8 9 10 Next ›