English
Related papers

Related papers: Glottal source estimation robustness: A comparison…

200 papers

Imprecise vowel articulation can be observed in people with Parkinson's disease (PD). Acoustic features measuring vowel articulation have been demonstrated to be effective indicators of PD in its assessment. Standard clinical vowel…

Audio and Speech Processing · Electrical Eng. & Systems 2021-08-18 Yuanyuan Liu , Nelly Penttilä , Tiina Ihalainen , Juulia Lintula , Rachel Convey , Okko Räsänen

State-of-the-art under-determined audio source separation systems rely on supervised end-end training of carefully tailored neural network architectures operating either in the time or the spectral domain. However, these methods are…

Audio and Speech Processing · Electrical Eng. & Systems 2020-05-29 Vivek Narayanaswamy , Jayaraman J. Thiagarajan , Rushil Anirudh , Andreas Spanias

This work presents a method for estimation of the acoustic intensity, the energy density and the associated sound field diffuseness around the origin, when the sound field is weighted with a spatial filter. The method permits energetic DOA…

Sound · Computer Science 2016-09-14 Archontis Politis , Ville Pulkki

Based on the analysis of existing acoustic methods and instruments, a prototype of an automated instrument has been developed to perform joint measurements in situ of two parameters: sound speed and ultrasound attenuation. The device is…

Signal Processing · Electrical Eng. & Systems 2021-09-21 Aleksandr N. Grekov , Nikolay A. Grekov , Evgeniy Sychov , K. A. Kuzmin

This paper contributes to the understanding of vocal folds oscillation during phonation. In order to test theoretical models of phonation, a new experimental set-up using a deformable vocal folds replica is presented. The replica is shown…

Classical Physics · Physics 2007-10-24 Nicolas Ruty , Annemie Van Hirtum , Xavier Pelorson , Ines Lopez-Arteaga , Avraham Hirschberg

Jitter and shimmer measurements have shown to be carriers of voice quality and prosodic information which enhance the performance of tasks like speaker recognition, diarization or automatic speech recognition (ASR). However, such features…

Computation and Language · Computer Science 2021-12-22 Guillermo Cámbara , Jordi Luque , Mireia Farrús

The estimation of the decay rate of a signal section is an integral component of both blind and non-blind reverberation time estimation methods. Several decay rate estimators have previously been proposed, based on, e.g., linear regression…

Sound · Computer Science 2015-10-02 Christian Schüldt , Peter Händel

Automatic objective non-invasive detection of pathological voice based on computerized analysis of acoustic signals can play an important role in early diagnosis, progression tracking and even effective treatment of pathological voices. In…

Spotforming is a target-speaker extraction technique that uses multiple microphone arrays. This method applies beamforming (BF) to each microphone array, and the common components among the BF outputs are estimated as the target source.…

Sound · Computer Science 2024-07-15 Shoma Ayano , Li Li , Shogo Seki , Daichi Kitamura

Pre-aspiration is defined as the period of glottal friction occurring in sequences of vocalic/consonantal sonorants and phonetically voiceless obstruents. We propose two machine learning methods for automatic measurement of pre-aspiration…

Computation and Language · Computer Science 2017-06-16 Yaniv Sheena , Míša Hejná , Yossi Adi , Joseph Keshet

We report in this paper the progresses on the determination of the Boltzmann constant using the acoustic gas thermometer (AGT) of fixed-length cylindrical cavities. First, we present the comparison of the molar masses of pure argon gases…

Classical Physics · Physics 2015-01-13 X. J. Feng , J. T. Zhang , H. Lin , K. A. Gillis , M. R. Moldover

This paper presents an overview and evaluation of some of the end-to-end ASR models on long-form audios. We study three categories of Automatic Speech Recognition(ASR) models based on their core architecture: (1) convolutional, (2)…

Audio and Speech Processing · Electrical Eng. & Systems 2023-09-22 Nithin Rao Koluguri , Samuel Kriman , Georgy Zelenfroind , Somshubra Majumdar , Dima Rekesh , Vahid Noroozi , Jagadeesh Balam , Boris Ginsburg

Holter monitoring, a long-term ECG recording (24-hours and more), contains a large amount of valuable diagnostic information about the patient. Its interpretation becomes a difficult and time-consuming task for the doctor who analyzes them…

Signal Processing · Electrical Eng. & Systems 2020-11-19 Konstantin Egorov , Elena Sokolova , Manvel Avetisian , Alexander Tuzhilin

Minimum Variance Distortionless Response (MVDR) is a classical adaptive beamformer that theoretically ensures the distortionless transmission of signals in the target direction, which makes it popular in real applications. Its noise…

Sound · Computer Science 2024-09-16 Jinglin Bai , Hao Li , Xueliang Zhang , Fei Chen

The objective of deep learning methods based on encoder-decoder architectures for music source separation is to approximate either ideal time-frequency masks or spectral representations of the target music source(s). The spectral…

End-to-end Automatic Speech Recognition (ASR) systems based on neural networks have seen large improvements in recent years. The availability of large scale hand-labeled datasets and sufficient computing resources made it possible to train…

Computer Vision and Pattern Recognition · Computer Science 2023-01-05 Maxime Burchi , Radu Timofte

Automatic Speech Recognition (ASR) systems must be robust to the myriad types of noises present in real-world environments including environmental noise, room impulse response, special effects as well as attacks by malicious actors…

Sound · Computer Science 2024-09-26 Muhammad A. Shah , Bhiksha Raj

The performance of voice-controlled systems is usually influenced by accented speech. To make these systems more robust, the frontend accent recognition (AR) technologies have received increased attention in recent years. As accent is a…

Audio and Speech Processing · Electrical Eng. & Systems 2021-05-06 Zhan Zhang , Xi Chen , Yuehai Wang , Jianyi Yang

We propose a method named AudioFormer,which learns audio feature representations through the acquisition of discrete acoustic codes and subsequently fine-tunes them for audio classification tasks. Initially,we introduce a novel perspective…

Sound · Computer Science 2023-08-28 Zhaohui Li , Haitao Wang , Xinghua Jiang

Grip force is commonly used as an overall health indicator in older adults and is valuable for tracking progress in physical training and rehabilitation. Existing methods for wearable grip force measurement are cumbersome and…

Human-Computer Interaction · Computer Science 2025-07-29 Kian Mahmoodi , Yudong Xie , Tan Gemicioglu , Chi-Jung Lee , Jiwan Kim , Cheng Zhang
‹ Prev 1 4 5 6 7 8 10 Next ›