English
Related papers

Related papers: Speech Polarity Detection Using Hilbert Phase Info…

200 papers

We use a publicly available numerical wave-propagation simulation of Hartlep et al. 2011 to test the ability of helioseismic holography to detect signatures of a compact, fully submerged, 5% sound-speed perturbation placed at a depth of 50…

Solar and Stellar Astrophysics · Physics 2015-06-11 Douglas C. Braun

With the forthcoming release of high precision polarization measurements, such as from the Planck satellite, it becomes critical to evaluate the performance of estimators for the polarization fraction and angle. These two physical…

Instrumentation and Methods for Astrophysics · Physics 2015-02-11 L. Montier , S. Plaszczynski , F. Levrier , M. Tristram , D. Alina , I. Ristorcelli , J. -P. Bernard , V. Guillet

Many applications of speech communication and speaker identification suffer from the problem of co-channel speech. This paper deals with a multi-resolution dyadic wavelet transform method for usable segments of co-channel speech detection…

Sound · Computer Science 2013-01-03 Wajdi Ghezaiel , Amel Ben Slimane Rahmouni , Ezzedine Ben Braiek

Methods that can generate synthetic speech which is perceptually indistinguishable from speech recorded by a human speaker, are easily available. Several incidents report misuse of synthetic speech generated from these methods to commit…

Computer Vision and Pattern Recognition · Computer Science 2024-04-18 Amit Kumar Singh Yadav , Kratika Bhagtani , Davide Salvi , Paolo Bestagini , Edward J. Delp

Self-supervised learning of speech representations from large amounts of unlabeled data has enabled state-of-the-art results in several speech processing tasks. Aggregating these speech representations across time is typically approached by…

Audio and Speech Processing · Electrical Eng. & Systems 2022-10-19 Themos Stafylakis , Ladislav Mosner , Sofoklis Kakouros , Oldrich Plchot , Lukas Burget , Jan Cernocky

Higher criticism is a method for detecting signals that are both sparse and weak. Although first proposed in cases where the noise variables are independent, higher criticism also has reasonable performance in settings where those variables…

Statistics Theory · Mathematics 2010-10-05 Peter Hall , Jiashun Jin

Previous speech pre-training methods, such as wav2vec2.0 and HuBERT, pre-train a Transformer encoder to learn deep representations from audio data, with objectives predicting either elements from latent vector quantized space or…

Sound · Computer Science 2022-04-08 Shuo Ren , Shujie Liu , Yu Wu , Long Zhou , Furu Wei

This study demonstrates whether financial text is useful for tactical asset allocation using stocks by using natural language processing to create polarity indexes in financial news. In this study, we performed clustering of the created…

Computational Engineering, Finance, and Science · Computer Science 2024-08-14 Rei Taguchi , Hiroki Sakaji , Kiyoshi Izumi

We design statistical hypothesis tests for performing leak detection in water pipeline channels. By applying an appropriate model for signal propagation, we show that the detection problem becomes one of distinguishing signal from noise,…

Signal Processing · Electrical Eng. & Systems 2022-10-25 Liusha Yang , Matthew R. McKay , Xun Wang

Autonomous robotics is critically affected by the robustness of its scene understanding algorithms. We propose a two-axis pipeline based on polarization indices to analyze dynamic urban scenes. As robots evolve in unknown environments, they…

Computer Vision and Pattern Recognition · Computer Science 2021-06-04 Marc Blanchon , Désiré Sidibé , Olivier Morel , Ralph Seulin , Fabrice Meriaudeau

This paper proposes attentive statistics pooling for deep speaker embedding in text-independent speaker verification. In conventional speaker embedding, frame-level features are averaged over all the frames of a single utterance to form an…

Audio and Speech Processing · Electrical Eng. & Systems 2019-02-27 Koji Okabe , Takafumi Koshinaka , Koichi Shinoda

We have studied the implications of high sensitivity polarization measurements of objects from the WMAP point source catalogue made using the VLA at 8.4, 22 and 43 GHz. The fractional polarization of sources is almost independent of…

Cosmology and Nongalactic Astrophysics · Physics 2015-05-18 R. A. Battye , I. W. A. Browne , M. W. Peel , N. J. Jackson , C. Dickinson

Speech self-supervised models such as wav2vec 2.0 and HuBERT are making revolutionary progress in Automatic Speech Recognition (ASR). However, they have not been totally proven to produce better performance on tasks other than ASR. In this…

Computation and Language · Computer Science 2022-10-05 Yingzhi Wang , Abdelmoumene Boumadane , Abdelwahab Heba

Rhythmic activity is ubiquitous in biological systems from the cellular to organism level. Reconstructing the instantaneous phase is the first step in analyzing the essential mechanism leading to a synchronization state from the observed…

Adaptation and Self-Organizing Systems · Physics 2022-09-02 Akari Matsuki , Hiroshi Kori , Ryota Kobayashi

We introduce a new automatic evaluation method for speaker similarity assessment, that is consistent with human perceptual scores. Modern neural text-to-speech models require a vast amount of clean training data, which is why many solutions…

Sound · Computer Science 2022-07-04 Deja Kamil , Sanchez Ariadna , Roth Julian , Cotescu Marius

This paper presents a self-supervised method for visual detection of the active speaker in a multi-person spoken interaction scenario. Active speaker detection is a fundamental prerequisite for any artificial cognitive system attempting to…

Computer Vision and Pattern Recognition · Computer Science 2019-07-19 Kalin Stefanov , Jonas Beskow , Giampiero Salvi

With increasing globalization and immigration, various studies have estimated that about half of the world population is bilingual. Consequently, individuals concurrently use two or more languages or dialects in casual conversational…

Computation and Language · Computer Science 2022-11-01 Saurav K. Aryal , Howard Prioleau , Gloria Washington

This paper proposes a novel Wavelet Packet based feature extraction approach for the task of text independent speaker recognition. The features are extracted by using the combination of Mel Frequency Cepstral Coefficient (MFCC) and Wavelet…

Informed speaker extraction aims to extract a target speech signal from a mixture of sources given prior knowledge about the desired speaker. Recent deep learning-based methods leverage a speaker discriminative model that maps a reference…

Audio and Speech Processing · Electrical Eng. & Systems 2022-02-17 Mohamed Elminshawi , Wolfgang Mack , Emanuël A. P. Habets

Parkinson's Disease (PD) affects over 10 million people worldwide, with speech impairments in up to 89% of patients. Current speech-based detection systems analyze entire utterances, potentially overlooking the diagnostic value of specific…

Computation and Language · Computer Science 2025-10-07 Ilias Tougui , Mehdi Zakroum , Mounir Ghogho