English
Related papers

Related papers: Glottal Closure and Opening Instant Detection from…

200 papers

We examine an analytic variational inference scheme for the Gaussian Process State Space Model (GPSSM) - a probabilistic model for system identification and time-series modelling. Our approach performs variational inference over both the…

Machine Learning · Statistics 2018-12-11 Alessandro Davide Ialongo , Mark van der Wilk , Carl Edward Rasmussen

With the development of internet of things technologies, tremendous sensor audio data has been produced, which poses great challenges to audio-based event detection in smart cities. In this paper, we target a challenging audio-based event…

Sound · Computer Science 2023-12-27 Haoyu Tang , Yunxiao Wang , Jihua Zhu , Shuaike Zhang , Mingzhu Xu , Qinghai Zheng , Yupeng Hu

This paper presents a novel method for fault detection in vibration/acoustic signals contaminated with non-Gaussian noise, specifically addressing the challenge of random impulsive and wideband disturbances in industrial measurements. While…

Signal Processing · Electrical Eng. & Systems 2025-02-18 A Drewnicka , A Michalak , R Zimroz , A Kumar , A Wyłomańska , J Wodecki

We present an algorithm for the identification of transient noise artifacts (glitches) in cross-correlation searches for long O(10s) gravitational-wave transients. The algorithm utilizes the auto-power in each detector as a discriminator…

Instrumentation and Methods for Astrophysics · Physics 2013-09-05 Tanner Prestegard , Eric Thrane , Nelson L. Christensen , Michael W. Coughlin , Ben Hubbert , Shivaraj Kandhasamy , Evan MacAyeal , Vuk Mandic

Recently, increasing attention has been directed to the study of the speech emotion recognition, in which global acoustic features of an utterance are mostly used to eliminate the content differences. However, the expression of speech…

Human-Computer Interaction · Computer Science 2018-11-21 Haotian Guan , Zhilei Liu , Longbiao Wang , Jianwu Dang , Ruiguo Yu

Time- and pitch-scale modifications of speech signals find important applications in speech synthesis, playback systems, voice conversion, learning/hearing aids, etc.. There is a requirement for computationally efficient and real-time…

Audio and Speech Processing · Electrical Eng. & Systems 2018-01-22 Sunil Rudresh , Aditya Vasisht , Karthika Vijayan , Chandra Sekhar Seelamantula

This paper presents the Multimodal Laryngoscopic Video Analyzing System (MLVAS), a novel system that leverages both audio and video data to automatically extract key video segments and metrics from raw laryngeal videostroboscopic videos for…

Sound · Computer Science 2026-03-10 Yucong Zhang , Xin Zou , Jinshan Yang , Wenjun Chen , Juan Liu , Faya Liang , Ming Li

We study a well-known estimator of the fractal index of a stochastic process. Our framework is very general and encompasses many models of interest; we show how to extend the theory of the estimator to a large class of non-Gaussian…

Statistics Theory · Mathematics 2020-09-02 Mikkel Bennedsen

Detecting the correct speech polarity is a necessary step prior to several speech processing techniques. An error on its determination could have a dramatic detrimental impact on their performance. As current systems have to deal with…

Audio and Speech Processing · Electrical Eng. & Systems 2020-06-02 Thomas Drugman

Owing to the forecasted improved sensitivity of ground-based gravitational-wave detectors, new research avenues will become accessible. This is the case for gravitational-wave strong lensing, predicted with a non-negligible observation rate…

General Relativity and Quantum Cosmology · Physics 2023-10-18 Justin Janquart , K. Haris , Otto A. Hannuksela , Chris Van Den Broeck

We introduce a "loosely coherent" method for detection of continuous gravitational waves that bridges the gap between semi-coherent and purely coherent methods. Explicit control over accepted families of signals is used to increase…

General Relativity and Quantum Cosmology · Physics 2015-05-18 Vladimir Dergachev

This paper presents an efficient numerical sensitivity-estimation method and implementation for continuous-gravitational-wave searches, extending and generalizing an earlier analytic approach by Wette [1]. This estimation framework applies…

General Relativity and Quantum Cosmology · Physics 2018-11-07 Christoph Dreissigacker , Reinhard Prix , Karl Wette

Silent speech interfaces (SSI) has been an exciting area of recent interest. In this paper, we present a non-invasive silent speech interface that uses inaudible acoustic signals to capture people's lip movements when they speak. We exploit…

Audio and Speech Processing · Electrical Eng. & Systems 2020-11-24 Jian Luo , Jianzong Wang , Ning Cheng , Guilin Jiang , Jing Xiao

This paper presents a hybrid approach to achieve iris localization based on a Laplacian of Gaussian (LoG) filter, region growing, and zero-crossings of the LoG filter. In the proposed method, an LoG filter with region growing is used to…

Computer Vision and Pattern Recognition · Computer Science 2022-01-19 Tariq M. Khan , Donald G. bailey , Yinan Kong

From a machine learning perspective, the human ability localize sounds can be modeled as a non-parametric and non-linear regression problem between binaural spectral features of sound received at the ears (input) and their sound-source…

Sound · Computer Science 2015-02-12 Yuancheng Luo , Dmitry N. Zotkin , Ramani Duraiswami

Audio tagging aims to assign predefined tags to audio clips to indicate the class information of audio events. Sequential audio tagging (SAT) means detecting both the class information of audio events, and the order in which they occur…

Sound · Computer Science 2022-10-25 Yuanbo Hou , Yun Wang , Wenwu Wang , Dick Botteldooren

This text is a compilation of some of the notes that the author has written during the development of the low-order model "DICO" [2, 8, 10, 11] for vowel phonation and the even more rudimentary glottal flow model [9] for processing…

Fluid Dynamics · Physics 2019-11-13 Jarmo Malinen

This paper introduces a novel Russian speech dataset called Golos, a large corpus suitable for speech research. The dataset mainly consists of recorded audio files manually annotated on the crowd-sourcing platform. The total duration of the…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-21 Nikolay Karpov , Alexander Denisenko , Fedor Minkin

This paper presents a methodology for early detection of audio events from audio streams. Early detection is the ability to infer an ongoing event during its initial stage. The proposed system consists of a novel inference step coupled with…

Sound · Computer Science 2019-04-09 Huy Phan , Philipp Koch , Ian McLoughlin , Alfred Mertins

We tackle the challenge of open-vocabulary segmentation, where we need to identify objects from a wide range of categories in different environments, using text prompts as our input. To overcome this challenge, existing methods often use…

Computer Vision and Pattern Recognition · Computer Science 2024-12-16 Yu-Jhe Li , Xinyang Zhang , Kun Wan , Lantao Yu , Ajinkya Kale , Xin Lu
‹ Prev 1 3 4 5 6 7 10 Next ›