中文
相关论文

相关论文: Glottal Closure and Opening Instant Detection from…

200 篇论文

We examine an analytic variational inference scheme for the Gaussian Process State Space Model (GPSSM) - a probabilistic model for system identification and time-series modelling. Our approach performs variational inference over both the…

机器学习 · 统计学 2018-12-11 Alessandro Davide Ialongo , Mark van der Wilk , Carl Edward Rasmussen

With the development of internet of things technologies, tremendous sensor audio data has been produced, which poses great challenges to audio-based event detection in smart cities. In this paper, we target a challenging audio-based event…

声音 · 计算机科学 2023-12-27 Haoyu Tang , Yunxiao Wang , Jihua Zhu , Shuaike Zhang , Mingzhu Xu , Qinghai Zheng , Yupeng Hu

This paper presents a novel method for fault detection in vibration/acoustic signals contaminated with non-Gaussian noise, specifically addressing the challenge of random impulsive and wideband disturbances in industrial measurements. While…

信号处理 · 电气工程与系统科学 2025-02-18 A Drewnicka , A Michalak , R Zimroz , A Kumar , A Wyłomańska , J Wodecki

We present an algorithm for the identification of transient noise artifacts (glitches) in cross-correlation searches for long O(10s) gravitational-wave transients. The algorithm utilizes the auto-power in each detector as a discriminator…

Recently, increasing attention has been directed to the study of the speech emotion recognition, in which global acoustic features of an utterance are mostly used to eliminate the content differences. However, the expression of speech…

人机交互 · 计算机科学 2018-11-21 Haotian Guan , Zhilei Liu , Longbiao Wang , Jianwu Dang , Ruiguo Yu

Time- and pitch-scale modifications of speech signals find important applications in speech synthesis, playback systems, voice conversion, learning/hearing aids, etc.. There is a requirement for computationally efficient and real-time…

音频与语音处理 · 电气工程与系统科学 2018-01-22 Sunil Rudresh , Aditya Vasisht , Karthika Vijayan , Chandra Sekhar Seelamantula

This paper presents the Multimodal Laryngoscopic Video Analyzing System (MLVAS), a novel system that leverages both audio and video data to automatically extract key video segments and metrics from raw laryngeal videostroboscopic videos for…

声音 · 计算机科学 2026-03-10 Yucong Zhang , Xin Zou , Jinshan Yang , Wenjun Chen , Juan Liu , Faya Liang , Ming Li

We study a well-known estimator of the fractal index of a stochastic process. Our framework is very general and encompasses many models of interest; we show how to extend the theory of the estimator to a large class of non-Gaussian…

统计理论 · 数学 2020-09-02 Mikkel Bennedsen

Detecting the correct speech polarity is a necessary step prior to several speech processing techniques. An error on its determination could have a dramatic detrimental impact on their performance. As current systems have to deal with…

音频与语音处理 · 电气工程与系统科学 2020-06-02 Thomas Drugman

Owing to the forecasted improved sensitivity of ground-based gravitational-wave detectors, new research avenues will become accessible. This is the case for gravitational-wave strong lensing, predicted with a non-negligible observation rate…

广义相对论与量子宇宙学 · 物理学 2023-10-18 Justin Janquart , K. Haris , Otto A. Hannuksela , Chris Van Den Broeck

We introduce a "loosely coherent" method for detection of continuous gravitational waves that bridges the gap between semi-coherent and purely coherent methods. Explicit control over accepted families of signals is used to increase…

广义相对论与量子宇宙学 · 物理学 2015-05-18 Vladimir Dergachev

This paper presents an efficient numerical sensitivity-estimation method and implementation for continuous-gravitational-wave searches, extending and generalizing an earlier analytic approach by Wette [1]. This estimation framework applies…

广义相对论与量子宇宙学 · 物理学 2018-11-07 Christoph Dreissigacker , Reinhard Prix , Karl Wette

Silent speech interfaces (SSI) has been an exciting area of recent interest. In this paper, we present a non-invasive silent speech interface that uses inaudible acoustic signals to capture people's lip movements when they speak. We exploit…

音频与语音处理 · 电气工程与系统科学 2020-11-24 Jian Luo , Jianzong Wang , Ning Cheng , Guilin Jiang , Jing Xiao

This paper presents a hybrid approach to achieve iris localization based on a Laplacian of Gaussian (LoG) filter, region growing, and zero-crossings of the LoG filter. In the proposed method, an LoG filter with region growing is used to…

计算机视觉与模式识别 · 计算机科学 2022-01-19 Tariq M. Khan , Donald G. bailey , Yinan Kong

From a machine learning perspective, the human ability localize sounds can be modeled as a non-parametric and non-linear regression problem between binaural spectral features of sound received at the ears (input) and their sound-source…

声音 · 计算机科学 2015-02-12 Yuancheng Luo , Dmitry N. Zotkin , Ramani Duraiswami

Audio tagging aims to assign predefined tags to audio clips to indicate the class information of audio events. Sequential audio tagging (SAT) means detecting both the class information of audio events, and the order in which they occur…

声音 · 计算机科学 2022-10-25 Yuanbo Hou , Yun Wang , Wenwu Wang , Dick Botteldooren

This text is a compilation of some of the notes that the author has written during the development of the low-order model "DICO" [2, 8, 10, 11] for vowel phonation and the even more rudimentary glottal flow model [9] for processing…

流体动力学 · 物理学 2019-11-13 Jarmo Malinen

This paper introduces a novel Russian speech dataset called Golos, a large corpus suitable for speech research. The dataset mainly consists of recorded audio files manually annotated on the crowd-sourcing platform. The total duration of the…

音频与语音处理 · 电气工程与系统科学 2021-06-21 Nikolay Karpov , Alexander Denisenko , Fedor Minkin

This paper presents a methodology for early detection of audio events from audio streams. Early detection is the ability to infer an ongoing event during its initial stage. The proposed system consists of a novel inference step coupled with…

声音 · 计算机科学 2019-04-09 Huy Phan , Philipp Koch , Ian McLoughlin , Alfred Mertins

We tackle the challenge of open-vocabulary segmentation, where we need to identify objects from a wide range of categories in different environments, using text prompts as our input. To overcome this challenge, existing methods often use…

计算机视觉与模式识别 · 计算机科学 2024-12-16 Yu-Jhe Li , Xinyang Zhang , Kun Wan , Lantao Yu , Ajinkya Kale , Xin Lu