中文
相关论文

相关论文: Glottal Closure and Opening Instant Detection from…

200 篇论文

Gravitational wave detectors now under construction are sensitive to the phase of the incident gravitational waves. Correspondingly, the signals from the different detectors can be combined, in the analysis, to simulate a single detector of…

广义相对论与量子宇宙学 · 物理学 2009-12-31 Lee Samuel Finn

The search for continuous gravitational-wave signals requires the development of techniques that can effectively explore the low-significance regions of the candidate set. In this paper we present the methods that were developed for a…

广义相对论与量子宇宙学 · 物理学 2015-03-11 Berit Behnke , Maria Alessandra Papa , Reinhard Prix

We propose a block-online algorithm of guided source separation (GSS). GSS is a speech separation method that uses diarization information to update parameters of the generative model of observation signals. Previous studies have shown that…

音频与语音处理 · 电气工程与系统科学 2020-11-17 Shota Horiguchi , Yusuke Fujita , Kenji Nagamatsu

The goal of this work is to recognise phrases and sentences being spoken by a talking face, with or without the audio. Unlike previous works that have focussed on recognising a limited number of words or phrases, we tackle lip reading as an…

计算机视觉与模式识别 · 计算机科学 2018-12-27 Triantafyllos Afouras , Joon Son Chung , Andrew Senior , Oriol Vinyals , Andrew Zisserman

Automatic detection of voice pathology enables objective assessment and earlier intervention for the diagnosis. This study provides a systematic analysis of glottal source features and investigates their effectiveness in voice pathology…

音频与语音处理 · 电气工程与系统科学 2023-10-18 Sudarsana Reddy Kadiri , Paavo Alku

Although speech and gesture recognition has been studied extensively, all the successful attempts of combining them in the unified framework were semantically motivated, e.g., keyword-gesture cooccurrence. Such formulations inherited the…

计算机视觉与模式识别 · 计算机科学 2007-05-23 Sanshzar Kettebekov , Mohammed Yeasin , Rajeev Sharma

This thesis presents advancements in the detection of gravitational waves from compact binary coalescences, utilising the most sensitive observatories constructed to date. The research focuses on enhancing gravitational-wave signal searches…

广义相对论与量子宇宙学 · 物理学 2026-01-27 Arthur Tolley

This paper introduces GlOttal-flow LPC Filter (GOLF), a novel method for singing voice synthesis (SVS) that exploits the physical characteristics of the human voice using differentiable digital signal processing. GOLF employs a glottal…

音频与语音处理 · 电气工程与系统科学 2024-10-21 Chin-Yun Yu , György Fazekas

Recent studies have shown that text-to-speech synthesis quality can be improved by using glottal vocoding. This refers to vocoders that parameterize speech into two parts, the glottal excitation and vocal tract, that occur in the human…

音频与语音处理 · 电气工程与系统科学 2019-03-15 Bajibabu Bollepalli , Lauri Juvela , Paavo Alku

The goal of this project is to develop a limited lip reading algorithm for a subset of the English language. We consider a scenario in which no audio information is available. The raw video is processed and the position of the lips in each…

计算机视觉与模式识别 · 计算机科学 2017-08-04 Jithin Donny George , Ronan Keane , Conor Zellmer

Complex cepstrum is known in the literature for linearly separating causal and anticausal components. Relying on advances achieved by the Zeros of the Z-Transform (ZZT) technique, we here investigate the possibility of using complex…

声音 · 计算机科学 2020-01-01 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

An inversion of the speech polarity may have a dramatic detrimental effect on the performance of various techniques of speech processing. An automatic method for determining the speech polarity (which is dependent upon the recording setup)…

声音 · 计算机科学 2020-05-19 Thomas Drugman , Thierry Dutoit

We present a new ${\it{gating}}$ method to remove non-Gaussian noise transients in gravitational wave data. The method does not rely on any a-priori knowledge on the amplitude or duration of the transient events. In light of the character…

广义相对论与量子宇宙学 · 物理学 2022-02-02 Benjamin Steltner , Maria Alessandra Papa , Heinz-Bernd Eggenstein

Starting with a collection of traces generated by process executions, process discovery is the task of constructing a simple model that describes the process, where simplicity is often measured in terms of model size. The challenge of…

人工智能 · 计算机科学 2024-04-17 Hanan Alkhammash , Artem Polyvyanyy , Alistair Moffat

Open-vocabulary human-object interaction (HOI) detection, which is concerned with the problem of detecting novel HOIs guided by natural language, is crucial for understanding human-centric scenes. However, prior zero-shot HOI detectors…

计算机视觉与模式识别 · 计算机科学 2024-04-11 Ting Lei , Shaofeng Yin , Yang Liu

A sound source was proposed for acoustic measurements of physical models of the human vocal tract. The physical models are produced by Fast Prototyping, based on Magnetic Resonance Imaging during prolonged vowel production. The sound…

仪器与探测器 · 物理学 2017-11-22 Antti Hannukainen , Juha Kuortti , Jarmo Malinen , Antti Ojalammi

We present a new method for the classification of transient noise signals (or glitches) in advanced gravitational-wave interferometers. The method uses learned dictionaries (a supervised machine learning algorithm) for signal denoising, and…

天体物理仪器与方法 · 物理学 2019-05-22 Miquel Llorens-Monteagudo , Alejandro Torres-Forné , José A. Font , Antonio Marquina

Speech intelligibility assessment is essential for many speech-related applications. However, most objective intelligibility metrics are intrusive, as they require clean reference speech in addition to the degraded or processed signal for…

声音 · 计算机科学 2025-12-23 Wenyu Luo , Jinhui Chen

We propose an objective intelligibility measure (OIM), called the Gammachirp Envelope Similarity Index (GESI), which can predict the speech intelligibility (SI) of simulated hearing loss (HL) sounds for normal hearing (NH) listeners. GESI…

音频与语音处理 · 电气工程与系统科学 2024-03-15 Ayako Yamamoto , Toshio Irino , Fuki Miyazaki , Honoka Tamaru

We introduce a deep learning model for speech denoising, a long-standing challenge in audio analysis arising in numerous applications. Our approach is based on a key observation about human speech: there is often a short pause between each…

声音 · 计算机科学 2020-10-26 Ruilin Xu , Rundi Wu , Yuko Ishiwaka , Carl Vondrick , Changxi Zheng