中文
相关论文

相关论文: Glottal Source Estimation using an Automatic Chirp…

200 篇论文

We introduce ZeroSCROLLS, a zero-shot benchmark for natural language understanding over long texts, which contains only test and small validation sets, without training data. We adapt six tasks from the SCROLLS benchmark, and add four new…

计算与语言 · 计算机科学 2023-12-19 Uri Shaham , Maor Ivgi , Avia Efrat , Jonathan Berant , Omer Levy

This paper proposes an Incremental Disentanglement-based Environment-Aware zero-shot text-to-speech (TTS) method, dubbed IDEA-TTS, that can synthesize speech for unseen speakers while preserving the acoustic characteristics of a given…

音频与语音处理 · 电气工程与系统科学 2024-12-24 Ye-Xin Lu , Hui-Peng Du , Zheng-Yan Sheng , Yang Ai , Zhen-Hua Ling

We propose a new method to measure parameters of SMBHBs as individual resolvable evolving single GWs sources, using the timing data of three or more non-simultaneous pulsars. These parameters include the sky position of the SMBHB and a…

高能天体物理现象 · 物理学 2013-04-15 Shuxu Yi

This paper is concerned with reconstructing an acoustic obstacle and its excitation sources from the phaseless near-field measurements. By supplementing some artificial sources to the inverse scattering system, this co-inversion problem can…

数值分析 · 数学 2022-12-20 Deyue Zhang , Yue Wu , Yukun Guo

This paper presents an Expert Decision Support System for the identification of time-invariant, aeroacoustic source types. The system comprises two steps: first, acoustic properties are calculated based on spectral and spatial information.…

声音 · 计算机科学 2022-03-09 Armin Goudarzi , Carsten Spehr , Steffen Herbold

Phase processing has been replaced by group delay processing for the extraction of source and system parameters from speech. Group delay functions are ill-behaved when the transfer function has zeros that are close to unit circle in the…

声音 · 计算机科学 2016-03-18 Rajeev Rajan , Hema A. Murthy

As computational tools for X-ray computed tomography (CT) become more quantitatively accurate, knowledge of the source-detector spectral response is critical for quantitative system-independent reconstruction and material characterization…

图像与视频处理 · 电气工程与系统科学 2023-02-28 Wenrui Li , Venkatesh Sridhar , K. Aditya Mohan , Saransh Singh , Jean-Baptiste Forien , Xin Liu , Gregery T. Buzzard , Charles A. Bouman

We proposed a novel unsupervised methodology named Disarranged Zone Learning (DZL) to automatically recognize stenosis in coronary angiography. The methodology firstly disarranges the frames in a video, secondly it generates an effective…

图像与视频处理 · 电气工程与系统科学 2021-10-05 Yanan Dai , Pengxiong Zhu , Bangde Xue , Yun Ling , Xibao Shi , Liang Geng , Qi Zhang , Jun Liu

While most research into speech synthesis has focused on synthesizing high-quality speech for in-dataset speakers, an equally essential yet unsolved problem is synthesizing speech for unseen speakers who are out-of-dataset with limited…

声音 · 计算机科学 2023-08-28 Wenbin Wang , Yang Song , Sanjay Jha

We develop a new approach for estimating the expected values of nonlinear functions applied to multivariate random variables with arbitrary distributions. Rather than assuming a particular distribution, we assume that we are only given the…

数值分析 · 数学 2020-06-25 Deanna Easley , Tyrus Berry

Evaluating disfluency removal in speech requires more than aggregate token-level scores. Traditional word-based metrics such as precision, recall, and F1 (E-Scores) capture overall performance but cannot reveal why models succeed or fail.…

We present EdiTTS, an off-the-shelf speech editing methodology based on score-based generative modeling for text-to-speech synthesis. EdiTTS allows for targeted, granular editing of audio, both in terms of content and pitch, without the…

声音 · 计算机科学 2022-07-12 Jaesung Tae , Hyeongju Kim , Taesu Kim

We report the first analytical expression purely constructed by a machine to determine photometric redshifts ($z_{\rm phot}$) of galaxies. A simple and reliable functional form is derived using $41,214$ galaxies from the Sloan Digital Sky…

宇宙学与河外天体物理 · 物理学 2015-06-16 A. Krone-Martins , E. E. O. Ishida , R. S. de Souza

With the increased applications of automatic speech recognition (ASR) in recent years, it is essential to automatically insert punctuation marks and remove disfluencies in transcripts, to improve the readability of the transcripts as well…

计算与语言 · 计算机科学 2020-03-04 Qian Chen , Mengzhe Chen , Bo Li , Wen Wang

Zero pronouns (ZPs) are frequently omitted in pro-drop languages, but should be recalled in non-pro-drop languages. This discourse phenomenon poses a significant challenge for machine translation (MT) when translating texts from pro-drop to…

计算与语言 · 计算机科学 2019-09-04 Longyue Wang , Zhaopeng Tu , Xing Wang , Shuming Shi

Pre-trained vision-language models (e.g., CLIP) have shown promising zero-shot generalization in many downstream tasks with properly designed text prompts. Instead of relying on hand-engineered prompts, recent works learn prompts using the…

计算机视觉与模式识别 · 计算机科学 2022-09-16 Manli Shu , Weili Nie , De-An Huang , Zhiding Yu , Tom Goldstein , Anima Anandkumar , Chaowei Xiao

Analysis of signals with oscillatory modes with crossover instantaneous frequencies is a challenging problem in time series analysis. One way to handle this problem is lifting the 2-dimensional time-frequency representation to a…

数值分析 · 数学 2022-06-22 Ziyu Chen , Hau-Tieng Wu

To determine the magnification of an extended source caused by gravitational lensing one has to perform a two-dimensional integral over point-source magnifications in general. Since the point-source magnification jumps to an infinite value…

天体物理学 · 物理学 2011-05-23 M. Dominik

We present a continuous-time probabilistic approach for estimating the chirp signal and its instantaneous frequency function when the true forms of these functions are not accessible. Our model represents these functions by non-linearly…

机器学习 · 统计学 2023-03-22 Zheng Zhao , Simo Särkkä , Jens Sjölund , Thomas B. Schön

Voiced segments of speech are assumed to be composed of non-stationary acoustic objects which can be described as stationary response of a non-stationary fundamental drive (FD) process and which are furthermore suited to reconstruct the…

声音 · 计算机科学 2007-05-23 Friedhelm R. Drepper