中文
相关论文

相关论文: pch2csd: an application for converting Nord Modula…

200 篇论文

Modulations are a critical part of sound design and music production, enabling the creation of complex and evolving audio. Modern synthesizers provide envelopes, low frequency oscillators (LFOs), and more parameter automation tools that…

声音 · 计算机科学 2025-10-08 Christopher Mitcheltree , Hao Hao Tan , Joshua D. Reiss

Submodular functions have many applications. Matchings have many applications. The bitext word alignment problem can be modeled as the problem of maximizing a nonnegative, monotone, submodular function constrained to matchings in a complete…

数据结构与算法 · 计算机科学 2013-01-14 Sagar Kale

Neural audio codecs are initially introduced to compress audio data into compact codes to reduce transmission latency. Researchers recently discovered the potential of codecs as suitable tokenizers for converting continuous audio into…

音频与语音处理 · 电气工程与系统科学 2024-02-21 Haibin Wu , Xuanjun Chen , Yi-Cheng Lin , Kai-wei Chang , Ho-Lam Chung , Alexander H. Liu , Hung-yi Lee

Component-based development (CBD) is a name, with which software development professionals are quite familiar. There are several models which have been proposed for CBD in last few years. They contain good features but there are some…

软件工程 · 计算机科学 2012-02-14 M. Rizwan Jameel Qureshi , M. E. Sandhu

Encouraged by recent interest in traditional Chinese instruments this work proposes a computational sound synthesis model for a traditional Chinese instrument, the guzheng. Digital waveguide model and modal synthesis are the most popular…

信号处理 · 电气工程与系统科学 2019-10-15 Enda Zhang , Gopal Gupta , Charles Greif , Andrew Paplinski

This paper advances the design of CTC-based all-neural (or end-to-end) speech recognizers. We propose a novel symbol inventory, and a novel iterated-CTC method in which a second system is used to transform a noisy initial output into a…

计算与语言 · 计算机科学 2022-02-24 G. Zweig , C. Yu , J. Droppo , A. Stolcke

Code-switching is a data augmentation scheme mixing words from multiple languages into source lingual text. It has achieved considerable generalization performance of cross-lingual transfer tasks by aligning cross-lingual contextual word…

计算与语言 · 计算机科学 2024-06-21 Zhuoran Li , Chunming Hu , Junfan Chen , Zhijun Chen , Xiaohui Guo , Richong Zhang

Synthesizer is a type of electronic musical instrument that is now widely used in modern music production and sound design. Each parameters configuration of a synthesizer produces a unique timbre and can be viewed as a unique instrument.…

声音 · 计算机科学 2022-07-29 Zui Chen , Yansen Jing , Shengcheng Yuan , Yifei Xu , Jian Wu , Hang Zhao

Neural codecs have demonstrated strong performance in high-fidelity compression of audio signals at low bitrates. The token-based representations produced by these codecs have proven particularly useful for generative modeling. While much…

音频与语音处理 · 电气工程与系统科学 2025-04-16 Patrick O'Reilly , Prem Seetharaman , Jiaqi Su , Zeyu Jin , Bryan Pardo

Modifying the pitch and timing of an audio signal are fundamental audio editing operations with applications in speech manipulation, audio-visual synchronization, and singing voice editing and synthesis. Thus far, methods for pitch-shifting…

音频与语音处理 · 电气工程与系统科学 2021-10-07 Max Morrison , Zeyu Jin , Nicholas J. Bryan , Juan-Pablo Caceres , Bryan Pardo

Utterances by L2 speakers can be unintelligible due to mispronunciation and improper prosody. In computer-aided language learning systems, textual feedback is often provided using a speech recognition engine. However, an ideal form of…

声音 · 计算机科学 2024-10-04 Haopeng Geng , Daisuke Saito , Nobuaki Minematsu

In view of the High Luminosity LHC, the current CMS tracking detector will have to be replaced during Long Shutdown 3 to cope with the higher radiation environment and to withstand an increased data rate. To prepare for the so-called CMS…

仪器与探测器 · 物理学 2025-05-13 Giorgia Bonomelli

Recently, sequence-to-sequence (seq-to-seq) models have been successfully applied in text-to-speech (TTS) to synthesize speech for single-language text. To synthesize speech for multiple languages usually requires multi-lingual speech from…

声音 · 计算机科学 2022-11-18 Haitong Zhang , Yue Lin

This paper describes an experimental system designed for development of real time voice synthesis applications. The system is composed from a DSP coprocessor card, equipped with an TMS320C25 or TMS320C50 chip, voice acquisition module…

声音 · 计算机科学 2008-03-04 Radu Arsinte , Attila Ferencz , Costin Miron

We construct modular invariant partition functions for strings propagating on non-compact manifolds of G_2 holonomy. Our amplitudes involve a pair of N=1 minimal models M_m, M_{m+2} (m=3,4,...) and are identified as describing strings on…

高能物理 - 理论 · 物理学 2009-11-07 Tohru Eguchi , Yuji Sugawara

This paper introduces a nonlinear string sound synthesizer, based on a finite difference simulation of the dynamic behavior of strings under various excitations. The presented synthesizer features a versatile string simulation engine…

声音 · 计算机科学 2024-01-09 Jin Woo Lee , Min Jun Choi , Kyogu Lee

P300 is an Event-Related Potential widely used in Brain-Computer Interfaces, but its detection is challenging due to inter-subject and temporal variability. This work introduces a clustering methodology based on Normalized Compression…

机器学习 · 计算机科学 2025-02-04 Guillermo Sarasa , Ana Granados , Francisco B Rodríguez

The premise of this article is that a basic understanding of the composition and functioning of large language models is critically urgent. To that end, we extract a representational map of OpenAI's GPT-2 with what we articulate as two…

计算与语言 · 计算机科学 2023-04-20 Minh Hua , Rita Raley

Audio synthesis has broad applications in multimedia. Recent advancements have made it possible to generate relevant audios from inputs describing an audio scene, such as images or texts. However, the immersiveness and expressiveness of the…

多媒体 · 计算机科学 2025-08-13 Wei Guo , Heng Wang , Jianbo Ma , Weidong Cai

Deep learning models require large amounts of clean data to acheive good performance. To avoid the cost of expensive data acquisition, researchers use the abundant data available on the internet. This raises significant privacy concerns on…

声音 · 计算机科学 2024-01-05 Vignesh Gokul , Shlomo Dubnov