English
Related papers

Related papers: An Efficient Modal-based Approach Towards Guzheng …

200 papers

Most contemporary music tagging systems rely on large volumes of annotated data. As an alternative, we investigate the extent to which synthetically generated music excerpts can improve tagging systems when only small annotated collections…

Sound · Computer Science 2024-07-03 Nadine Kroher , Steven Manangu , Aggelos Pikrakis

Recently, denoising diffusion models have demonstrated remarkable performance among generative models in various domains. However, in the speech domain, the application of diffusion models for synthesizing time-varying audio faces…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-13 Ji-Sang Hwang , Sang-Hoon Lee , Seong-Whan Lee

Classical Chinese poetry is a vital and enduring part of Chinese literature, conveying profound emotional resonance. Existing studies analyze sentiment based on textual meanings, overlooking the unique rhythmic and visual features inherent…

Computation and Language · Computer Science 2025-05-20 Xiaocong Du , Haoyu Pei , Haipeng Zhang

We introduce two rule-based models to modify the prosody of speech synthesis in order to modulate the emotion to be expressed. The prosody modulation is based on speech synthesis markup language (SSML) and can be used with any commercial…

Sound · Computer Science 2023-07-06 Felix Burkhardt , Uwe Reichel , Florian Eyben , Björn Schuller

This paper presents XiaoiceSing, a high-quality singing voice synthesis system which employs an integrated network for spectrum, F0 and duration modeling. We follow the main architecture of FastSpeech while proposing some singing-specific…

Audio and Speech Processing · Electrical Eng. & Systems 2020-06-12 Peiling Lu , Jie Wu , Jian Luan , Xu Tan , Li Zhou

Modulations are a critical part of sound design and music production, enabling the creation of complex and evolving audio. Modern synthesizers provide envelopes, low frequency oscillators (LFOs), and more parameter automation tools that…

Sound · Computer Science 2025-10-08 Christopher Mitcheltree , Hao Hao Tan , Joshua D. Reiss

Whether literally or suggestively, the concept of soundscape is alluded in both modern and ancient music. In this study, we examine whether we can analyze and compare Western and Chinese classical music based on soundscape models. We…

Sound · Computer Science 2020-02-24 Jianyu Fan , Yi-Hsuan Yang , Kui Dong , Philippe Pasquier

Classical parametric speech coding techniques provide a compact representation for speech signals. This affords a very low transmission rate but with a reduced perceptual quality of the reconstructed signals. Recently, autoregressive deep…

Audio and Speech Processing · Electrical Eng. & Systems 2019-07-02 Ahmed Mustafa , Arijit Biswas , Christian Bergler , Julia Schottenhamml , Andreas Maier

Expressive speech synthesis aims to generate speech that captures a wide range of para-linguistic features, including emotion and articulation, though current research primarily emphasizes emotional aspects over the nuanced articulatory…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-24 Zehua Kcriss Li , Meiying Melissa Chen , Yi Zhong , Pinxin Liu , Zhiyao Duan

Poetry generation is an interesting research topic in the field of text generation. As one of the most valuable literary and cultural heritages of China, Chinese classical poetry is very familiar and loved by Chinese people from generation…

Computation and Language · Computer Science 2020-03-26 Jinyi Hu , Maosong Sun

A novel multi-channel artificial wind noise generator based on a fluid dynamics model, namely the Corcos model, is proposed. In particular, the model is used to approximate the complex coherence function of wind noise signals measured with…

Audio and Speech Processing · Electrical Eng. & Systems 2019-12-13 Daniele Mirabilii , Emanuël A. P. Habets

Physical models of rigid bodies are used for sound synthesis in applications from virtual environments to music production. Traditional methods such as modal synthesis often rely on computationally expensive numerical solvers, while recent…

Sound · Computer Science 2022-10-31 Rodrigo Diaz , Ben Hayes , Charalampos Saitis , György Fazekas , Mark Sandler

Exterior sound field interpolation is a challenging problem that often requires specific array configurations and prior knowledge on the source conditions. We propose an interpolation method based on Gaussian processes using a point source…

Audio and Speech Processing · Electrical Eng. & Systems 2026-02-06 Juliano G. C. Ribeiro , Ryo Matsuda , Jorge Trevino

We introduce Gull, a generative multifunctional audio codec. Gull is a general purpose neural audio compression and decompression model which can be applied to a wide range of tasks and applications such as real-time communication, audio…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-10 Yi Luo , Jianwei Yu , Hangting Chen , Rongzhi Gu , Chao Weng

Acoustic feedback is a critical indicator for assessing the contact condition between the tool and the workpiece when humans perform grinding tasks with rotary tools. In contrast, robotic grinding systems typically rely on force sensing,…

Robotics · Computer Science 2026-04-07 Zongyuan Zhang , Christopher Lehnert , Will N. Browne , Jonathan M. Roberts

Noise can induce coherent oscillations in excitable systems without periodic orbits. Here, we establish a method to derive a hybrid system approximating the noise-induced coherent oscillations in excitable systems and further perform phase…

Adaptation and Self-Organizing Systems · Physics 2022-05-26 Jinjie Zhu , Yuzuru Kato , Hiroya Nakao

Automatic synthesis of realistic co-speech gestures is an increasingly important yet challenging task in artificial embodied agent creation. Previous systems mainly focus on generating gestures in an end-to-end manner, which leads to…

Sound · Computer Science 2023-05-05 Tenglong Ao , Qingzhe Gao , Yuke Lou , Baoquan Chen , Libin Liu

Generating sound effects with controllable variations is a challenging task, traditionally addressed using sophisticated physical models that require in-depth knowledge of signal processing parameters and algorithms. In the era of…

Sound · Computer Science 2024-12-30 Yunyi Liu , Craig Jin

In this paper, one of the major shortcomings of the conventional numerical approaches is alleviated by introducing the probabilistic nature of molecular transitions into the framework of classical computational electrodynamics. The main aim…

Classical Physics · Physics 2020-01-24 Ali Reza Hashemi , Mahmood Hosseini-Farzad

Speech-driven gesture generation is highly challenging due to the random jitters of human motion. In addition, there is an inherent asynchronous relationship between human speech and gestures. To tackle these challenges, we introduce a…

Human-Computer Interaction · Computer Science 2023-05-19 Sicheng Yang , Zhiyong Wu , Minglei Li , Zhensong Zhang , Lei Hao , Weihong Bao , Haolin Zhuang