English

Score and Lyrics-Free Singing Voice Generation

Sound 2020-07-22 v2 Machine Learning Audio and Speech Processing

Abstract

Generative models for singing voice have been mostly concerned with the task of ``singing voice synthesis,'' i.e., to produce singing voice waveforms given musical scores and text lyrics. In this work, we explore a novel yet challenging alternative: singing voice generation without pre-assigned scores and lyrics, in both training and inference time. In particular, we outline three such generation schemes, and propose a pipeline to tackle these new tasks. Moreover, we implement such models using generative adversarial networks and evaluate them both objectively and subjectively.

Keywords

Cite

@article{arxiv.1912.11747,
  title  = {Score and Lyrics-Free Singing Voice Generation},
  author = {Jen-Yu Liu and Yu-Hua Chen and Yin-Cheng Yeh and Yi-Hsuan Yang},
  journal= {arXiv preprint arXiv:1912.11747},
  year   = {2020}
}

Comments

Accepted by International Conference on Computational Creativity (ICCC) 2020

R2 v1 2026-06-23T12:56:33.808Z