中文
相关论文

相关论文: Creating an A Cappella Singing Audio Dataset for A…

200 篇论文

Computational harmony analysis is important for MIR tasks such as automatic segmentation, corpus analysis and automatic chord label estimation. However, recent research into the ambiguous nature of musical harmony, causing limited…

声音 · 计算机科学 2023-10-18 Hendrik Vincent Koops , Gianluca Micchi , Ilaria Manco , Elio Quinton

Musical audio is generally composed of three physical properties: frequency, time and magnitude. Interestingly, human auditory periphery also provides neural codes for each of these dimensions to perceive music. Inspired by these intrinsic…

声音 · 计算机科学 2021-06-16 Shuai Yu , Xiaoheng Sun , Yi Yu , Wei Li

We propose a data cleansing method that utilizes a neural analysis and synthesis (NANSY++) framework to train an end-to-end neural diarization model (EEND) for singer diarization. Our proposed model converts song data with choral singing…

音频与语音处理 · 电气工程与系统科学 2024-06-25 Hokuto Munakata , Ryo Terashima , Yusuke Fujita

Fine-grained entity typing is a challenging task with wide applications. However, most existing datasets for this task are in English. In this paper, we introduce a corpus for Chinese fine-grained entity typing that contains 4,800 mentions…

计算与语言 · 计算机科学 2020-04-21 Chin Lee , Hongliang Dai , Yangqiu Song , Xin Li

Singing voice beat and downbeat tracking posses several applications in automatic music production, analysis and manipulation. Among them, some require real-time processing, such as live performance processing and auto-accompaniment for…

音频与语音处理 · 电气工程与系统科学 2023-06-06 Mojtaba Heydari , Ju-Chiang Wang , Zhiyao Duan

Recognizing human non-speech vocalizations is an important task and has broad applications such as automatic sound transcription and health condition monitoring. However, existing datasets have a relatively small number of vocal sound…

声音 · 计算机科学 2022-06-22 Yuan Gong , Jin Yu , James Glass

Creating a pop song melody according to pre-written lyrics is a typical practice for composers. A computational model of how lyrics are set as melodies is important for automatic composition systems, but an end-to-end lyric-to-melody model…

音频与语音处理 · 电气工程与系统科学 2023-01-05 Daiyu Zhang , Ju-Chiang Wang , Katerina Kosta , Jordan B. L. Smith , Shicen Zhou

We are investigating the broader concept of using AI-based generative music systems to generate training data for Music Information Retrieval (MIR) tasks. To kick off this line of work, we ran an initial experiment in which we trained a…

声音 · 计算机科学 2023-11-16 Nadine Kroher , Helena Cuesta , Aggelos Pikrakis

Obtaining large-scale human-labeled datasets to train acoustic representation models is a very challenging task. On the contrary, we can easily collect data with machine-generated labels. In this work, we propose to exploit…

计算机视觉与模式识别 · 计算机科学 2020-01-03 Shaoyong Jia , Xin Shu , Yang Yang , Dawei Liang , Qiyue Liu , Junhui Liu

In this paper we describe an approach to identify the name of a piece of piano music, based on a short audio excerpt of a performance. Given only a description of the pieces in text format (i.e. no score information is provided), a…

信息检索 · 计算机科学 2017-08-03 Andreas Arzt , Gerhard Widmer

We propose a novel procedure to generate pseudo mandarin speech data named as CAMP (character audio mix up), which aims at generating audio from a character scale. We also raise a method for building a mandarin character scale audio…

声音 · 计算机科学 2022-10-25 Zeping Min , Qian Ge , Zhong Li

Assessing spoken language is challenging, and quantifying pronunciation metrics for machine learning models is even harder. However, for the Holy Quran, this task is simplified by the rigorous recitation rules (tajweed) established by…

音频与语音处理 · 电气工程与系统科学 2025-09-03 Abdullah Abdelfattah , Mahmoud I. Khalil , Hazem Abbas

Intonation is one of the important factors affecting the teaching language arts, so it is an urgent problem to be addressed by evaluating the teachers' intonation through artificial intelligence technology. However, the lack of an…

声音 · 计算机科学 2023-12-15 Shuhua Liu , Chunyu Zhang , Binshuai Li , Niantong Qin , Huanting Cheng , Huayu Zhang

We introduce Jamendo-QA, a large-scale dataset for Music Question Answering (Music-QA). The dataset is built on freely licensed tracks from the Jamendo platform and is automatically annotated using the Qwen-Omni model. Jamendo-QA provides…

多媒体 · 计算机科学 2025-09-22 Junyoung Koh , Soo Yong Kim , Yongwon Choi , Gyu Hyeong Choi

Recent advances in deep learning accelerated the development of content-based automatic music tagging systems. Music information retrieval (MIR) researchers proposed various architecture designs, mainly based on convolutional neural…

音频与语音处理 · 电气工程与系统科学 2020-06-02 Minz Won , Andres Ferraro , Dmitry Bogdanov , Xavier Serra

Procedural audio, often referred to as "digital Foley", generates sound from scratch using computational processes. It represents an innovative approach to sound-effects creation. However, the development and adoption of procedural audio…

声音 · 计算机科学 2025-01-30 Nelly Garcia , Joshua Reiss

Computational historical linguistics seeks to systematically understand processes of sound change, including during periods at which little to no formal recording of language is attested. At the same time, few computational resources exist…

计算与语言 · 计算机科学 2024-04-26 Stephen Bothwell , Brian DuSell , David Chiang , Brian Krostenko

This paper addresses the challenge of learning to recite the Quran for non-Arabic speakers. We explore the possibility of crowdsourcing a carefully annotated Quranic dataset, on top of which AI models can be built to simplify the learning…

声音 · 计算机科学 2024-05-07 Raghad Salameh , Mohamad Al Mdfaa , Nursultan Askarbekuly , Manuel Mazzara

Word similarity computation is a widely recognized task in the field of lexical semantics. Most proposed tasks test on similarity of word pairs of single morpheme, while few works focus on words of two morphemes or more morphemes. In this…

计算与语言 · 计算机科学 2019-07-02 Junjie Huang , Fanchao Qi , Chenghao Yang , Zhiyuan Liu , Maosong Sun

In this work we propose a new task: artistic visualization of classical Chinese poems, where the goal is to generatepaintings of a certain artistic style for classical Chinese poems. For this purpose, we construct a new dataset called…

计算机视觉与模式识别 · 计算机科学 2021-09-29 Dan Li , Shuai Wang , Jie Zou , Chang Tian , Elisha Nieuwburg , Fengyuan Sun , Evangelos Kanoulas
‹ 上一页 1 8 9 10 下一页 ›