中文
相关论文

相关论文: Guitar Tone Morphing by Diffusion-based Model

200 篇论文

We present a fast and high-fidelity method for music generation, based on specified f0 and loudness, such that the synthesized audio mimics the timbre and articulation of a target instrument. The generation process consists of learned…

音频与语音处理 · 电气工程与系统科学 2020-09-08 Michael Michelashvili , Lior Wolf

Device-guided music transfer adapts playback across unseen devices for users who lack them. Existing methods mainly focus on modifying the timbre, rhythm, harmony, or instrumentation to mimic genres or artists, overlooking the diverse…

声音 · 计算机科学 2025-11-24 Manh Pham Hung , Changshuo Hu , Ting Dang , Dong Ma

Moire engineering in two-dimensional transition metal dichalcogenides enables access to correlated quantum phenomena. Realizing such effects demands simultaneous control over twist angle and material composition to modulate phonons,…

This paper explores a simple extension of diffusion-based rectified flow Transformers for text-to-music generation, termed as FluxMusic. Generally, along with design in advanced Flux\footnote{https://github.com/black-forest-labs/flux}…

声音 · 计算机科学 2024-12-23 Zhengcong Fei , Mingyuan Fan , Changqian Yu , Junshi Huang

Timbre transfer techniques aim at converting the sound of a musical piece generated by one instrument into the same one as if it was played by another instrument, while maintaining as much as possible the content in terms of musical…

音频与语音处理 · 电气工程与系统科学 2023-07-31 Luca Comanducci , Fabio Antonacci , Augusto Sarti

Consumer-grade music recordings such as those captured by mobile devices typically contain distortions in the form of background noise, reverb, and microphone-induced EQ. This paper presents a deep learning approach to enhance low-quality…

声音 · 计算机科学 2022-04-29 Nikhil Kandpal , Oriol Nieto , Zeyu Jin

A method is proposed which enables one to produce musical compositions by using transposition in place of harmonic progression. A transposition scale is introduced to provide a set of intervals commensurate with the musical scale, such as…

声音 · 计算机科学 2016-01-12 Andrei V Smirnov

This paper presents a web application for visualizing the tonality of a piece of music -- the organization of its chords and scales -- at a high level of abstraction and with coordinated playback. The application applies the discrete…

声音 · 计算机科学 2022-03-25 Daniel Harasim , Giovanni Affatato , Fabian C. Moss

An algorithm called MUSIC-like algorithm was originally proposed as an alternative method to the MUltiple SIgnal Classification (MUSIC) algorithm for direction-of-arrival (DOA) estimation. Without requiring explicit model order estimation,…

信号处理 · 电气工程与系统科学 2018-11-20 Narong Borijindargoon , Boon Poh Ng

A modular method was suggested before to recover a band limited signal from the sample and hold and linearly interpolated (or, in general, an nth-order-hold) version of the regular samples. In this paper a novel approach for compensating…

多媒体 · 计算机科学 2010-11-12 Ali Ayremlou , Mohammad Tofighi , Farokh Marvasti

Recent advancements in music large language models (LLMs) have significantly improved music understanding tasks, which involve the model's ability to analyze and interpret various musical elements. These improvements primarily focused on…

声音 · 计算机科学 2025-09-24 Zhuoyuan Mao , Mengjie Zhao , Qiyu Wu , Hiromi Wakaki , Yuki Mitsufuji

Deep generative models are now able to synthesize high-quality audio signals, shifting the critical aspect in their development from audio quality to control capabilities. Although text-to-music generation is getting largely adopted by the…

声音 · 计算机科学 2024-08-02 Nils Demerlé , Philippe Esling , Guillaume Doras , David Genova

Deep learning approaches for black-box modelling of audio effects have shown promise, however, the majority of existing work focuses on nonlinear effects with behaviour on relatively short time-scales, such as guitar amplifiers and…

声音 · 计算机科学 2023-05-11 Marco Comunità , Christian J. Steinmetz , Huy Phan , Joshua D. Reiss

In this paper we present the first steps towards the creation of a tool which enables artists to create music visualizations using pre-trained, generative, machine learning models. First, we investigate the application of network bending,…

声音 · 计算机科学 2024-07-01 Luke Dzwonczyk , Carmine Emanuele Cella , David Ban

In this paper, graph theory is used to explore the musical notion of tonal modulation, in theory and application. We define (pivot) modulation graphs based on the common scales used in popular music. Properties and parameters of these…

声音 · 计算机科学 2023-06-27 Jason I. Brown , Ian George

A modular method was suggested before to recover a band limited signal from the sample and hold and linearly interpolated (or, in general, an nth-order-hold) version of the regular samples. In this paper a novel approach for compensating…

计算机视觉与模式识别 · 计算机科学 2012-05-15 Mohammad Tofighi , Ali Ayremlou , Farokh Marvasti

The objective of this paper is to understand the critical parameters that need to be addressed while designing a guitar tuner. The focus of the design lies in developing a suitable algorithm to accurately detect the fundamental frequency of…

声音 · 计算机科学 2009-12-07 Mary Lourde R. , Anjali Kuppayil Saji

Recent developments in MIR have led to several benchmark deep learning models whose embeddings can be used for a variety of downstream tasks. At the same time, the vast majority of these models have been trained on Western pop/rock music…

声音 · 计算机科学 2023-07-20 Charilaos Papaioannou , Emmanouil Benetos , Alexandros Potamianos

This thesis develops a Transformer model based on Whisper, which extracts melodies and chords from music audio and records them into ABC notation. A comprehensive data processing workflow is customized for ABC notation, including data…

声音 · 计算机科学 2024-10-23 Hongyao Zhang , Bohang Sun

Common temporal models for automatic chord recognition model chord changes on a frame-wise basis. Due to this fact, they are unable to capture musical knowledge about chord progressions. In this paper, we propose a temporal model that…

声音 · 计算机科学 2018-08-17 Filip Korzeniowski , Gerhard Widmer