English
Related papers

Related papers: Guitar Tone Morphing by Diffusion-based Model

200 papers

In this paper a novel hybrid approach for compensating the distortion of any interpolation has been proposed. In this hybrid method, a modular approach was incorporated in an iterative fashion. By using this approach we can get drastic…

Multimedia · Computer Science 2010-09-21 A. ParandehGheibi , M. A. Akhaee , A. Ayremlou , M. A. Rahimian , F. Marvasti

Many music AI models learn a map between music content and human-defined labels. However, many annotations, such as chords, can be naturally expressed within the music modality itself, e.g., as sequences of symbolic notes. This observation…

Sound · Computer Science 2025-09-30 Junyan Jiang , Daniel Chin , Liwei Lin , Xuanjie Liu , Gus Xia

We demonstrate midinfrared second-harmonic generation as a highly sensitive phonon spectroscopy technique that we exemplify using $\alpha$-quartz (SiO$_2$) as a model system. A midinfrared free-electron laser provides direct access to…

A music glove instrument equipped with force sensitive, flex and IMU sensors is trained on an electric piano to learn note sequences based on a time series of sensor inputs. Once trained, the glove is used on any surface to generate the…

Human-Computer Interaction · Computer Science 2020-01-28 Joseph Bakarji

Deep diffusion models excel at realistic image synthesis but demand large training sets-an obstacle in data-scarce domains like transesophageal echocardiography (TEE). While synthetic augmentation has boosted performance in transthoracic…

Image and Video Processing · Electrical Eng. & Systems 2025-08-19 Emmanuel Oladokun , Yuxuan Ou , Anna Novikova , Daria Kulikova , Sarina Thomas , Jurica Šprem , Vicente Grau

This paper introduces Open-Amp, a synthetic data framework for generating large-scale and diverse audio effects data. Audio effects are relevant to many musical audio processing and Music Information Retrieval (MIR) tasks, such as modelling…

Audio and Speech Processing · Electrical Eng. & Systems 2024-11-25 Alec Wright , Alistair Carson , Lauri Juvela

Deep generative models have been used in style transfer tasks for images. In this study, we adapt and improve CycleGAN model to perform music style transfer on Jazz and Classic genres. By doing so, we aim to easily generate new songs, cover…

Sound · Computer Science 2025-03-31 Fidan Samet , Oguz Bakir , Adnan Fidan

Lyric interpretations can help people understand songs and their lyrics quickly, and can also make it easier to manage, retrieve and discover songs efficiently from the growing mass of music archives. In this paper we propose BART-fusion, a…

Sound · Computer Science 2022-08-25 Yixiao Zhang , Junyan Jiang , Gus Xia , Simon Dixon

In this work, we propose a permutation invariant language model, SymphonyNet, as a solution for symbolic symphony music generation. We propose a novel Multi-track Multi-instrument Repeatable (MMR) representation for symphonic music and…

Sound · Computer Science 2022-09-19 Jiafeng Liu , Yuanliang Dong , Zehua Cheng , Xinran Zhang , Xiaobing Li , Feng Yu , Maosong Sun

There has been fascinating work on creating artistic transformations of images by Gatys. This was revolutionary in how we can in some sense alter the 'style' of an image while generally preserving its 'content'. In our work, we present a…

Sound · Computer Science 2024-12-24 Prateek Verma , Julius O. Smith

An optimized phonon approach for the numerical diagonalization of interacting electron-phonon systems is proposed. The variational method is based on an expansion in coherent states that leads to a dramatic truncation in the phonon space.…

Strongly Correlated Electrons · Physics 2009-11-10 V. Cataudella , G. De Filippis , F. Martone , C. A. Perroni

Despite deep learning's remarkable advances in style transfer across various domains, generating controllable performance-level musical style transfer for complete symbolically represented musical works remains a challenging area of…

The recent surge in the popularity of diffusion models for image synthesis has attracted new attention to their potential for generation tasks in other domains. However, their applications to symbolic music generation remain largely…

Sound · Computer Science 2025-05-07 Jincheng Zhang , György Fazekas , Charalampos Saitis

Songs, as a central form of musical art, exemplify the richness of human intelligence and creativity. While recent advances in generative modeling have enabled notable progress in long-form song generation, current systems for full-length…

Audio and Speech Processing · Electrical Eng. & Systems 2025-07-25 Huakang Chen , Yuepeng Jiang , Guobin Ma , Chunbo Hao , Shuai Wang , Jixun Yao , Ziqian Ning , Meng Meng , Jian Luan , Lei Xie

Machine-learning interatomic potentials are widely used as computationally efficient surrogates for density functional theory in atomistic simulations, enabling large-scale, long-time modeling of materials systems. We investigate how…

Materials Science · Physics 2026-04-13 Jonas Grandel , Philipp Benner , Janine George

The tunability of the interlayer coupling by twisting one layer with respect to another layer of two-dimensional materials provides a unique way to manipulate the phonons and related properties. We refer to this engineering of phononic…

Materials Science · Physics 2020-03-25 Indrajit Maity , Mit H. Naik , Prabal K Maiti , H. R. Krishnamurthy , Manish Jain

Topological phase transitions occur when the electronic bands change their topological properties, typically featuring the closing of the bandgap. While the influence of topological phase transitions on electronic and optical properties has…

Materials Science · Physics 2021-01-04 Shengying Yue , Bowen Deng , Yanming Liu , Yujie Quan , Runqing Yang , Bolin Liao

This paper explores the innovative application of the Fractional Fourier Transform (FrFT) in sound synthesis, highlighting its potential to redefine time-frequency analysis in audio processing. As an extension of the classical Fourier…

Sound · Computer Science 2025-06-12 Esteban Gutiérrez , Rodrigo Cádiz , Carlos Sing Long , Frederic Font , Xavier Serra

Image-to-Image (I2I) multi-domain translation models are usually evaluated also using the quality of their semantic interpolation results. However, state-of-the-art models frequently show abrupt changes in the image appearance during…

Computer Vision and Pattern Recognition · Computer Science 2021-06-17 Yahui Liu , Enver Sangineto , Yajing Chen , Linchao Bao , Haoxian Zhang , Nicu Sebe , Bruno Lepri , Wei Wang , Marco De Nadai

Text-to-music generation has advanced rapidly, with modern autoregressive and diffusion-based models producing convincing music from natural-language prompts. However, much of this progress relies on large-scale training data and external…

Sound · Computer Science 2026-05-21 Junyoung Koh
‹ Prev 1 8 9 10 Next ›