中文
相关论文

相关论文: Musical Score Following using Statistical Inferenc…

200 篇论文

We present a probabilistic generative model for timing deviations in expressive music performance. The structure of the proposed model is equivalent to a switching state space model. The switch variables correspond to discrete note…

人工智能 · 计算机科学 2011-06-27 A. T. Cemgil , B. Kappen

This application-oriented study concerns computational musicology, which makes use of grammar systems. We define multi-generative rule-synchronized scattered-context grammar systems (without erasing rules) and demonstrates how to…

形式语言与自动机理论 · 计算机科学 2025-07-22 Jozef Makiš , Alexander Meduna , Zbyněk Křivka

Transmission spectroscopy, which consists of measuring the wavelength-dependent absorption of starlight by a planet's atmosphere during a transit, is a powerful probe of atmospheric composition. However, the expected signal is typically…

地球与行星天体物理 · 物理学 2015-05-30 N. P. Gibson , S. Aigrain , S. Roberts , T. M. Evans , M. Osborne , F. Pont

Starting with a collection of traces generated by process executions, process discovery is the task of constructing a simple model that describes the process, where simplicity is often measured in terms of model size. The challenge of…

人工智能 · 计算机科学 2024-04-17 Hanan Alkhammash , Artem Polyvyanyy , Alistair Moffat

Automatic piano transcription models are typically evaluated using simple frame- or note-wise information retrieval (IR) metrics. Such benchmark metrics do not provide insights into the transcription quality of specific musical aspects such…

声音 · 计算机科学 2024-10-10 Patricia Hu , Lukáš Samuel Marták , Carlos Cancino-Chacón , Gerhard Widmer

Modern music streaming services are heavily based on recommendation engines to serve content to users. Sequential recommendation -- continuously providing new items within a single session in a contextually coherent manner -- has been an…

信息检索 · 计算机科学 2024-09-12 Pavan Seshadri , Shahrzad Shashaani , Peter Knees

Consistency models have exhibited remarkable capabilities in facilitating efficient image/video generation, enabling synthesis with minimal sampling steps. It has proven to be advantageous in mitigating the computational burdens associated…

声音 · 计算机科学 2024-04-23 Zhengcong Fei , Mingyuan Fan , Junshi Huang

The proliferation of capable and efficient machine learning (ML) models marks one of the strongest methodological shifts in signal processing (SP) in its nearly 100-year history. ML models support the development of SP systems that…

信号处理 · 电气工程与系统科学 2026-05-01 Daniel Waxman , Fernando Llorente , Petar M. Djurić

This paper introduces a novel recurrent model for music composition that is tailored to the structure of polyphonic music. We propose an efficient new conditional probabilistic factorization of musical scores, viewing a score as a…

声音 · 计算机科学 2019-11-28 John Thickstun , Zaid Harchaoui , Dean P. Foster , Sham M. Kakade

Choosing a proper set of kernel functions is an important problem in learning Gaussian Process (GP) models since each kernel structure has different model complexity and data fitness. Recently, automatic kernel composition methods provide…

机器学习 · 计算机科学 2021-02-25 Anh Tong , Toan Tran , Hung Bui , Jaesik Choi

Musical expressivity and coherence are indispensable in music composition and performance, while often neglected in modern AI generative models. In this work, we introduce a listening-based data-processing technique that captures the…

声音 · 计算机科学 2025-03-18 Jingwei Liu

Score-based generative modeling (SGM) is a highly successful approach for learning a probability distribution from data and generating further samples. We prove the first polynomial convergence guarantees for the core mechanic behind SGM:…

机器学习 · 计算机科学 2023-05-04 Holden Lee , Jianfeng Lu , Yixin Tan

Professional athletes increasingly use automated analysis of meta- and signal data to improve their training and game performance. As in other related human-to-human research fields, signal data, in particular, contain important…

声音 · 计算机科学 2022-02-21 Lukas Stappen , Manuel Milling , Valentin Munst , Korakot Hoffmann , Bjorn W. Schuller

Music Information Retrieval (MIR) has seen a recent surge in deep learning-based approaches, which often involve encoding symbolic music (i.e., music represented in terms of discrete note events) in an image-like or language like fashion.…

音频与语音处理 · 电气工程与系统科学 2023-09-12 Huan Zhang , Emmanouil Karystinaios , Simon Dixon , Gerhard Widmer , Carlos Eduardo Cancino-Chacón

We introduce a film score generation framework to harmonize visual pixels and music melodies utilizing a latent diffusion model. Our framework processes film clips as input and generates music that aligns with a general theme while offering…

多媒体 · 计算机科学 2024-11-13 F. Qi , L. Ni , C. Xu

Simulation-based inference (SBI) has become a widely used framework in applied sciences for estimating the parameters of stochastic models that best explain experimental observations. A central question in this setting is how to effectively…

机器学习 · 统计学 2026-01-07 Camille Touron , Gabriel V. Cardoso , Julyan Arbel , Pedro L. C. Rodrigues

Statistical models and information theory have provided a useful set of tools for studying music from a quantitative perspective. These approaches have been employed to generate compositions, analyze structural patterns, and model cognitive…

物理与社会 · 物理学 2025-09-30 Linus Chen-Plotkin , Suman S. Kulkarni , Dani S. Bassett

Diffusion models are widely used in applications ranging from image generation to inverse problems. However, training diffusion models typically requires clean ground-truth images, which are unavailable in many applications. We introduce…

图像与视频处理 · 电气工程与系统科学 2025-05-20 Chicago Y. Park , Shirin Shoushtari , Hongyu An , Ulugbek S. Kamilov

Diffusion models (DMs) have emerged as powerful tools for modeling complex data distributions and generating realistic new samples. Over the years, advanced architectures and sampling methods have been developed to make these models…

机器学习 · 计算机科学 2025-12-11 Roi Benita , Michael Elad , Joseph Keshet

Automatic music transcription converts audio recordings into symbolic representations, facilitating music analysis, retrieval, and generation. A musical note is characterized by pitch, onset, and offset in an audio domain, whereas it is…

声音 · 计算机科学 2025-02-19 Leekyung Kim , Sungwook Jeon , Wan Heo , Jonghun Park