中文
相关论文

相关论文: Deep Polyphonic ADSR Piano Note Transcription

200 篇论文

Automatic music transcription (AMT) has achieved high accuracy for piano due to the availability of large, high-quality datasets such as MAESTRO and MAPS, but comparable datasets are not yet available for other instruments. In recent work,…

音频与语音处理 · 电气工程与系统科学 2024-02-26 Xavier Riley , Drew Edwards , Simon Dixon

Since the introduction of the Transformer architecture for large language models, the softmax-based attention layer has faced increasing scrutinity due to its quadratic-time computational complexity. Attempts have been made to replace it…

机器学习 · 计算机科学 2026-02-02 Robert Forchheimer

In a recent paper, the authors proposed a new class of low-complexity iterative thresholding algorithms for reconstructing sparse signals from a small set of linear measurements \cite{DMM}. The new algorithms are broadly referred to as AMP,…

信息论 · 计算机科学 2009-11-24 David L. Donoho , Arian Maleki , Andrea Montanari

Recent work in word spotting in handwritten documents has yielded impressive results. This progress has largely been made by supervised learning systems, which are dependent on manually annotated data, making deployment to new collections a…

计算机视觉与模式识别 · 计算机科学 2020-03-26 Tomas Wilkinson , Carl Nettelblad

Viewing polyphonic piano transcription as a multitask learning problem, where we need to simultaneously predict onsets, intermediate frames and offsets of notes, we investigate the performance impact of additional prediction targets, using…

声音 · 计算机科学 2019-02-13 Rainer Kelz , Sebastian Böck , Gerhard Widmer

Reliable detection of the prodromal stages of Alzheimer's disease (AD) remains difficult even today because, unlike other neurocognitive impairments, there is no definitive diagnosis of AD in vivo. In this context, existing research has…

机器学习 · 计算机科学 2021-08-03 Amish Mittal , Sourav Sahoo , Arnhav Datar , Juned Kadiwala , Hrithwik Shalu , Jimson Mathew

We consider the task of learning mappings from sequential data to real-valued responses. We present and evaluate an approach to learning a type of hidden Markov model (HMM) for regression. The learning process involves inferring the…

机器学习 · 计算机科学 2012-06-18 Keith Noto , Mark Craven

In this paper, we present a neural network approach for synchronizing audio recordings of human piano performances with their corresponding loosely aligned MIDI files. The task is addressed using a Convolutional Recurrent Neural Network…

声音 · 计算机科学 2025-06-30 Sebastian Murgul , Moritz Reiser , Michael Heizmann , Christoph Seibert

Sparse model is widely used in hyperspectral image classification.However, different of sparsity and regularization parameters has great influence on the classification results.In this paper, a novel adaptive sparse deep network based on…

图像与视频处理 · 电气工程与系统科学 2019-10-22 Jingwen Yan , Zixin Xie , Jingyao Chen , Yinan Liu , Lei Liu

This paper presents a deep learning assisted synthesis approach for direct end-to-end generation of RF/mm-wave passive matching network with 3D EM structures. Different from prior approaches that synthesize EM structures from target circuit…

机器学习 · 计算机科学 2022-01-07 Siawpeng Er , Edward Liu , Minshuo Chen , Yan Li , Yuqi Liu , Tuo Zhao , Hua Wang

This paper presents a statistical method for use in music transcription that can estimate score times of note onsets and offsets from polyphonic MIDI performance signals. Because performed note durations can deviate largely from…

人工智能 · 计算机科学 2017-07-10 Eita Nakamura , Kazuyoshi Yoshii , Simon Dixon

We investigate nonlinear regression for nonstationary sequential data. In most real-life applications such as business domains including finance, retail, energy and economy, timeseries data exhibits nonstationarity due to the temporally…

机器学习 · 计算机科学 2020-06-19 Fatih Ilhan , Oguzhan Karaahmetoglu , Ismail Balaban , Suleyman Serdar Kozat

A differentiable digital signal processing (DDSP) autoencoder is a musical sound synthesizer that combines a deep neural network (DNN) and spectral modeling synthesis. It allows us to flexibly edit sounds by changing the fundamental…

Automatic music transcription (AMT) is the task of transcribing audio recordings into symbolic representations. Recently, neural network-based methods have been applied to AMT, and have achieved state-of-the-art results. However, many…

声音 · 计算机科学 2021-08-03 Qiuqiang Kong , Bochen Li , Xuchen Song , Yuan Wan , Yuxuan Wang

We describe a generalization of the Hierarchical Dirichlet Process Hidden Markov Model (HDP-HMM) which is able to encode prior information that state transitions are more likely between "nearby" states. This is accomplished by defining a…

机器学习 · 统计学 2017-07-24 Colin Reimer Dawson , Chaofan Huang , Clayton T. Morrison

The proliferation of malware variants poses a significant challenges to traditional malware detection approaches, such as signature-based methods, necessitating the development of advanced machine learning techniques. In this research, we…

机器学习 · 计算机科学 2024-12-30 Ritik Mehta , Olha Jureckova , Mark Stamp

Audio-to-score alignment (A2SA) is a multimodal task consisting in the alignment of audio signals to music scores. Recent literature confirms the benefits of Automatic Music Transcription (AMT) for A2SA at the frame-level. In this work, we…

声音 · 计算机科学 2022-01-03 Federico Simonetta , Stavros Ntalampiras , Federico Avanzini

Building competitive hybrid hidden Markov model~(HMM) systems for automatic speech recognition~(ASR) requires a complex multi-stage pipeline consisting of several training criteria. The recent sequence-to-sequence models offer the advantage…

声音 · 计算机科学 2023-06-19 Tina Raissi , Christoph Lüscher , Moritz Gunz , Ralf Schlüter , Hermann Ney

We consider the problem of estimating the maximum posterior probability (MAP) state sequence for a finite state and finite emission alphabet hidden Markov model (HMM) in the Bayesian setup, where both emission and transition matrices have…

机器学习 · 统计学 2020-04-20 Alexey Koloydenko , Kristi Kuljus , Jüri Lember

There have been significant advances in deep learning for music demixing in recent years. However, there has been little attention given to how these neural networks can be adapted for real-time low-latency applications, which could be…

音频与语音处理 · 电气工程与系统科学 2024-02-28 Satvik Venkatesh , Arthur Benilov , Philip Coleman , Frederic Roskam