English
Related papers

Related papers: Extract fundamental frequency based on CNN combine…

200 papers

We propose the multi-layered cepstrum (MLC) method to estimate multiple fundamental frequencies (MF0) of a signal under challenging contamination such as high-pass filter noise. Taking the operation of cepstrum (i.e., Fourier transform,…

Audio and Speech Processing · Electrical Eng. & Systems 2019-02-05 Chin-Yun Yu , Li Su

This monograph introduces a novel approach to polyphonic music generation by addressing the "Missing Middle" problem through structural inductive bias. Focusing on Beethoven's piano sonatas as a case study, we empirically verify the…

Machine Learning · Computer Science 2026-04-10 Joonwon Seo

Deep learning has become increasingly popular in both supervised and unsupervised machine learning thanks to its outstanding empirical performance. However, because of their intrinsic complexity, most deep learning methods are largely…

Machine Learning · Computer Science 2018-09-07 Yang Young Lu , Yingying Fan , Jinchi Lv , William Stafford Noble

We propose a novel decentralized feature extraction approach in federated learning to address privacy-preservation issues for speech recognition. It is built upon a quantum convolutional neural network (QCNN) composed of a quantum circuit…

Single-particle trajectories measured in microscopy experiments contain important information about dynamic processes undergoing in a range of materials including living cells and tissues. However, extracting that information is not a…

Quantitative Methods · Quantitative Biology 2019-09-25 Patrycja Kowalek , Hanna Loch-Olszewska , Janusz Szwabiński

Music source separation with deep neural networks typically relies only on amplitude features. In this paper we show that additional phase features can improve the separation performance. Using the theoretical relationship between STFT…

In live and studio recordings unexpected sound events often lead to interferences in the signal. For non-stationary interferences, sound source separation techniques can be used to reduce the interference level in the recording. In this…

Sound · Computer Science 2017-11-01 Delia Fano Yela , Sebastian Ewert , Derry FitzGerald , Mark Sandler

In this research, Piano performances have been analyzed only based on visual information. Computer vision algorithms, e.g., Hough transform and binary thresholding, have been applied to find where the keyboard and specific keys are located.…

Computer Vision and Pattern Recognition · Computer Science 2019-10-29 Seongjae Kang , Jaeyoon Kim , Sung-eui Yoon

We propose a method for the blind separation of sounds of musical instruments in audio signals. We describe the individual tones via a parametric model, training a dictionary to capture the relative amplitudes of the harmonics. The model…

Audio and Speech Processing · Electrical Eng. & Systems 2021-08-10 Sören Schulze , Johannes Leuschner , Emily J. King

Deep convolutional neural networks are a powerful model class for a range of computer vision problems, but it is difficult to interpret the image filtering process they implement, given their sheer size. In this work, we introduce a method…

Computer Vision and Pattern Recognition · Computer Science 2023-04-18 Chris Hamblin , Talia Konkle , George Alvarez

We present a computational imaging mode for large scale electron microscopy data, which retrieves a complex wave from noisy/sparse intensity recordings using a deep learning approach and subsequently reconstructs an image of the specimen…

Materials Science · Physics 2022-02-28 Thomas Friedrich , Chu-Ping Yu , Johan Verbeek , Timothy Pennycook , Sandra Van Aert

Minutiae play a major role in fingerprint identification. Extracting reliable minutiae is difficult for latent fingerprints which are usually of poor quality. As the limitation of traditional handcrafted features, a fully convolutional…

Computer Vision and Pattern Recognition · Computer Science 2017-09-08 Yao Tang , Fei Gao , Jufu Feng

Deep learning methods have been widely used for Human Activity Recognition (HAR) using recorded signals from Iner-tial Measurement Units (IMUs) sensors that are installed on various parts of the human body. For this type of HAR, sev-eral…

Machine Learning · Computer Science 2025-02-10 Saeed Arabzadeh , Farshad Almasganj , Mohammad Mahdi Ahmadi

Feature extraction in noisy image datasets presents many challenges in model reliability. In this paper, we use the discrete Fourier transform in conjunction with persistent homology analysis to extract specific frequencies that correspond…

Computer Vision and Pattern Recognition · Computer Science 2025-12-09 Anil Chintapalli , Peter Tenholder , Henry Chen , Arjun Rao

Piano fingering -- knowing which finger to use to play each note in a musical piece, is a hard and important skill to master when learning to play the piano. While some sheet music is available with expert-annotated fingering information,…

Computer Vision and Pattern Recognition · Computer Science 2023-03-08 Amit Moryossef , Yanai Elazar , Yoav Goldberg

Hierarchical quantum classifiers, such as quantum convolutional neural networks (QCNNs), represent recent progress toward designing effective and feasible architectures for quantum classification. However, their performance on near-term…

Quantum Physics · Physics 2026-02-26 Taehyun Kim , Israel F. Araujo , Daniel K. Park

This paper proposes a deep convolutional neural network for performing note-level instrument assignment. Given a polyphonic multi-instrumental music signal along with its ground truth or predicted notes, the objective is to assign an…

Sound · Computer Science 2021-07-30 Carlos Lordelo , Emmanouil Benetos , Simon Dixon , Sven Ahlbäck

In this paper, a new algorithm for extracting features from sequences of multidimensional observations is presented. The independently developed Dynamic Mode Decomposition and Matrix Pencil methods provide a least-squares model-based…

Numerical Analysis · Mathematics 2018-04-20 Leonid Pogorelyuk , Clarence W. Rowley

The analysis of speech production based on physical models of the vocal folds and vocal tract is essential for studies on vocal-fold behavior and linguistic research. This paper proposes a speech production analysis method using…

Sound · Computer Science 2025-11-04 Kazuya Yokota , Ryosuke Harakawa , Masaaki Baba , Masahiro Iwahashi

The work presented here applies deep learning to the task of automated cardiac auscultation, i.e. recognizing abnormalities in heart sounds. We describe an automated heart sound classification algorithm that combines the use of…

Sound · Computer Science 2017-10-20 Jonathan Rubin , Rui Abreu , Anurag Ganguli , Saigopal Nelaturi , Ion Matei , Kumar Sricharan