English
Related papers

Related papers: Upsampling layers for music source separation

200 papers

Unlike ordinary computer vision tasks that focus more on the semantic content of images, the image manipulation detection task pays more attention to the subtle information of image manipulation. In this paper, the noise image extracted by…

Computer Vision and Pattern Recognition · Computer Science 2022-07-05 Zhongyuan Zhang , Yi Qian , Yanxiang Zhao , Lin Zhu , Jinjin Wang

Fully-supervised models for source separation are trained on parallel mixture-source data and are currently state-of-the-art. However, such parallel data is often difficult to obtain, and it is cumbersome to adapt trained models to mixtures…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-30 Ge Zhu , Jordan Darefsky , Fei Jiang , Anton Selitskiy , Zhiyao Duan

This work investigates a particular class of artefacts, or ghost sources, in radio interferometric images. Earlier observations with (and simulations of) the Westerbork Synthesis Radio Telescope (WSRT) suggested that these were due to…

Instrumentation and Methods for Astrophysics · Physics 2014-04-22 T. G. Grobler , C. D. Nunhokee , O. M. Smirnov , A. J. van Zyl , A. G. de Bruyn

Self-supervised representation learning often uses data augmentations to induce some invariance to "style" attributes of the data. However, with downstream tasks generally unknown at training time, it is difficult to deduce a priori which…

In this paper, we propose to improve image decomposition algorithms in the case of noisy images. In \cite{gilles1,aujoluvw}, the authors propose to separate structures, textures and noise from an image. Unfortunately, the use of separable…

Image and Video Processing · Electrical Eng. & Systems 2024-11-12 Jerome Gilles

The Deepfake technology has raised serious concerns regarding privacy breaches and trust issues. To tackle these challenges, Deepfake detection technology has emerged. Current methods over-rely on the global feature space, which contains…

Computer Vision and Pattern Recognition · Computer Science 2024-10-16 Weijie Zhou , Xiaoqing Luo , Zhancheng Zhang , Jiachen He , Xiaojun Wu

Audiovisual representation learning typically relies on the correspondence between sight and sound. However, there are often multiple audio tracks that can correspond with a visual scene. Consider, for example, different conversations on…

Sound · Computer Science 2024-06-11 Nikhil Singh , Chih-Wei Wu , Iroro Orife , Mahdi Kalayeh

Beamforming is a signal processing technique. It has been studied in many areas such as radar, sonar, seismology and wireless communications, to name but a few. It can be used for a myriad of purposes, such as detecting the presence of a…

Other Computer Science · Computer Science 2012-12-27 Hidri Adel , Meddeb Souad , Abdulqadir Alaqeeli , Amiri Hamid

The human visual system contains a hierarchical sequence of modules that take part in visual perception at different levels of abstraction, i.e., superordinate, basic, and subordinate levels. One important question is to identify the…

Neurons and Cognition · Quantitative Biology 2018-03-12 Matin N. Ashtiani , Saeed Reza Kheradpisheh , Timothée Masquelier , Mohammad Ganjtabesh

Music tag words that describe music audio by text have different levels of abstraction. Taking this issue into account, we propose a music classification approach that aggregates multi-level and multi-scale features using pre-trained…

Sound · Computer Science 2017-06-22 Jongpil Lee , Juhan Nam

Sound synthesizers are widespread in modern music production but they increasingly require expert skills to be mastered. This work focuses on interpolation between presets, i.e., sets of values of all sound synthesis parameters, to enable…

Sound · Computer Science 2023-03-10 Gwendal Le Vaillant , Thierry Dutoit

Visual artifacts remain a persistent challenge in diffusion models, even with training on massive datasets. Current solutions primarily rely on supervised detectors, yet lack understanding of why these artifacts occur in the first place. In…

Computer Vision and Pattern Recognition · Computer Science 2025-03-21 Yu Cao , Zengqun Zhao , Ioannis Patras , Shaogang Gong

Deepfake speech attribution remains challenging for existing solutions. Classifier-based solutions often fail to generalize to domain-shifted samples, and watermarking-based solutions are easily compromised by distortions like codec…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-16 Wanying Ge , Xin Wang , Junichi Yamagishi

Color theme or color palette can deeply influence the quality and the feeling of a photograph or a graphical design. Although color palettes may come from different sources such as online crowd-sourcing, photographs and graphical designs,…

Computer Vision and Pattern Recognition · Computer Science 2017-03-20 Huy Q. Phan , Hongbo Fu , Antoni B. Chan

In ultrasound imaging the appearance of homogeneous regions of tissue is subject to speckle, which for certain applications can make the detection of tissue irregularities difficult. To cope with this, it is common practice to apply speckle…

Image and Video Processing · Electrical Eng. & Systems 2022-08-02 Rüdiger Göbl , Christoph Hennersperger , Nassir Navab

In medical imaging, outliers can contain hypo/hyper-intensities, minor deformations, or completely altered anatomy. To detect these irregularities it is helpful to learn the features present in both normal and abnormal images. However this…

Computer Vision and Pattern Recognition · Computer Science 2024-03-15 Jeremy Tan , Benjamin Hou , James Batten , Huaqi Qiu , Bernhard Kainz

Techniques to improve the data quality of interferometric radio observations are considered. Fundaments of fringe frequencies in the uv-plane are discussed and filters are used to attenuate radio-frequency interference (RFI) and off-axis…

Instrumentation and Methods for Astrophysics · Physics 2012-04-02 A. R. Offringa , A. G. de Bruyn , S. Zaroubi

The ever-increasing use of synthetically generated content in different sectors of our everyday life, one for all media information, poses a strong need for deepfake detection tools in order to avoid the proliferation of altered messages.…

Computer Vision and Pattern Recognition · Computer Science 2024-04-18 Andrea Ciamarra , Roberto Caldelli , Federico Becattini , Lorenzo Seidenari , Alberto Del Bimbo

The identification of structural differences between a music performance and the score is a challenging yet integral step of audio-to-score alignment, an important subtask of music information retrieval. We present a novel method to detect…

Sound · Computer Science 2021-02-16 Ruchit Agrawal , Daniel Wolff , Simon Dixon

With the increased use of virtual and augmented reality applications, the importance of point cloud data rises. High-quality capturing of point clouds is still expensive and thus, the need for point cloud super-resolution or point cloud…

Image and Video Processing · Electrical Eng. & Systems 2022-05-04 Viktoria Heimann , Andreas Spruck , André Kaup