中文
相关论文

相关论文: ArtifactNet: Detecting AI-Generated Music via Fore…

200 篇论文

With the advance in user-friendly and powerful video editing tools, anyone can easily manipulate videos without leaving prominent visual traces. Frame-rate up-conversion (FRUC), a representative temporal-domain operation, increases the…

多媒体 · 计算机科学 2021-03-26 Minseok Yoon , Seung-Hun Nam , In-Jae Yu , Wonhyuk Ahn , Myung-Joon Kwon , Heung-Kyu Lee

AI-generated image detection has become increasingly important with the rapid advancement of generative AI. However, detectors built on Vision Foundation Models (VFMs, \emph{e.g.}, CLIP) often struggle to generalize to images created using…

计算机视觉与模式识别 · 计算机科学 2026-03-11 Chao Shuai , Zhenguang Liu , Shaojing Fan , Bin Gong , Weichen Lian , Xiuli Bi , Zhongjie Ba , Kui Ren

Context-based detection methods such as DetectGPT achieve strong generalization in identifying AI-generated text by evaluating content compatibility with a model's learned distribution. In contrast, existing image detectors rely on…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Minsuk Jang , Hyunseo Jeong , Minseok Son , Changick Kim

Electroencephalography (EEG) measures the electrical brain activity in real-time by using sensors placed on the scalp. Artifacts, due to eye movements and blink, muscular/cardiac activity and generic electrical disturbances, have to be…

计算机视觉与模式识别 · 计算机科学 2021-04-27 Giuseppe Placidi , Luigi Cinque , Matteo Polsinelli

Electroencephalograms (EEG) are often contaminated by artifacts which make interpreting them more challenging for clinicians. Hence, automated artifact recognition systems have the potential to aid the clinical workflow. In this abstract,…

信号处理 · 电气工程与系统科学 2019-03-20 Subhrajit Roy

The rapid progress of generative models has intensified the need for reliable and robust detection under real-world conditions. However, existing detectors often overfit to generator-specific artifacts and remain highly sensitive to…

计算机视觉与模式识别 · 计算机科学 2025-12-25 Ruiqi Liu , Yi Han , Zhengbo Zhang , Liwei Yao , Zhiyuan Yan , Jialiang Shen , ZhiJin Chen , Boyi Sun , Lubin Weng , Jing Dong , Yan Wang , Shu Wu

With the proliferation of Large Language Model (LLM) based deepfake audio, there is an urgent need for effective detection methods. Previous deepfake audio generation methods typically involve a multi-step generation process, with the final…

Purpose: Off-resonance artifact correction by deep-learning, to facilitate rapid pediatric body imaging with a scan time efficient 3D cones trajectory. Methods: A residual convolutional neural network to correct off-resonance artifacts…

计算机视觉与模式识别 · 计算机科学 2018-10-04 David Y Zeng , Jamil Shaikh , Dwight G Nishimura , Shreyas S Vasanawala , Joseph Y Cheng

This research paper presents a novel audio fingerprinting system for Automatic Content Recognition (ACR). By using signal processing techniques and statistical transformations, our proposed method generates compact fingerprints of audio…

声音 · 计算机科学 2023-05-18 Anoubhav Agarwaal , Prabhat Kanaujia , Sartaki Sinha Roy , Susmita Ghose

This paper shows that it is possible to train large and deep convolutional neural networks (CNN) for JPEG compression artifacts reduction, and that such networks can provide significantly better reconstruction quality compared to previously…

计算机视觉与模式识别 · 计算机科学 2016-05-03 Pavel Svoboda , Michal Hradis , David Barina , Pavel Zemcik

Classifying a weapon based on its muzzle blast is a challenging task that has significant applications in various security and military fields. Most of the existing works rely on ad-hoc deployment of spatially diverse microphone sensors to…

音频与语音处理 · 电气工程与系统科学 2021-03-02 Simone Raponi , Isra Ali , Gabriele Oligeri

Drawing an analogy with automatic image completion systems, we propose Music SketchNet, a neural network framework that allows users to specify partial musical ideas guiding automatic music generation. We focus on generating the missing…

机器学习 · 计算机科学 2021-03-31 Ke Chen , Cheng-i Wang , Taylor Berg-Kirkpatrick , Shlomo Dubnov

With the proliferation of Audio Language Model (ALM) based deepfake audio, there is an urgent need for generalized detection methods. ALM-based deepfake audio currently exhibits widespread, high deception, and type versatility, posing a…

The rapid development of generative models has imposed an urgent demand for detection schemes with strong generalization capabilities. However, existing detection methods generally suffer from overfitting to specific source models, leading…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Xiangyu Hu , Yicheng Hong , Hongchuang Zheng , Wenjun Zeng , Bingyao Liu

The availability of high-quality, AI-generated audio raises security challenges such as misinformation campaigns and voice-cloning fraud. A key defense against the misuse of AI-generated audio is by watermarking it, so that it can be easily…

声音 · 计算机科学 2026-05-20 Kexin Li , Xiao Hu , Ilya Grishchenko , David Lie

With the rapid advancement of Large Language Models (LLMs), AI-driven music generation has become a vibrant and fruitful area of research. However, the representation of musical data remains a significant challenge. To address this, a…

机器学习 · 计算机科学 2025-09-16 Cheng-Yang Tsai , Tzu-Wei Huang , Shao-Yu Wei , Guan-Wei Chen , Hung-Ying Chu , Yu-Cheng Lin

Accurately detecting and classifying damage in analogue media such as paintings, photographs, textiles, mosaics, and frescoes is essential for cultural heritage preservation. While machine learning models excel in correcting degradation if…

计算机视觉与模式识别 · 计算机科学 2024-12-09 Daniela Ivanova , Marco Aversa , Paul Henderson , John Williamson

The advancement of generation models has led to the emergence of highly realistic artificial intelligence (AI)-generated videos. Malicious users can easily create non-existent videos to spread false information. This letter proposes an…

计算机视觉与模式识别 · 计算机科学 2024-03-26 Jianfa Bai , Man Lin , Gang Cao

Most of the current supervised automatic music transcription (AMT) models lack the ability to generalize. This means that they have trouble transcribing real-world music recordings from diverse musical genres that are not presented in the…

声音 · 计算机科学 2021-07-30 Kin Wai Cheuk , Dorien Herremans , Li Su

With the rapid advancement of vision generation models, the potential security risks stemming from synthetic visual content have garnered increasing attention, posing significant challenges for AI-generated image detection. Existing methods…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Xinghan Li , Yue Yu , Xue Song , Haijun Shan , Jingjing Chen