English
Related papers

Related papers: ArtifactNet: Detecting AI-Generated Music via Fore…

200 papers

With the advancement of IPTV and HDTV technology, previous subtle errors in videos are now becoming more prominent because of the structure oriented and compression based artifacts. In this paper, we focus towards the development of a…

Computer Vision and Pattern Recognition · Computer Science 2018-08-31 Md Mehedi Hasan , Tasneem Rahman , Kiok Ahn , Oksam Chae

Artifact detectors have been shown to enhance the performance of image-generative models by serving as reward models during fine-tuning. These detectors enable the generative model to improve overall output fidelity and aesthetics. However,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-25 Dennis Menn , Feng Liang , Diana Marculescu

Machine-generated music (MGM) has become a groundbreaking innovation with wide-ranging applications, such as music therapy, personalised editing, and creative inspiration within the music industry. However, the unregulated proliferation of…

Sound · Computer Science 2026-04-30 Yupei Li , Qiyang Sun , Hanqian Li , Lucia Specia , Björn W. Schuller

In this paper, AzarNet, a deep neural network (DNN), is proposed to recognizing seven different Dastgahs of Iranian classical music in Maryam Iranian classical music (MICM) dataset. Over the last years, there has been remarkable interest in…

Sound · Computer Science 2019-01-10 Shahla RezezadehAzar , Ali Ahmadi , Saber Malekzadeh , Maryam Samami

Audio classification is the task of identifying the sound categories that are associated with a given audio signal. This paper presents an investigation on large-scale audio classification based on the recently released AudioSet database.…

Sound · Computer Science 2018-10-31 Yuzhong Wu , Tan Lee

Depression, a common mental disorder, significantly influences individuals and imposes considerable societal impacts. The complexity and heterogeneity of the disorder necessitate prompt and effective detection, which nonetheless, poses a…

Sound · Computer Science 2023-08-25 Xiao Xu , Yang Wang , Xinru Wei , Fei Wang , Xizhe Zhang

Recent generative models show impressive performance in generating photographic images. Humans can hardly distinguish such incredibly realistic-looking AI-generated images from real ones. AI-generated images may lead to ubiquitous…

Computer Vision and Pattern Recognition · Computer Science 2024-03-08 Nan Zhong , Yiran Xu , Sheng Li , Zhenxing Qian , Xinpeng Zhang

Diffusion model-generated images can appear indistinguishable from authentic photographs, but these images often contain artifacts and implausibilities that reveal their AI-generated provenance. Given the challenge to public trust in media…

Human-Computer Interaction · Computer Science 2025-02-18 Negar Kamali , Karyn Nakamura , Aakriti Kumar , Angelos Chatzimparmpas , Jessica Hullman , Matthew Groh

Modern streaming services are increasingly labeling videos based on their visual or audio content. This typically augments the use of technologies such as AI and ML by allowing to use natural speech for searching by keywords and video…

Sound · Computer Science 2021-09-22 Ievgeniia Kuzminykh , Dan Shevchuk , Stavros Shiaeles , Bogdan Ghita

Synthetic video generation is progressing very rapidly. The latest models can produce very realistic high-resolution videos that are virtually indistinguishable from real ones. Although several video forensic detectors have been recently…

Computer Vision and Pattern Recognition · Computer Science 2025-11-07 Riccardo Corvi , Davide Cozzolino , Ekta Prashnani , Shalini De Mello , Koki Nagano , Luisa Verdoliva

Computed tomography (CT) images are frequently degraded by acquisition artifacts, including noise, blur, streaking, aliasing, and metal artifacts. Yet CT enhancement is still largely evaluated using image quality metrics with limited…

Despite the popularity of guitar effects, there is very little existing research on classification and parameter estimation of specific plugins or effect units from guitar recordings. In this paper, convolutional neural networks were used…

Sound · Computer Science 2022-04-19 Marco Comunità , Dan Stowell , Joshua D. Reiss

Generative Artificial Intelligence (Gen-AI) models are increasingly used to produce content across domains, including text, images, and audio. While these models represent a major technical breakthrough, they gain their generative…

Machine Learning · Computer Science 2024-12-13 Pascal Epple , Igor Shilov , Bozhidar Stevanoski , Yves-Alexandre de Montjoye

The National Institute of Standards and Technology (NIST) Computer Forensic Tool Testing (CFTT) programme has become the de facto standard for providing digital forensic tool testing and validation. However to date, no comprehensive…

Cryptography and Security · Computer Science 2025-12-22 Akila Wickramasekara , Tharusha Mihiranga , Aruna Withanage , Buddhima Weerasinghe , Frank Breitinger , John Sheppard , Mark Scanlon

This paper introduces the jazznet Dataset, a dataset of fundamental jazz piano music patterns for developing machine learning (ML) algorithms in music information retrieval (MIR). The dataset contains 162520 labeled piano patterns,…

Sound · Computer Science 2023-02-20 Tosiron Adegbija

Pattern recognition from audio signals is an active research topic encompassing audio tagging, acoustic scene classification, music classification, and other areas. Spectrogram and mel-frequency cepstral coefficients (MFCC) are among the…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-18 Md. Istiaq Ansari , Taufiq Hasan

An ideal audio retrieval system efficiently and robustly recognizes a short query snippet from an extensive database. However, the performance of well-known audio fingerprinting systems falls short at high signal distortion levels. This…

Audio and Speech Processing · Electrical Eng. & Systems 2024-11-22 Anup Singh , Kris Demuynck , Vipul Arora

Professionally generated content (PGC) streamed online can contain visual artefacts that degrade the quality of user experience. These artefacts arise from different stages of the streaming pipeline, including acquisition, post-production,…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Chen Feng , Duolikun Danier , Fan Zhang , Alex Mackin , Andy Collins , David Bull

For nearly a decade, deepfake detection has been framed as a classification task: given an audio or video clip, decide whether it is real or synthetic. Top detectors often report high accuracy on standard benchmarks; however, performance…

Computers and Society · Computer Science 2026-05-12 Jessee Ho , Shweta Khushu , Shaina Raza

As the labor force decreases, the demand for labor-saving automatic anomalous sound detection technology that conducts maintenance of industrial equipment has grown. Conventional approaches detect anomalies based on the reconstruction…

Audio and Speech Processing · Electrical Eng. & Systems 2020-05-20 Kaori Suefusa , Tomoya Nishida , Harsh Purohit , Ryo Tanabe , Takashi Endo , Yohei Kawaguchi
‹ Prev 1 8 9 10 Next ›