English
Related papers

Related papers: BBAND Index: A No-Reference Banding Artifact Predi…

200 papers

While the BD-rate performance of recent learned video codec models in both low-delay and random-access modes exceed that of respective modes of traditional codecs on average over common benchmarks, the performance improvements for…

Image and Video Processing · Electrical Eng. & Systems 2025-10-13 Ahmet Bilican , M. Akın Yılmaz , A. Murat Tekalp

Anomaly detection and localization in images is a growing field in computer vision. In this area, a seemingly understudied problem is anomaly clustering, i.e., identifying and grouping different types of anomalies in a fully unsupervised…

Computer Vision and Pattern Recognition · Computer Science 2024-04-19 Andrei-Timotei Ardelean , Tim Weyrich

Occlusion handling is one of the challenges of object detection and segmentation, and scene understanding. Because objects appear differently when they are occluded in varying degree, angle, and locations. Therefore, determining the…

Computer Vision and Pattern Recognition · Computer Science 2022-04-28 Kaziwa Saleh , Zoltan Vamossy

In this paper, we address a novel image restoration problem relevant to machine learning dataset curation: the detection and removal of noisy mirrored padding artifacts. While data augmentation techniques like padding are necessary for…

Computer Vision and Pattern Recognition · Computer Science 2025-04-07 Lucas Choi , Ross Greer

Recent deepfake detection methods demonstrate improved cross-dataset generalization, yet the underlying mechanisms remain underexplored. We introduce the Alpha Blending Hypothesis, positing that state-of-the-art frame-based detectors…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Andrii Yermakov , Jan Cech , Mario Fritz , Jiri Matas

Modern reconstruction techniques can effectively model complex 3D scenes from sparse 2D views. However, automatically assessing the quality of novel views and identifying artifacts is challenging due to the lack of ground truth images and…

Computer Vision and Pattern Recognition · Computer Science 2025-07-30 Nicolai Hermann , Jorge Condor , Piotr Didyk

Humans have the capacity to question what we see and to recognize when our vision is unreliable (e.g., when we realize that we are experiencing a visual illusion). Inspired by this capacity, we present MetaCOG: a hierarchical probabilistic…

Artificial Intelligence · Computer Science 2024-07-10 Marlene D. Berke , Zhangir Azerbayev , Mario Belledonne , Zenna Tavares , Julian Jara-Ettinger

This paper presents a blind detection and compensation technique for camera lens geometric distortions. The lens distortion introduces higher-order correlations in the frequency domain and in turn it can be detected using higher-order…

Computer Vision and Pattern Recognition · Computer Science 2007-05-23 Lili Ma , YangQuan Chen , Kevin L. Moore

While low-level image features have proven to be effective representations for visual recognition tasks such as object recognition and scene classification, they are inadequate to capture complex semantic meaning required to solve…

Multimedia · Computer Science 2014-06-17 Tim Althoff , Hyun Oh Song , Trevor Darrell

Video salient object detection aims to find the most visually distinctive objects in a video. To explore the temporal dependencies, existing methods usually resort to recurrent neural networks or optical flow. However, these approaches…

Computer Vision and Pattern Recognition · Computer Science 2021-11-04 Yi-Wen Chen , Xiaojie Jin , Xiaohui Shen , Ming-Hsuan Yang

We introduce FakeParts, a new class of deepfakes characterized by subtle, localized manipulations to specific spatial regions or temporal segments of otherwise authentic videos. Unlike fully synthetic content, these partial manipulations -…

Computer Vision and Pattern Recognition · Computer Science 2025-12-22 Ziyi Liu , Firas Gabetni , Awais Hussain Sani , Xi Wang , Soobash Daiboo , Gaetan Brison , Gianni Franchi , Vicky Kalogeiton

Digital images contain a lot of redundancies, therefore, compressions are applied to reduce the image size without the loss of reasonable image quality. The same become more prominent in the case of videos that contains image sequences and…

Image and Video Processing · Electrical Eng. & Systems 2021-12-30 Nisar Ahmed , Hafiz Muhammad Shahzad Asif , Hassan Khalid

Existing saliency models have been designed and evaluated for predicting the saliency in distortion-free images. However, in practice, the image quality is affected by a host of factors at several stages of the image processing pipeline…

Computer Vision and Pattern Recognition · Computer Science 2016-04-14 Milind S. Gide , Samuel F. Dodge , Lina J. Karam

Query-based video grounding is an important yet challenging task in video understanding, which aims to localize the target segment in an untrimmed video according to a sentence query. Most previous works achieve significant progress by…

Computer Vision and Pattern Recognition · Computer Science 2022-03-09 Shentong Mo , Daizong Liu , Wei Hu

Deepfake is the manipulated video made with a generative deep learning technique such as Generative Adversarial Networks (GANs) or Auto Encoder that anyone can utilize. Recently, with the increase of Deepfake videos, some classifiers…

Computer Vision and Pattern Recognition · Computer Science 2021-04-06 Young-Jin Heo , Young-Ju Choi , Young-Woon Lee , Byung-Gyu Kim

This paper develops a new video compression approach based on underdetermined blind source separation. Underdetermined blind source separation, which can be used to efficiently enhance the video compression ratio, is combined with various…

Multimedia · Computer Science 2012-05-22 Jing Liu , Fei Qiao , Qi Wei , Huazhong Yang

Endoscopy is a routine imaging technique used for both diagnosis and minimally invasive surgical treatment. Artifacts such as motion blur, bubbles, specular reflections, floating objects and pixel saturation impede the visual interpretation…

Computer Vision and Pattern Recognition · Computer Science 2020-11-24 Sharib Ali , Felix Zhou , Adam Bailey , Barbara Braden , James East , Xin Lu , Jens Rittscher

Highlight detection in sports videos has a broad viewership and huge commercial potential. It is thus imperative to detect highlight scenes more suitably for human interest with high temporal accuracy. Since people instinctively suppress…

Computer Vision and Pattern Recognition · Computer Science 2020-07-03 Tamami Nakano , Atsuya Sakata , Akihiro Kishimoto

In this paper, we propose a novel end-to-end architecture that could generate a variety of plausible video sequences correlating two given discontinuous frames. Our work is inspired by the human ability of inference. Specifically, given two…

Computer Vision and Pattern Recognition · Computer Science 2019-12-17 Weimian Li , Baoyang Chen , Wenmin Wang

Music retrieval and recommendation applications often rely on content features encoded as embeddings, which provide vector representations of items in a music dataset. Numerous complementary embeddings can be derived from processing items…

Information Retrieval · Computer Science 2023-08-15 Andres Ferraro , Jaehun Kim , Sergio Oramas , Andreas Ehmann , Fabien Gouyon