English
Related papers

Related papers: Predicting Chroma from Luma in AV1

200 papers

The upcoming video coding standard, Versatile Video Coding (VVC), has shown great improvement compared to its predecessor, High Efficiency Video Coding (HEVC), in terms of bitrate saving. Despite its substantial performance, compressed…

Image and Video Processing · Electrical Eng. & Systems 2021-12-09 Fatemeh Nasiri , Wassim Hamidouche , Luce Morin , Nicolas Dhollande , Gildas Cocherel

It has recently been discovered that using a pre-trained vision-language model (VLM), e.g., CLIP, to align a whole query image with several finer text descriptions generated by a large language model can significantly enhance zero-shot…

Computer Vision and Pattern Recognition · Computer Science 2024-06-06 Jinhao Li , Haopeng Li , Sarah Erfani , Lei Feng , James Bailey , Feng Liu

Low-light image enhancement (LLIE) aims to improve the illuminance of images due to insufficient light exposure. Recently, various lightweight learning-based LLIE methods have been proposed to handle the challenges of unfavorable prevailing…

Computer Vision and Pattern Recognition · Computer Science 2023-05-24 Yuantong Zhang , Baoxin Teng , Daiqin Yang , Zhenzhong Chen , Haichuan Ma , Gang Li , Wenpeng Ding

Color Filter Arrays (CFA) are optical filters in digital cameras that capture specific color channels. Current commercial CFAs are hand-crafted patterns with different physical and application-specific considerations. This study proposes a…

Image and Video Processing · Electrical Eng. & Systems 2024-06-21 Cemre Omer Ayna , Bahadir Kursat Gunturk , Ali Cafer Gurbuz

Light field photography has been studied thoroughly in recent years. One of its drawbacks is the need for multi-lens in the imaging. To compensate that, compressed light field photography has been proposed to tackle the trade-offs between…

Computer Vision and Pattern Recognition · Computer Science 2019-02-22 Ofir Nabati , David Mendlovic , Raja Giryes

Recent work in image and video generation has been adopting the autoregressive LLM architecture due to its generality and potentially easy integration into multi-modal systems. The crux of applying autoregressive training in language…

Computation and Language · Computer Science 2024-08-22 Xiaochuang Han , Marjan Ghazvininejad , Pang Wei Koh , Yulia Tsvetkov

In this study, we propose a feature extraction framework based on contrastive learning with adaptive positive and negative samples (CL-FEFA) that is suitable for unsupervised, supervised, and semi-supervised single-view feature extraction.…

Machine Learning · Computer Science 2022-01-12 Hongjie Zhang

Recently, semantic communication has been widely applied in wireless image transmission systems as it can prioritize the preservation of meaningful semantic information in images over the accuracy of transmitted symbols, leading to improved…

Information Theory · Computer Science 2023-04-20 Shunpu Tang , Qianqian Yang , Lisheng Fan , Xianfu Lei , Yansha Deng , Arumugam Nallanathan

The Alliance for Open Media (AOMedia) has developed the AV2 video coding standard to supersede AV1, aiming for substantial compression efficiency gains across diverse media applications. This paper details the quality and performance…

Image and Video Processing · Electrical Eng. & Systems 2026-05-18 Zhijun Lei , Vibhoothi Vibhoothi , Dzung Hoang , Yixin Du , Ramzi Khsib

We present Dark from Light (DfL) - a novel method to infer the dark sector in wide-field galaxy surveys, leveraging a machine learning approach trained on contemporary cosmological simulations. The aim of this algorithm is to provide a…

Cosmology and Nongalactic Astrophysics · Physics 2025-08-28 Asa F. L. Bluck , Joanna M. Piotrowska , Paul Goubert , Roberto Maiolino , Camilo Casimiro , Thomas Pinto Franco , Nicolas Cea

We present Cycle-Contrastive Learning (CCL), a novel self-supervised method for learning video representation. Following a nature that there is a belong and inclusion relation of video and its frames, CCL is designed to find correspondences…

Computer Vision and Pattern Recognition · Computer Science 2020-10-29 Quan Kong , Wenpeng Wei , Ziwei Deng , Tomoaki Yoshinaga , Tomokazu Murakami

Sparse code multiple access (SCMA), as a code-domain non-orthogonal multiple access (NOMA) scheme, has received considerable research attention for enabling massive connectivity in future wireless communication systems. In this paper, we…

Signal Processing · Electrical Eng. & Systems 2022-01-11 Saumya Chaturvedi , Dil Nashin Anwar , Vivek Ashok Bohara , Anand Srivastava , Zilong Liu

Analyzing the behavior of cryptographic functions in stripped binaries is a challenging but essential task. Cryptographic algorithms exhibit greater logical complexity compared to typical code, yet their analysis is unavoidable in areas…

Cryptography and Security · Computer Science 2025-04-28 Xiuwei Shang , Guoqiang Chen , Shaoyin Cheng , Shikai Guo , Yanming Zhang , Weiming Zhang , Nenghai Yu

The intracluster light (ICL) is a luminous component of galaxy clusters composed of stars that are gravitationally bound to the cluster potential but do not belong to the individual galaxies. Previous studies of the ICL have shown that its…

In the rapidly evolving field of artificial intelligence, multimodal models, e.g., integrating vision and language into visual-language models (VLMs), have become pivotal for many applications, ranging from image captioning to multimodal…

Machine Learning · Computer Science 2024-04-24 Duy Phuong Nguyen , J. Pablo Munoz , Ali Jannesari

CCTV safety monitoring demands anomaly detectors combine reliable clip-level accuracy with predictable per-clip latency despite weak supervision. This work investigates compact vision-language models (VLMs) as practical detectors for this…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Kirill Borodin , Kirill Kondrashov , Nikita Vasiliev , Ksenia Gladkova , Inna Larina , Mikhail Gorodnichev , Grach Mkrtchian

Contrastive Language-Image Pre-training (CLIP) relies on Vision Transformers whose attention mechanism is susceptible to spurious correlations, and scales quadratically with resolution. To address these limitations, We present CLIMP, the…

Computer Vision and Pattern Recognition · Computer Science 2026-01-13 Nimrod Shabtay , Itamar Zimerman , Eli Schwartz , Raja Giryes

Utilizing large language models (LLMs) to compose off-the-shelf visual tools represents a promising avenue of research for developing robust visual assistants capable of addressing diverse visual tasks. However, these methods often overlook…

Computer Vision and Pattern Recognition · Computer Science 2024-04-11 Zhi Gao , Yuntao Du , Xintong Zhang , Xiaojian Ma , Wenjuan Han , Song-Chun Zhu , Qing Li

Recent years have witnessed an increasing interest in end-to-end learned video compression. Most previous works explore temporal redundancy by detecting and compressing a motion map to warp the reference frame towards the target frame. Yet,…

Image and Video Processing · Electrical Eng. & Systems 2022-11-21 Ren Yang , Radu Timofte , Luc Van Gool

This paper proposes image-adaptive contrast limited adaptive histogram equalization (IA-CLAHE). Conventional CLAHE is widely used to boost the performance of various computer vision tasks and to improve visual quality for human perception…

Computer Vision and Pattern Recognition · Computer Science 2026-04-20 Rikuto Otsuka , Yuho Shoji , Yuka Ogino , Takahiro Toizumi , Atsushi Ito
‹ Prev 1 3 4 5 6 7 10 Next ›