English
Related papers

Related papers: Predicting Chroma from Luma in AV1

200 papers

Modern video codecs including the newly developed AOMedia Video 1 (AV1) utilize hybrid coding techniques to remove spatial and temporal redundancy. However, efficient exploitation of statistical dependencies measured by a mean squared error…

Image and Video Processing · Electrical Eng. & Systems 2019-08-09 Di Chen , Chichen Fu , Zoe Liu , Fengqing Zhu

We combine a high-resolution hydro-simulation of the LambdaCDM cosmology with two radiative transfer schemes (for continuum and line radiation) to predict the properties, spectra and spatial distribution of fluorescent Ly-alpha emission at…

Astrophysics · Physics 2009-11-10 Sebastiano Cantalupo , Cristiano Porciani , Simon J. Lilly , Francesco Miniati

Conventional video encoders typically employ a fixed chroma subsampling format, such as YUV420, which may not optimally reflect variations in chroma detail across different types of content. This can lead to suboptimal chroma quality and…

Image and Video Processing · Electrical Eng. & Systems 2026-02-09 Amritha Premkumar , Christian Herglotz

Recent advances in end-to-end video compression have shown promising results owing to their unified end-to-end learning optimization. However, such generalized frameworks often lack content-specific adaptation, leading to suboptimal…

Computer Vision and Pattern Recognition · Computer Science 2025-12-16 Tiange Zhang , Xiandong Meng , Siwei Ma

Robot manipulation relying on learned object-centric descriptors became popular in recent years. Visual descriptors can easily describe manipulation task objectives, they can be learned efficiently using self-supervision, and they can…

Computer Vision and Pattern Recognition · Computer Science 2024-06-19 David B. Adrian , Andras Gabor Kupcsik , Markus Spies , Heiko Neumann

With the rapid development of Vision-Language Models (VLMs) and the growing demand for their applications, efficient compression of the image inputs has become increasingly important. Existing VLMs predominantly digest and understand…

Computer Vision and Pattern Recognition · Computer Science 2026-05-25 Zifu Zhang , Tongda Xu , Siqi Li , Shengxi Li , Yue Zhang , Mai Xu , Yan Wang

We have recently witnessed that ``Intelligence" and `` Compression" are the two sides of the same coin, where the language large model (LLM) with unprecedented intelligence is a general-purpose lossless compressor for various data…

Computer Vision and Pattern Recognition · Computer Science 2024-11-25 Kecheng Chen , Pingping Zhang , Hui Liu , Jie Liu , Yibing Liu , Jiaxin Huang , Shiqi Wang , Hong Yan , Haoliang Li

Contrastive Language-Image Pre-training (CLIP) has achieved success on multiple downstream tasks by aligning image and text modalities. However, the nature of global contrastive learning limits CLIP's ability to comprehend compositional…

Computer Vision and Pattern Recognition · Computer Science 2025-08-27 Xiaoxing Hu , Kaicheng Yang , Jun Wang , Haoran Xu , Ziyong Feng , Yupei Wang

Rapidly increasing demand for high speed data is pushing 6G wireless networks to support larger link scales, lower latency, and higher spectral efficiency. Visible light communications (VLC) is a strong complement to radio frequency (RF)…

Signal Processing · Electrical Eng. & Systems 2025-10-24 Xuesong Wang

This paper introduces a new machine learning-assisted chromatic dispersion compensation filter, demonstrating its superior power efficiency compared to conventional FFT-based filters for metro link distances. Validations on FPGA confirmed…

Signal Processing · Electrical Eng. & Systems 2024-09-23 Geraldo Gomes , Pedro Freire , Jaroslaw E. Prilepsky , Sergei K. Turitsyn

Latent inpainting in diffusion models still relies almost universally on linearly interpolating VAE latents under a downsampled mask. We propose a key principle for compositing image latents: Pixel-Equivalent Latent Compositing (PELC). An…

Computer Vision and Pattern Recognition · Computer Science 2025-12-08 Rowan Bradbury , Dazhi Zhong

Despite the vast success of standard planar convolutional neural networks, they are not the most efficient choice for analyzing signals that lie on an arbitrarily curved manifold, such as a cylinder. The problem arises when one performs a…

Computer Vision and Pattern Recognition · Computer Science 2021-07-28 Bahar Azari , Deniz Erdogmus

Actively tunable optical filters based on chalcogenide phase-change materials (PCMs) are an emerging technology with applications across chemical spectroscopy and thermal imaging. The refractive index of an embedded PCM thin film is…

Instrumentation and Detectors · Physics 2021-02-23 David Bombara , Calum Williams , Stephen Borg , Hyun Jung Kim

Recent advances in video compression have seen significant coding performance improvements with the development of new standards and learning-based video codecs. However, most of these works focus on application scenarios that allow a…

Multimedia · Computer Science 2025-02-18 Siyue Teng , Yuxuan Jiang , Ge Gao , Fan Zhang , Thomas Davis , Zoe Liu , David Bull

This paper investigates methods for calculating the chromatic symmetric function (CSF) of a graph in chromatic-bases and the $m_\lambda$-basis. Our key contributions include a novel approach for calculating the CSF in chromatic-bases…

Combinatorics · Mathematics 2025-02-25 Nima Amoei Mobaraki , Yasaman Gerivani , Sina Ghasemi Nezhad

In the past decade, SIFT descriptor has been witnessed as one of the most robust local invariant feature descriptors and widely used in various vision tasks. Most traditional image classification systems depend on the luminance-based SIFT…

Computer Vision and Pattern Recognition · Computer Science 2013-10-01 Chen Junzhou , Li Qing , Peng Qiang , Kin Hong Wong

We simulate the flux emitted from galaxy halos in order to quantify the brightness of the circumgalactic medium (CGM). We use dedicated zoom-in cosmological simulations with the hydrodynamical Adaptive Mesh Refinement code RAMSES, which are…

Few-shot learning (FSL) often requires effective adaptation of models using limited labeled data. However, most existing FSL methods rely on entangled representations, requiring the model to implicitly recover the unmixing process to obtain…

Computer Vision and Pattern Recognition · Computer Science 2025-08-06 Tianjiao Jiang , Zhen Zhang , Yuhang Liu , Javen Qinfeng Shi

CLIP (Contrastive Language-Image Pre-Training) has shown remarkable zero-shot transfer capabilities in cross-modal correlation tasks such as visual classification and image retrieval. However, its performance in cross-modal generation tasks…

Computer Vision and Pattern Recognition · Computer Science 2022-11-15 Junyang Wang , Yi Zhang , Ming Yan , Ji Zhang , Jitao Sang

In this paper, we propose $\text{HF}^2$-VAD, a Hybrid framework that integrates Flow reconstruction and Frame prediction seamlessly to handle Video Anomaly Detection. Firstly, we design the network of ML-MemAE-SC (Multi-Level Memory modules…

Computer Vision and Pattern Recognition · Computer Science 2021-08-17 Zhian Liu , Yongwei Nie , Chengjiang Long , Qing Zhang , Guiqing Li
‹ Prev 1 4 5 6 7 8 10 Next ›