English

A Combined Deep Learning based End-to-End Video Coding Architecture for YUV Color Space

Computer Vision and Pattern Recognition 2021-04-05 v1 Artificial Intelligence Machine Learning Multimedia

Abstract

Most of the existing deep learning based end-to-end video coding (DLEC) architectures are designed specifically for RGB color format, yet the video coding standards, including H.264/AVC, H.265/HEVC and H.266/VVC developed over past few decades, have been designed primarily for YUV 4:2:0 format, where the chrominance (U and V) components are subsampled to achieve superior compression performances considering the human visual system. While a broad number of papers on DLEC compare these two distinct coding schemes in RGB domain, it is ideal to have a common evaluation framework in YUV 4:2:0 domain for a more fair comparison. This paper introduces a new DLEC architecture for video coding to effectively support YUV 4:2:0 and compares its performance against the HEVC standard under a common evaluation framework. The experimental results on YUV 4:2:0 video sequences show that the proposed architecture can outperform HEVC in intra-frame coding, however inter-frame coding is not as efficient on contrary to the RGB coding results reported in recent papers.

Cite

@article{arxiv.2104.00807,
  title  = {A Combined Deep Learning based End-to-End Video Coding Architecture for YUV Color Space},
  author = {Ankitesh K. Singh and Hilmi E. Egilmez and Reza Pourreza and Muhammed Coban and Marta Karczewicz and Taco S. Cohen},
  journal= {arXiv preprint arXiv:2104.00807},
  year   = {2021}
}

Comments

5 pages, submitted to as a conference paper. arXiv admin note: text overlap with arXiv:2103.01760

R2 v1 2026-06-24T00:47:34.231Z