中文
相关论文

相关论文: A Study on the Effect of Color Spaces in Learned I…

200 篇论文

While convolution and self-attention are extensively used in learned image compression (LIC) for transform coding, this paper proposes an alternative called Contextual Clustering based LIC (CLIC) which primarily relies on clustering…

图像与视频处理 · 电气工程与系统科学 2024-01-23 Yichi Zhang , Zhihao Duan , Ming Lu , Dandan Ding , Fengqing Zhu , Zhan Ma

In this paper, a unified transformation method in learned image compression(LIC) is proposed from the perspective of modulation. Firstly, the quantization in LIC is considered as a generalized channel with additive uniform noise. Moreover,…

图像与视频处理 · 电气工程与系统科学 2024-03-13 Youneng Bao , Fangyang Meng , Wen Tan , Chao Li , Yonghong Tian , Yongsheng Liang

Recent advances in learning-based methods have markedly enhanced the capabilities of image compression. However, these methods struggle with high bit-depth volumetric medical images, facing issues such as degraded performance, increased…

图像与视频处理 · 电气工程与系统科学 2024-10-24 Kai Wang , Yuanchao Bai , Daxin Li , Deming Zhai , Junjun Jiang , Xianming Liu

Image compression has been the subject of extensive research for several decades, resulting in the development of well-known standards such as JPEG, JPEG2000, and H.264/AVC. However, recent advancements in deep learning have led to the…

图像与视频处理 · 电气工程与系统科学 2024-02-20 Gaocheng Ma , Yinfeng Chai , Tianhao Jiang , Ming Lu , Tong Chen

The learned image compression (LIC) methods have already surpassed traditional techniques in compressing natural scene (NS) images. However, directly applying these methods to screen content (SC) images, which possess distinct…

图像与视频处理 · 电气工程与系统科学 2025-02-24 Shiqi Jiang , Hui Yuan , Shuai Li , Huanqiang Zeng , Sam Kwong

We propose a method for lossy image compression based on recurrent, convolutional neural networks that outperforms BPG (4:2:0 ), WebP, JPEG2000, and JPEG as measured by MS-SSIM. We introduce three improvements over previous research that…

计算机视觉与模式识别 · 计算机科学 2017-03-30 Nick Johnston , Damien Vincent , David Minnen , Michele Covell , Saurabh Singh , Troy Chinen , Sung Jin Hwang , Joel Shor , George Toderici

Learned image compression has achieved extraordinary rate-distortion performance in PSNR and MS-SSIM compared to traditional methods. However, it suffers from intensive computation, which is intolerable for real-world applications and leads…

图像与视频处理 · 电气工程与系统科学 2022-08-01 Hongjiu Yu , Qiancheng Sun , Jin Hu , Xingyuan Xue , Jixiang Luo , Dailan He , Yilong Li , Pengbo Wang , Yuanyuan Wang , Yaxu Dai , Yan Wang , Hongwei Qin

This paper introduces a learned hierarchical B-frame coding scheme in response to the Grand Challenge on Neural Network-based Video Coding at ISCAS 2023. We address specifically three issues, including (1) B-frame coding, (2) YUV 4:2:0…

计算机视觉与模式识别 · 计算机科学 2023-01-02 Mu-Jung Chen , Hong-Sheng Xie , Cheng Chien , Wen-Hsiao Peng , Hsueh-Ming Hang

Learned image compression (LIC) techniques have achieved remarkable progress; however, effectively integrating high-level semantic information remains challenging. In this work, we present a \underline{S}emantic-\underline{E}nhanced…

应用统计 · 统计学 2025-04-03 Haisheng Fu , Jie Liang , Zhenman Fang , Jingning Han

Learned image compression techniques have achieved considerable development in recent years. In this paper, we find that the performance bottleneck lies in the use of a single hyperprior decoder, in which case the ternary Gaussian model…

计算机视觉与模式识别 · 计算机科学 2021-11-02 Zhao Zan , Chao Liu , Heming Sun , Xiaoyang Zeng , Yibo Fan

Recent advances in learned image compression (LIC) have achieved remarkable performance improvements over traditional codecs. Notably, the MLIC series-LICs equipped with multi-reference entropy models-have substantially surpassed…

图像与视频处理 · 电气工程与系统科学 2026-02-26 Wei Jiang , Yongqi Zhai , Jiayu Yang , Feng Gao , Ronggang Wang

In video compression, coding efficiency is improved by reusing pixels from previously decoded frames via motion and residual compensation. We define two levels of hierarchical redundancy in video frames: 1) first-order: redundancy in pixel…

图像与视频处理 · 电气工程与系统科学 2022-09-21 Reza Pourreza , Hoang Le , Amir Said , Guillaume Sautiere , Auke Wiggers

Object-centric architectures can learn to extract distinct object representations from visual scenes, enabling downstream applications on the object level. Similarly to autoencoder-based image models, object-centric approaches have been…

计算机视觉与模式识别 · 计算机科学 2024-12-20 Bastian Jäckl , Yannick Metz , Udo Schlegel , Daniel A. Keim , Maximilian T. Fischer

Raw images preserve linear sensor measurements and high bit-depth information crucial for advanced vision tasks and photography applications, yet their storage remains challenging due to large file sizes, varying bit depths, and…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Chunhang Zheng , Tongda Xu , Mingli Xie , Yan Wang , Dou Li

One of the major differentiators unlocked by learned codecs relative to their hard-coded traditional counterparts is their ability to be optimized directly to appeal to the human visual system. Despite this potential, a perceptual yet…

计算机视觉与模式识别 · 计算机科学 2026-05-07 Kedar Tatwawadi , Parisa Rahimzadeh , Zhanghao Sun , Zhiqi Chen , Ziyun Yang , Sanjay Nair , Divija Hasteer , Oren Rippel

Learned image compression (LIC) has reached the traditional hand-crafted methods such as JPEG2000 and BPG in terms of the coding gain. However, the large model size of the network prohibits the usage of LIC on resource-limited embedded…

图像与视频处理 · 电气工程与系统科学 2020-07-10 Heming Sun , Zhengxue Cheng , Masaru Takeuchi , Jiro Katto

Modern cameras typically offer two types of image states: a minimally processed linear raw RGB image representing the raw sensor data, and a highly-processed non-linear image state, such as the sRGB state. The CIE-XYZ color space is a…

图像与视频处理 · 电气工程与系统科学 2024-05-22 Shir Barzel , Moshe Salhov , Ofir Lindenbaum , Amir Averbuch

Learning effective visual representations that generalize well without human supervision is a fundamental problem in order to apply Machine Learning to a wide variety of tasks. Recently, two families of self-supervised methods, contrastive…

机器学习 · 计算机科学 2021-12-07 Kuang-Huei Lee , Anurag Arnab , Sergio Guadarrama , John Canny , Ian Fischer

The incorporation of LiDAR technology into some high-end smartphones has unlocked numerous possibilities across various applications, including photography, image restoration, augmented reality, and more. In this paper, we introduce a novel…

图像与视频处理 · 电气工程与系统科学 2024-06-28 Alessandro Gnutti , Stefano Della Fiore , Mattia Savardi , Yi-Hsin Chen , Riccardo Leonardi , Wen-Hsiao Peng

Recently, there has been much interest in deep learning techniques to do image compression and there have been claims that several of these produce better results than engineered compression schemes (such as JPEG, JPEG2000 or BPG). A…

图像与视频处理 · 电气工程与系统科学 2019-08-13 Yash Patel , Srikar Appalaraju , R. Manmatha