中文
相关论文

相关论文: Rate-Distortion Optimization With Alternative Refe…

200 篇论文

Depth image based rendering techniques for multiview applications have been recently introduced for efficient view generation at arbitrary camera positions. Encoding rate control has thus to consider both texture and depth data. Due to…

计算机视觉与模式识别 · 计算机科学 2012-11-20 Boshra Rajaei , Thomas Maugey , Hamid-Reza Pourreza , Pascal Frossard

This paper addresses rate control for transmission of scalable video streams via Network Utility Maximization (NUM) formulation. Due to stringent QoS requirements of video streams and specific characterization of utility experienced by…

Video compression is a critical component of Internet video delivery. Recent work has shown that deep learning techniques can rival or outperform human-designed algorithms, but these methods are significantly less compute and…

计算机视觉与模式识别 · 计算机科学 2021-04-07 Mehrdad Khani , Vibhaalakshmi Sivaraman , Mohammad Alizadeh

Neural networks (NN) can improve standard video compression by pre- and post-processing the encoded video. For optimal NN training, the standard codec needs to be replaced with a codec proxy that can provide derivatives of estimated…

图像与视频处理 · 电气工程与系统科学 2023-01-25 Amir Said , Manish Kumar Singh , Reza Pourreza

We consider the semantic rate-distortion problem motivated by task-oriented video compression. The semantic information corresponding to the task, which is not observable to the encoder, shows impacts on the observations through a joint…

信息论 · 计算机科学 2022-08-15 Tao Guo , Yizhu Wang , Jie Han , Huihui Wu , Bo Bai , Wei Han

Recently, the field of Image Coding for Machines (ICM) has garnered heightened interest and significant advances thanks to the rapid progress of learning-based techniques for image compression and analysis. Previous studies often require…

计算机视觉与模式识别 · 计算机科学 2024-07-18 Jinming Liu , Ruoyu Feng , Yunpeng Qi , Qiuyu Chen , Zhibo Chen , Wenjun Zeng , Xin Jin

Most neural compression models are trained on large datasets of images or videos in order to generalize to unseen data. Such generalization typically requires large and expressive architectures with a high decoding complexity. Here we…

图像与视频处理 · 电气工程与系统科学 2023-12-06 Hyunjik Kim , Matthias Bauer , Lucas Theis , Jonathan Richard Schwarz , Emilien Dupont

3D Gaussian Splatting (3DGS) has become an emerging technique with remarkable potential in 3D representation and image rendering. However, the substantial storage overhead of 3DGS significantly impedes its practical applications. In this…

计算机视觉与模式识别 · 计算机科学 2024-10-22 Henan Wang , Hanxin Zhu , Tianyu He , Runsen Feng , Jiajun Deng , Jiang Bian , Zhibo Chen

This paper studies the rate-distortion-perception (RDP) tradeoff for a Gaussian vector source coding problem where the goal is to compress the multi-component source subject to distortion and perception constraints. Specifically, the RDP…

信息论 · 计算机科学 2025-03-18 Jingjing Qian , Sadaf Salehkalaibar , Jun Chen , Ashish Khisti , Wei Yu , Wuxian Shi , Yiqun Ge , Wen Tong

This paper focuses on the task of quality enhancement for compressed videos. Although deep network-based video restorers achieve impressive progress, most of the existing methods lack a structured design to optimally leverage the priors…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Hanchi Sun , Xiaohong Liu , Xinyang Jiang , Yifei Shen , Dongsheng Li , Xiongkuo Min , Guangtao Zhai

We frame the problem of selecting an optimal audio encoding scheme as a supervised learning task. Through uniform convergence theory, we guarantee approximately optimal codec selection while controlling for selection bias. We present…

声音 · 计算机科学 2018-12-20 Clayton Sanford , Cyrus Cousins , Eli Upfal

Conventional communication systems, including both separation-based coding and AI-driven joint source-channel coding (JSCC), are largely guided by Shannon's rate-distortion theory. However, relying on generic distortion metrics fails to…

信息论 · 计算机科学 2026-01-21 Tong Wu , Zhiyong Chen , Guo Lu , Li Song , Feng Yang , Meixia Tao , Wenjun Zhang

The joint source-channel coding (JSCC) framework leverages deep learning to learn from data the best codes for source and channel coding. When the output signal, rather than being binary, is directly mapped onto the IQ domain…

机器学习 · 计算机科学 2024-06-07 Junli Fang , João F. C. Mota , Baoshan Lu , Weicheng Zhang , Xuemin Hong

Recent years have witnessed the dramatic growth of Internet video traffic, where the video bitstreams are often compressed and delivered in low quality to fit the streamer's uplink bandwidth. To alleviate the quality degradation, it comes…

图像与视频处理 · 电气工程与系统科学 2023-03-09 Qihua Zhou , Ruibin Li , Song Guo , Peiran Dong , Yi Liu , Jingcai Guo , Zhenda Xu

We revisit the Gray-Wyner lossy source coding problem and derive the first-order asymptotic optimal rate-distortion-perception region when additional perception constraints are imposed on reproduced source sequences. The optimal trade-off…

信息论 · 计算机科学 2026-01-19 Yu Yang , Yingxin Zhang , Weijie Yuan , Lin Zhou

We consider a multiterminal source coding problem in which a source is estimated at a central processing unit from lossy-compressed remote observations. Each lossy-encoded observation is produced by a remote sensor which obtains a noisy…

信息论 · 计算机科学 2016-05-13 Ruiyang Song , Stefano Rini , Alon Kipnis , Andrea J. Goldsmith

Learned image compression methods generally optimize a rate-distortion loss, trading off improvements in visual distortion for added bitrate. Increasingly, however, compressed imagery is used as an input to deep learning networks for…

图像与视频处理 · 电气工程与系统科学 2022-02-02 Maxime Kawawa-Beaudan , Ryan Roggenkemper , Avideh Zakhor

We propose a rate-distortion optimization method for 3D videos based on visual discomfort estimation. We calculate visual discomfort in the encoded depth maps using two indexes: temporal outliers (TO) and spatial outliers (SO). These two…

图像与视频处理 · 电气工程与系统科学 2018-11-22 Dogancan Temel , Ghassan AlRegib

Optical camera communication (OCC) has emerged as a key enabling technology for the seamless operation of future autonomous vehicles. By leveraging the supreme performance of OCC, we can meet the stringent requirements of ultra-reliable and…

网络与互联网体系结构 · 计算机科学 2022-05-16 Amirul Islam , Leila Musavian , Nikolaos Thomos

Efficient 3D LiDAR point cloud compression (LPCC) and streaming are critical for edge server-assisted robotic systems, enabling real-time communication with compact data representations. A widely adopted approach represents LiDAR point…

图像与视频处理 · 电气工程与系统科学 2026-03-17 Shengqian Wang , Chang Tu , He Chen