中文
相关论文

相关论文: Modeling Generalized Rate-Distortion Functions

200 篇论文

Evaluating the realism of generated images remains a fundamental challenge in generative modeling. Existing distributional metrics such as the Frechet Inception Distance (FID) and CLIP-MMD (CMMD) compare feature distributions at a semantic…

计算机视觉与模式识别 · 计算机科学 2026-04-06 Joé Napolitano , Pascal Nguyen

We present the first unified framework for rate-distortion-optimized compression and segmentation of 3D Gaussian Splatting (3DGS). While 3DGS has proven effective for both real-time rendering and semantic scene understanding, prior works…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Yu-Jen Tseng , Chia-Hao Kao , Jing-Zhong Chen , Alessandro Gnutti , Shao-Yuan Lo , Yen-Yu Lin , Wen-Hsiao Peng

In rate-distortion optimization, the encoder settings are determined by maximizing a reconstruction quality measure subject to a constraint on the bit rate. One of the main challenges of this approach is to define a quality measure that can…

图像与视频处理 · 电气工程与系统科学 2021-08-11 Qi Liu , Hui Yuan , Raouf Hamzaoui , Honglei Su , Junhui Hou , Huan Yang

Iterative algorithms have many advantages for linear tomographic image reconstruction when compared to back-projection based methods. However, iterative methods tend to have significantly higher computational complexity. To overcome this,…

分布式、并行与集群计算 · 计算机科学 2019-03-29 Yushan Gao , Ander Biguri , Thomas Blumensath

High frame rate and accurate depth estimation plays an important role in several tasks crucial to robotics and automotive perception. To date, this can be achieved through ToF and LiDAR devices for indoor and outdoor applications,…

计算机视觉与模式识别 · 计算机科学 2024-09-13 Andrea Conti , Matteo Poggi , Valerio Cambareri , Stefano Mattoccia

Event cameras offer a high temporal resolution over traditional frame-based cameras, which makes them suitable for motion and structure estimation. However, it has been unclear how event-based 3D Gaussian Splatting (3DGS) approaches could…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Kai Kohyama , Yoshimitsu Aoki , Guillermo Gallego , Shintaro Shiba

We have deluge of data in time series format for numerous phenomena. The number of snapshots, resolution and many other factors come into play as we look to identify the dynamics in a given problem. The pre-processing and post-processing…

信号处理 · 电气工程与系统科学 2020-01-13 Mohammad N. Murshed , M. Monir Uddin

There has been a growing trend in compressing and transmitting videos from terminals for machine vision tasks. Nevertheless, most video coding optimization method focus on minimizing distortion according to human perceptual metrics,…

多媒体 · 计算机科学 2025-12-18 Fei Zhao , Mengxi Guo , Shijie Zhao , Junlin Li , Li Zhang , Xiaodong Xie

We consider the problem of recovering elements of a low-dimensional model from linear measurements. From signal and image processing to inverse problems in data science, this question has been at the center of many applications. Lately,…

信号处理 · 电气工程与系统科学 2025-05-15 Yann Traonmilin , Jean François Aujol , Antoine Guennec

In this work, we propose a no-reference video quality assessment method, aiming to achieve high-generalization capability in cross-content, -resolution and -frame rate quality prediction. In particular, we evaluate the quality of a video by…

图像与视频处理 · 电气工程与系统科学 2021-06-24 Baoliang Chen , Lingyu Zhu , Guo Li , Hongfei Fan , Shiqi Wang

Machine learning models struggle with generalization when encountering out-of-distribution (OOD) samples with unexpected distribution shifts. For vision tasks, recent studies have shown that test-time adaptation employing diffusion models…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Yun-Yun Tsai , Fu-Chen Chen , Albert Y. C. Chen , Junfeng Yang , Che-Chun Su , Min Sun , Cheng-Hao Kuo

Efficient and high-fidelity prior sampling and inversion for complex geological media is still a largely unsolved challenge. Here, we use a deep neural network of the variational autoencoder type to construct a parametric low-dimensional…

机器学习 · 统计学 2017-10-26 Eric Laloy , Romain Hérault , John Lee , Diederik Jacques , Niklas Linde

Recent advances in Rate-Distortion-Perception (RDP) theory highlight the importance of balancing compression level, reconstruction quality, and perceptual fidelity. While previous work has explored numerical approaches to approximate the…

信息论 · 计算机科学 2025-08-20 Chunhui Chen , Linyi Chen , Xueyan Niu , Hao Wu

The network transport of 3D video, which contains two views of a video scene, poses significant challenges due to the increased video data compared to conventional single-view video. Addressing these challenges requires a thorough…

多媒体 · 计算机科学 2013-11-25 Akshay Pulipaka , Patrick Seeling , Martin Reisslein , Lina J. Karam

In this paper, we introduce GoodDrag, a novel approach to improve the stability and image quality of drag editing. Unlike existing methods that struggle with accumulated perturbations and often result in distortions, GoodDrag introduces an…

计算机视觉与模式识别 · 计算机科学 2024-04-11 Zewei Zhang , Huan Liu , Jun Chen , Xiangyu Xu

In this paper, we propose a new full-reference quality metric for mobile 3D content. Our method is modeled around the Human Visual System, fusing the information of both left and right channels, considering color components, the cyclopean…

图像与视频处理 · 电气工程与系统科学 2018-03-19 Amin Banitalebi-Dehkordi , Mahsa T. Pourazad , Panos Nasiopoulos

In rate-distortion (RD) problems one seeks reduced representations of a source that meet a target distortion constraint. Such optimal representations undergo topological transitions at some critical rate values, when their cardinality or…

信息论 · 计算机科学 2023-10-09 Shlomi Agmon , Etam Benger , Or Ordentlich , Naftali Tishby

A deep image compression scheme is proposed in this paper, offering the state-of-the-art compression efficiency, against the traditional JPEG, JPEG2000, BPG and those popular learning based methodologies. This is achieved by a novel…

图像与视频处理 · 电气工程与系统科学 2019-02-28 Haojie Liu , Tong Chen , Peiyao Guo , Qiu Shen , Zhan Ma

Dynamic garment reconstruction from monocular video is an important yet challenging task due to the complex dynamics and unconstrained nature of the garments. Recent advancements in neural rendering have enabled high-quality geometric…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Soham Dasgupta , Shanthika Naik , Preet Savalia , Sujay Kumar Ingle , Avinash Sharma

High Dynamic Range (HDR) videos have enjoyed a surge in popularity in recent years due to their ability to represent a wider range of contrast and color than Standard Dynamic Range (SDR) videos. Although HDR video capture has seen…

图像与视频处理 · 电气工程与系统科学 2024-04-23 Abhinau K. Venkataramanan , Cosmin Stejerean , Ioannis Katsavounidis , Hassene Tmar , Alan C. Bovik