English
Related papers

Related papers: On Self-Adaptive Perception Loss Function for Sequ…

200 papers

The mean squared error (MSE) is a ubiquitous loss function for speech enhancement, but its problem is that the error cannot reflect the auditory perception quality. This is because MSE causes models to over-emphasize low-frequency…

Sound · Computer Science 2025-11-11 Zixuan Li , Xueliang Zhang , Changjiang Zhao , Shuai Gao , Lei Miao , Zhipeng Yan , Ying Sun , Chong Zhu

Single-image super-resolution (SISR) networks trained with perceptual and adversarial losses provide high-contrast outputs compared to those of networks trained with distortion-oriented losses, such as L1 or L2. However, it has been shown…

Computer Vision and Pattern Recognition · Computer Science 2023-03-14 Seung Ho Park , Young Su Moon , Nam Ik Cho

Lossy Image compression is necessary for efficient storage and transfer of data. Typically the trade-off between bit-rate and quality determines the optimal compression level. This makes the image quality metric an integral part of any…

Computer Vision and Pattern Recognition · Computer Science 2021-07-16 Juan Carlos Mier , Eddie Huang , Hossein Talebi , Feng Yang , Peyman Milanfar

Recent advances in machine learning-aided lossy compression are incorporating perceptual fidelity into the rate-distortion theory. In this paper, we study the rate-distortion-perception trade-off when the perceptual quality is measured by…

Information Theory · Computer Science 2023-05-23 Xueyan Niu , Deniz Gündüz , Bo Bai , Wei Han

We study lossy source coding under a distortion measure defined by the negative log-likelihood induced by a prescribed conditional distribution $P_{X|U}$. This \emph{log-likelihood distortion} models compression settings in which the…

Information Theory · Computer Science 2026-01-26 Anuj Kumar Yadav , Dan Song , Yanina Shkel , Ayfer Özgür

In the field of neural data compression, the prevailing focus has been on optimizing algorithms for either classical distortion metrics, such as PSNR or SSIM, or human perceptual quality. With increasing amounts of data consumed by machines…

Image and Video Processing · Electrical Eng. & Systems 2024-01-17 Dan Jacobellis , Daniel Cummings , Neeraja J. Yadwadkar

The impressive growth of data throughput in optical microscopy has triggered a widespread use of supervised learning (SL) models running on compressed image datasets for efficient automated analysis. However, since lossy image compression…

In this paper, we study the computation of the rate-distortion-perception function (RDPF) for discrete memoryless sources subject to a single-letter average distortion constraint and a perception constraint that belongs to the family of…

Information Theory · Computer Science 2023-05-09 Giuseppe Serra , Photios A. Stavrou , Marios Kountouris

The distortion-rate function of output-constrained lossy source coding with limited common randomness is analyzed for the special case of squared error distortion measure. An explicit expression is obtained when both source and…

Information Theory · Computer Science 2024-03-25 Li Xie , Liangyan Li , Jun Chen , Zhongshan Zhang

This paper considers the problem of lossy compression for the computation of a function of two correlated sources, both of which are observed at the encoder. Due to presence of observation costs, the encoder is allowed to observe only…

Information Theory · Computer Science 2013-07-22 Xi Liu , Osvaldo Simeone , Elza Erkip

Nowadays, deep-learning image coding solutions have shown similar or better compression efficiency than conventional solutions based on hand-crafted transforms and spatial prediction techniques. These deep-learning codecs require a large…

Multimedia · Computer Science 2023-11-13 Shima Mohammadi , Joao Ascenso

In the classical source coding problem, the compressed source is reconstructed at the decoder with respect to some distortion metric. Motivated by settings in which we are interested in more than simply reconstructing the compressed source,…

Information Theory · Computer Science 2023-10-03 Oğuzhan Kubilay Ülger , Elza Erkip

This paper introduces a sparse projection matrix composed of discrete (digital) periodic lines that create a pseudo-random (p.frac) sampling scheme. Our approach enables random Cartesian sampling whilst employing deterministic and…

Image and Video Processing · Electrical Eng. & Systems 2024-01-09 Marlon Bran Lorenzana , Benjamin Cottier , Matthew Marques , Andrew Kingston , Shekhar S. Chandra

New bounds on the rate distortion function of certain non-Gaussian sources, with a proportional-weighted mean-square error (MSE) distortion measure, are given. The growth, g, of the rate distortion function, as a result of changing from a…

Information Theory · Computer Science 2007-07-13 Jacob Binia

Self-supervised representation learning (SSL) has attained SOTA results on several downstream speech tasks, but SSL-based speech enhancement (SE) solutions still lag behind. To address this issue, we exploit three main ideas: (i)…

In this work, we propose a novel consistency-preserving loss function for recovering the phase information in the context of phase reconstruction (PR) and speech enhancement (SE). Different from conventional techniques that directly…

Audio and Speech Processing · Electrical Eng. & Systems 2024-09-25 Pin-Jui Ku , Chun-Wei Ho , Hao Yen , Sabato Marco Siniscalchi , Chin-Hui Lee

The rate-distortion function (RDF) has long been an information-theoretic benchmark for data compression. As its natural extension, the indirect rate-distortion function (iRDF) corresponds to the scenario where the encoder can only access…

Information Theory · Computer Science 2025-03-11 Zichao Yu , Qiang Sun , Wenyi Zhang

We examine the problem of selecting a small set of linear measurements for reconstructing high-dimensional signals. Well-established methods for optimizing such measurements include principal component analysis (PCA), independent component…

Computer Vision and Pattern Recognition · Computer Science 2025-08-06 Ling-Qi Zhang , Zahra Kadkhodaie , Eero P. Simoncelli , David H. Brainard

This paper proposes a probabilistic contrastive loss function for self-supervised learning. The well-known contrastive loss is deterministic and involves a temperature hyperparameter that scales the inner product between two normed feature…

Machine Learning · Computer Science 2021-12-06 Shen Li , Jianqing Xu , Bryan Hooi

We investigate lossy compression (source coding) of data in the form of permutations. This problem has direct applications in the storage of ordinal data or rankings, and in the analysis of sorting algorithms. We analyze the rate-distortion…

Information Theory · Computer Science 2016-11-18 Da Wang , Arya Mazumdar , Gregory Wornell