English
Related papers

Related papers: On the R\'{e}nyi Cross-Entropy

200 papers

Following a recent proof of Shannon's entropy power inequality (EPI), a comprehensive framework for deriving various EPIs for the R\'enyi entropy is presented that uses transport arguments from normal densities and a change of variable by…

Information Theory · Computer Science 2018-10-17 Olivier Rioul

Tsallis and R\'{e}nyi entropies, which are monotone transformations of each other, are deformations of the celebrated Shannon entropy. Maximization of these deformed entropies, under suitable constraints, leads to the $q$-exponential family…

Probability · Mathematics 2022-01-14 Ting-Kam Leonard Wong , Jun Zhang

Quantum key distribution requires tight and reliable bounds on the secret key rate to ensure robust security. This is particularly so for the regime of finite block sizes, where the optimization of generalized R\'enyi entropic quantities is…

Quantum Physics · Physics 2026-04-09 Rebecca R. B. Chung , Nelly H. Y. Ng , Yu Cai

Machine learning theory has mostly focused on generalization to samples from the same distribution as the training data. Whereas a better understanding of generalization beyond the training distribution where the observed distribution…

Machine Learning · Statistics 2019-05-29 Matias Vera , Pablo Piantanida , Leonardo Rey Vega

In this paper we derive explicit formulas of the R\'enyi information, Shannon entropy and Song measure for the invariant density of one dimensional ergodic diffusion processes. In particular, the diffusion models considered include the…

Probability · Mathematics 2007-11-13 Alessandro De Gregorio , Stefano Iacus

Entropy is a measure of self-information which is used to quantify losses. Entropy was developed in thermodynamics, but is also used to compare probabilities based on their deviating information content. Corresponding model uncertainty is…

Probability · Mathematics 2018-01-23 Alois Pichler , Ruben Schlotter

We consider the two-dimensional (2d) Ising model on a infinitely long cylinder and study the probabilities $p_i$ to observe a given spin configuration $i$ along a circular section of the cylinder. These probabilities also occur as…

Strongly Correlated Electrons · Physics 2010-11-02 Jean-Marie Stéphan , Grégoire Misguich , Vincent Pasquier

Given two networks with the same training loss on a dataset, when would they have drastically different test losses and errors? Better understanding of this question of generalization may improve practical applications of deep networks. In…

Machine Learning · Computer Science 2018-07-26 Qianli Liao , Brando Miranda , Andrzej Banburski , Jack Hidary , Tomaso Poggio

Estimating entropies from limited data series is known to be a non-trivial task. Naive estimations are plagued with both systematic (bias) and statistical errors. Here, we present a new 'balanced estimator' for entropy functionals Shannon,…

Statistical Mechanics · Physics 2008-04-30 Juan A. Bonachela , Haye Hinrichsen , Miguel A. Munoz

In many applications, the probability density function is subject to experimental errors. In this work the continuos dependence of a class of generalized entropies on the experimental errors is studied. This class includes the C. Shannon,…

Data Analysis, Statistics and Probability · Physics 2016-05-20 György Steinbrecher , Giorgio Sonnino

Modern deep learning is primarily an experimental science, in which empirical advances occasionally come at the expense of probabilistic rigor. Here we focus on one such example; namely the use of the categorical cross-entropy loss to model…

Machine Learning · Statistics 2020-11-11 Elliott Gordon-Rodriguez , Gabriel Loaiza-Ganem , Geoff Pleiss , John P. Cunningham

We study convexity properties of R\'{e}nyi entropy as function of $\alpha>0$ on finite alphabets. We also describe robustness of the R\'{e}nyi entropy on finite alphabets, and it turns out that the rate of respective convergence depends on…

Probability · Mathematics 2021-03-09 Filipp Buryak , Yuliya Mishura

The cross entropy loss is widely used due to its effectiveness and solid theoretical grounding. However, as training progresses, the loss tends to focus on hard to classify samples, which may prevent the network from obtaining gains in…

Machine Learning · Computer Science 2021-09-14 Barak Battash , Lior Wolf , Tamir Hazan

In the estimation theory context, we generalize the notion of Shannon's entropy power to the R\'{e}nyi-entropy setting. This not only allows to find new estimation inequalities, such as the R\'{e}nyi-entropy based De Bruijn identity,…

Quantum Physics · Physics 2021-04-07 Petr Jizba , Jacob Dunningham , Martin Prokš

By replacing linear averaging in Shannon entropy with Kolmogorov-Nagumo average (KN-averages) or quasilinear mean and further imposing the additivity constraint, R\'{e}nyi proposed the first formal generalization of Shannon entropy. Using…

Information Theory · Computer Science 2007-07-13 Ambedkar Dukkipati , M. Narasimha Murty , Shalabh Bhatnagar

We present simple and computationally efficient nonparametric estimators of R\'enyi entropy and mutual information based on an i.i.d. sample drawn from an unknown, absolutely continuous distribution over $\R^d$. The estimators are…

Machine Learning · Statistics 2010-10-27 Dávid Pál , Barnabás Póczos , Csaba Szepesvári

We present the Tamed Cross Entropy (TCE) loss function, a robust derivative of the standard Cross Entropy (CE) loss used in deep learning for classification tasks. However, unlike other robust losses, the TCE loss is designed to exhibit the…

Machine Learning · Computer Science 2018-10-12 Manuel Martinez , Rainer Stiefelhagen

A class of estimators of the R\'{e}nyi and Tsallis entropies of an unknown distribution $f$ in $\mathbb{R}^m$ is presented. These estimators are based on the $k$th nearest-neighbor distances computed from a sample of $N$ i.i.d. vectors with…

Statistics Theory · Mathematics 2012-11-16 Nikolai Leonenko , Luc Pronzato , Vippal Savani

Sharpness (of the loss minima) is a common measure to investigate the generalization of neural networks. Intuitively speaking, the flatter the landscape near the minima is, the better generalization might be. Unfortunately, the correlation…

Machine Learning · Computer Science 2025-10-17 Qiaozhe Zhang , Jun Sun , Ruijie Zhang , Yingzhuang Liu

We study the $p$-R\'{e}nyi entropy power inequality with a weight factor $t$ on two independent continuous random variables $X$ and $Y$. The extension essentially relies on a modulation on the sharp Young's inequality due to Bobkov and…

Quantum Physics · Physics 2023-11-28 Junseo Lee , Hyeonjun Yeo , Kabgyun Jeong