English
Related papers

Related papers: Generalized Decomposition of Multivariate Informat…

200 papers

The information that two random variables $Y$, $Z$ contain about a third random variable $X$ can have aspects of shared information (contained in both $Y$ and $Z$), of complementary information (only available from $(Y,Z)$ together) and of…

Information Theory · Computer Science 2015-03-05 Johannes Rauh , Nils Bertschinger , Eckehard Olbrich , Jürgen Jost

Partial information decomposition has recently found applications in biological signal processing and machine learning. Despite its impacts, the decomposition was introduced through an informal and heuristic route, and its exact operational…

Information Theory · Computer Science 2025-02-18 Chao Tian , Shlomo Shamai

The Information Bottleneck (IB) principle offers a compelling theoretical framework to understand how neural networks (NNs) learn. However, its practical utility has been constrained by unresolved theoretical ambiguities and significant…

Machine Learning · Computer Science 2026-02-02 Charles Westphal , Stephen Hailes , Mirco Musolesi

The study of irreducible higher-order interactions has become a core topic of study in complex systems. Two of the most well-developed frameworks, topological data analysis and multivariate information theory, aim to provide formal tools…

Information Theory · Computer Science 2025-04-15 Thomas F. Varley , Pedro A. M. Mediano , Alice Patania , Josh Bongard

In this paper, we delve deeper into the Kullback-Leibler (KL) Divergence loss and mathematically prove that it is equivalent to the Decoupled Kullback-Leibler (DKL) Divergence loss that consists of (1) a weighted Mean Square Error (wMSE)…

Machine Learning · Computer Science 2025-03-12 Jiequan Cui , Beier Zhu , Qingshan Xu , Zhuotao Tian , Xiaojuan Qi , Bei Yu , Hanwang Zhang , Richang Hong

This paper is devoted to the mathematical study of some divergences based on the mutual information well-suited to categorical random vectors. These divergences are generalizations of the "entropy distance" and "information distance". Their…

Statistics Theory · Mathematics 2016-08-16 Jean-François Coeurjolly , Rémy Drouilhet , Jean-François Robineau

We study a separable design for computing information measures, where the information measure is computed from learned feature representations instead of raw data. Under mild assumptions on the feature representations, we demonstrate that a…

Information Theory · Computer Science 2025-01-28 Xiangxiang Xu , Lizhong Zheng

We propose a new model selection method, the posterior averaging information criterion, for Bayesian model assessment from a predictive perspective. The theoretical foundation is built on the Kullback-Leibler divergence to quantify the…

Methodology · Statistics 2020-09-22 Shouhao Zhou

We examine the relationship between the mutual information between the output model and the empirical sample and the generalization of the algorithm in the context of stochastic convex optimization. Despite increasing interest in…

Machine Learning · Computer Science 2024-01-17 Roi Livni

How much one has learned from an experiment is quantifiable by the information gain, also known as the Kullback-Leibler divergence. The narrowing of the posterior parameter distribution $P(\theta|D)$ compared with the prior parameter…

Statistical Mechanics · Physics 2022-08-29 Johannes Buchner

We introduce a novel framework for decomposing interventional causal effects into synergistic, redundant, and unique components, building on the intuition of Partial Information Decomposition (PID) and the principle of M\"obius inversion.…

Artificial Intelligence · Computer Science 2025-09-22 Abel Jansma

Mutual information $I(X;Y)$ is a useful definition in information theory to estimate how much information the random variable $Y$ holds about the random variable $X$. One way to define the mutual information is by comparing the joint…

Information Theory · Computer Science 2022-04-14 Bulut Kuskonmaz , Jaron Skovsted Gundersen , Rafal Wisniewski

Information distance is a parameter-free similarity measure based on compression, used in pattern recognition, data mining, phylogeny, clustering, and classification. The notion of information distance is extended from pairs to multiples…

Computer Vision and Pattern Recognition · Computer Science 2009-05-21 Paul M. B. Vitanyi

A range of nonlinear image reconstruction procedures based on extremizing the generalized Shannon entropy, Kullback-Leibler cross-entropy and Renyi information measures and proposed by the author in early papers is presented. The…

Astrophysics · Physics 2007-05-23 Anisa T. Bajkova

The aim of this work is to provide bounds connecting two probability measures of the same event using R\'enyi $\alpha$-Divergences and Sibson's $\alpha$-Mutual Information, a generalization of respectively the Kullback-Leibler Divergence…

Information Theory · Computer Science 2020-01-20 Amedeo Roberto Esposito , Michael Gastpar , Ibrahim Issa

In the context of statistical learning, the Information Bottleneck method seeks a right balance between accuracy and generalization capability through a suitable tradeoff between compression complexity, measured by minimum description…

Information Theory · Computer Science 2021-02-16 Mohammad Mahdi Mahvari , Mari Kobayashi , Abdellatif Zaidi

The introduction of the partial information decomposition generated a flurry of proposals for defining an intersection information that quantifies how much of "the same information" two or more random variables specify about a target random…

Information Theory · Computer Science 2015-06-11 Virgil Griffith , Edwin K. P. Chong , Ryan G. James , Christopher J. Ellison , James P. Crutchfield

The concept of Generalized Inverse based Decoding (GID) is introduced, as an algebraic framework for the syndrome decoding problem (SDP) and low weight codeword problem (LWP). The framework has ground on two characterizations by generalized…

Information Theory · Computer Science 2022-02-18 Ferucio Laurentiu Tiplea , Vlad-Florin Dragoi

Of the various attempts to generalize information theory to multiple variables, the most widely utilized, interaction information, suffers from the problem that it is sometimes negative. Here we reconsider from first principles the general…

Information Theory · Computer Science 2010-04-16 Paul L. Williams , Randall D. Beer

Graph neural networks are widely used for node classification, but they remain vulnerable to out-of-distribution (OOD) shifts in node features and graph structure. Prior work established that methods trained with standard supervised…

Machine Learning · Computer Science 2026-05-15 Danny Wang , Ruihong Qiu , Zi Huang
‹ Prev 1 4 5 6 7 8 10 Next ›