English
Related papers

Related papers: Dimension-free Information Concentration via Exp-C…

200 papers

A thermodynamic formalism describing the efficiency of information learning is proposed, which is applicable for stochastic thermodynamic systems with multiple internal degree of freedom. The learning rate, entropy production rate (EPR),…

Statistical Mechanics · Physics 2023-05-31 Minghao Li , Shihao Xia , Youlin Wang , Minglong Lv , Shanhe Su

We define {\em predictive information} $I_{\rm pred} (T)$ as the mutual information between the past and the future of a time series. Three qualitatively different behaviors are found in the limit of large observation times $T$: $I_{\rm…

Data Analysis, Statistics and Probability · Physics 2011-11-10 William Bialek , Ilya Nemenman , Naftali Tishby

We study the continuity property of the generalized entropy as a function of the underlying probability distribution, defined with an action space and a loss function, and use this property to answer the basic questions in statistical…

Machine Learning · Computer Science 2022-01-04 Aolin Xu

An important theme in recent work in asymptotic geometric analysis is that many classical implications between different types of geometric or functional inequalities can be reversed in the presence of convexity assumptions. In this note,…

Probability · Mathematics 2015-07-22 Elizabeth S. Meckes , Mark W. Meckes

We investigate the asymptotic behavior of probability measures associated with stochastic dynamical systems featuring either globally contracting or $B_{r}$-contracting drift terms. While classical results often assume constant diffusion…

Dynamical Systems · Mathematics 2025-05-14 Simone Betteti , Francesco Bullo

We study the velocity of the propagation of information for a class of local dissipative quantum dynamics. This finite velocity is expressed by the so-called Lieb-Robinson bound. Besides the properties of the already studied dynamics, we…

Mathematical Physics · Physics 2013-09-17 Benoît Descamps

Multi-instance data, in which each object (bag) contains a collection of instances, are widespread in machine learning, computer vision, bioinformatics, signal processing, and social sciences. We present a maximum entropy (ME) framework for…

Machine Learning · Computer Science 2016-03-15 Behrouz Behmardi , Forrest Briggs , Xiaoli Z. Fern , Raviv Raich

Consider a random sample $X_1 , X_2 , ..., X_n$ drawn independently and identically distributed from some known sampling distribution $P_X$. Let $X_{(1)} \le X_{(2)} \le ... \le X_{(n)}$ represent the order statistics of the sample. The…

Information Theory · Computer Science 2020-09-28 Alex Dytso , Martina Cardone , Cynthia Rush

Strongly log-concave (SLC) distributions are a rich class of discrete probability distributions over subsets of some ground set. They are strictly more general than strongly Rayleigh (SR) distributions such as the well-known determinantal…

Machine Learning · Computer Science 2019-06-14 Joshua Robinson , Suvrit Sra , Stefanie Jegelka

Large Language Models (LLMs) are known to memorize portions of their training data, sometimes even reproduce content verbatim when prompted appropriately. Despite substantial interest, existing LLM memorization research has offered limited…

Computation and Language · Computer Science 2026-04-21 Yizhan Huang , Zhe Yang , Meifang Chen , Huang Nianchen , Jianping Zhang , Michael R. Lyu

The Levy-type distributions are derived using the principle of maximum Tsallis nonextensive entropy both in the full and half spaces. The rates of convergence to the exact Levy stable distributions are determined by taking the N-fold…

Statistical Mechanics · Physics 2009-10-31 Sumiyoshi Abe , A. K. Rajagopal

Stimulated by the need of describing useful notions related to information measures, we introduce the `pdf-related distributions'. These are defined in terms of transformation of absolutely continuous random variables through their own…

Probability · Mathematics 2024-05-02 Antonio Di Crescenzo , Luca Paolillo , Alfonso Suarez-Llorens

Learning systems acquire structured internal representations from data, yet classical information-theoretic results state that deterministic transformations do not increase information. This raises a fundamental question: how can learning…

Machine Learning · Computer Science 2026-01-29 Daisuke Okanohara

Learned image compression methods have attracted great research interest and exhibited superior rate-distortion performance to the best classical image compression standards of the present. The entropy model plays a key role in learned…

Computer Vision and Pattern Recognition · Computer Science 2025-05-16 Jingbo Lu , Leheng Zhang , Xingyu Zhou , Mu Li , Wen Li , Shuhang Gu

This presentation's Part 3 studies the evolutionary information processes and regularities of evolution dynamics, evaluated by an entropy functional (EF) of a random field (modeled by a diffusion information process) and an informational…

Adaptation and Self-Organizing Systems · Physics 2012-08-17 Vladimir S. Lerner

Is reduction always a good scientific strategy? Does it always lead to a gain in information? The very existence of the special sciences above and beyond physics seems to hint no. Previous research has shown that dimension reduction…

Information Theory · Computer Science 2021-04-28 Thomas Varley , Erik Hoel

Dataset Condensation (DC) seeks to select or distill samples from large datasets into smaller subsets while preserving performance on target tasks. Existing methods primarily focus on pruning or synthesizing data in the same format as the…

Computer Vision and Pattern Recognition · Computer Science 2026-03-11 Shaobo Wang , Youxin Jiang , Tianle Niu , Yantai Yang , Ruiji Zhang , Shuhao Hu , Shuaiyu Zhang , Chenghao Sun , Weiya Li , Conghui He , Xuming Hu , Linfeng Zhang

Recent work has shown that tight concentration of the entire spectrum of singular values of a deep network's input-output Jacobian around one at initialization can speed up learning by orders of magnitude. Therefore, to guide important…

Machine Learning · Statistics 2018-02-28 Jeffrey Pennington , Samuel S. Schoenholz , Surya Ganguli

To learn (statistical) dependencies among random variables requires exponentially large sample size in the number of observed random variables if any arbitrary joint probability distribution can occur. We consider the case that sparse data…

Machine Learning · Computer Science 2007-05-23 Dominik Janzing , Daniel Herrmann

This thesis details a class of partial orders on the space of probability distributions and the space of density operators which capture the idea of information content. Some links to domain theory and computational linguistics are also…

Logic in Computer Science · Computer Science 2017-01-25 John van de Wetering
‹ Prev 1 8 9 10 Next ›