English
Related papers

Related papers: Discrete Layered Entropy, Conditional Compression …

200 papers

One of the most influential results in neural network theory is the universal approximation theorem [1, 2, 3] which states that continuous functions can be approximated to within arbitrary accuracy by single-hidden-layer feedforward neural…

Machine Learning · Computer Science 2021-12-16 Clemens Hutter , Recep Gül , Helmut Bölcskei

We consider the problem of constructing prefix-free codes in which a designated symbol, a space, can only appear at the end of codewords. We provide a linear-time algorithm to construct almost-optimal codes with this property, meaning that…

Information Theory · Computer Science 2024-05-13 Roberto Bruno , Ugo Vaccaro

The minimum error entropy (MEE) criterion has been successfully used in fields such as parameter estimation, system identification and the supervised machine learning. There is in general no explicit expression for the optimal MEE estimate…

Information Theory · Computer Science 2015-04-14 Badong Chen , Guangmin Wang , Nanning Zheng , Jose C. Principe

In [1] it is shown that recurrent neural networks (RNNs) can learn - in a metric entropy optimal manner - discrete time, linear time-invariant (LTI) systems. This is effected by comparing the number of bits needed to encode the…

Dynamical Systems · Mathematics 2022-11-29 Clemens Hutter , Thomas Allard , Helmut Bölcskei

We study how much a linear program (LP) can be compressed when solved repeatedly, given prior knowledge about its objective function. Existing data-driven projection methods learn low-dimensional surrogate LPs with approximate…

Optimization and Control · Mathematics 2026-05-26 Yuhan Ye , Omar Bennouna

In this paper we will analyze discrete probability distributions in which probabilities of particular outcomes of some experiment (microstates) can be represented by the ratio of natural numbers (in other words, probabilities are…

Information Theory · Computer Science 2009-09-29 Marko V. Jankovic

We show that for log-concave real random variables with fixed variance the Shannon differential entropy is minimized for an exponential random variable. We apply this result to derive upper bounds on capacities of additive noise channels…

Probability · Mathematics 2024-03-19 James Melbourne , Piotr Nayar , Cyril Roberto

Suppose that you have $n$ colours and $m$ mutually independent dice, each of which has $r$ sides. Each dice lands on any of its sides with equal probability. You may colour the sides of each die in any way you wish, but there is one…

Probability · Mathematics 2018-08-14 Christos Pelekis

We present a detailed derivation of some estimators of Shannon entropy for discrete distributions. They hold for finite samples of N points distributed into M "boxes", with N and M -> oo, but N/M < oo. In the high sampling regime (<< 1…

Data Analysis, Statistics and Probability · Physics 2011-11-09 P. Grassberger

Learning, prediction, and compression are intimately connected: a model that accurately predicts the next symbol in a sequence can be coupled with a source coder to compress that sequence near its information-theoretic limit. When tokenized…

Information Theory · Computer Science 2026-05-05 Vishnu Teja Kunde , Jean-Francois Chamberland , Krishna R. Narayanan , Jamison Ebert

We consider the problem of approximating the empirical Shannon entropy of a high-frequency data stream under the relaxed strict-turnstile model, when space limitations make exact computation infeasible. An equivalent measure of entropy is…

Computation · Statistics 2013-04-18 Peter Clifford , Ioana Ada Cosma

The rapid scaling of artificial intelligence models has revealed a fundamental tension between model capacity (storage) and inference efficiency (computation). While classical information theory focuses on transmission and storage limits,…

Information Theory · Computer Science 2026-01-01 Jianfeng Xu , Zeyan Li

Shannon entropy in position ($S_{\rvec}$) and momentum ($S_{\pvec}$) spaces, along with their sum ($S_t$) are presented for unit-normalized densities of He, Li$^+$ and Be$^{2+}$ ions, spatially confined at the center of an impenetrable…

Quantum Physics · Physics 2021-03-01 Sangita Majumdar , Amlan K. Roy

We study the maximum achievable differential entropy at the output of a system assigning to each input X the sum X+N, with N a given noise with probability law absolutely continuous with respect to the Lebesgue measure and where the input…

Optimization and Control · Mathematics 2016-02-04 Francisco J. Piera

Semisupervised text classification has become a major focus of research over the past few years. Hitherto, most of the research has been based on supervised learning, but its main drawback is the unavailability of labeled data samples in…

Machine Learning · Computer Science 2021-11-17 Shivani Malhotra , Vinay Kumar , Alpana Agarwal

We have presented a new axiomatic derivation of Shannon Entropy for a discrete probability distribution on the basis of the postulates of additivity and concavity of the entropy function.We have then modified shannon entropy to take account…

Quantum Physics · Physics 2007-05-23 C. G. Chakrabarti , Indranil Chakrabarty

We propose an end-to-end learned image compression codec wherein the analysis transform is jointly trained with an object classification task. This study affirms that the compressed latent representation can predict human perceptual…

Computer Vision and Pattern Recognition · Computer Science 2024-01-17 Chen-Hsiu Huang , Ja-Ling Wu

Neural networks achieve remarkable performance through superposition: encoding multiple features as overlapping directions in activation space rather than dedicating individual neurons to each feature. This challenges interpretability, yet…

Machine Learning · Computer Science 2025-12-16 Leonard Bereska , Zoe Tzifa-Kratira , Reza Samavi , Efstratios Gavves

We present a technique for entropy optimization to calculate a distribution from its moments. The technique is based upon maximizing a discretized form of the Shannon entropy functional by mapping the problem onto a dual space where an…

Disordered Systems and Neural Networks · Physics 2009-11-10 K. Bandyopadhyay , A. K. Bhattacharya , Parthapratim Biswas , D. A. Drabold

We conclude a sequence of work by giving near-optimal sketching and streaming algorithms for estimating Shannon entropy in the most general streaming model, with arbitrary insertions and deletions. This improves on prior results that obtain…

Data Structures and Algorithms · Computer Science 2008-12-18 Nicholas J. A. Harvey , Jelani Nelson , Krzysztof Onak