English
Related papers

Related papers: Dimension-free Information Concentration via Exp-C…

200 papers

We use a formal correspondence between thermodynamics and inference, where the number of samples can be thought of as the inverse temperature, to study a quantity called ``learning capacity'' which is a measure of the effective…

Machine Learning · Computer Science 2024-10-22 Daiwei Chen , Wei-Kai Chang , Pratik Chaudhari

The field of complex networks studies a wide variety of interacting systems by representing them as networks. To understand their properties and mutual relations, the randomisation of network connections is a commonly used tool. However,…

Statistical Mechanics · Physics 2024-10-18 Noam Abadi , Franco Ruzzenenti

Most work on supervised learning research has focused on marginal predictions. In decision problems, joint predictive distributions are essential for good performance. Previous work has developed methods for assessing low-order predictive…

Machine Learning · Statistics 2022-03-01 Ian Osband , Zheng Wen , Seyed Mohammad Asghari , Vikranth Dwaracherla , Xiuyuan Lu , Benjamin Van Roy

We study the expected volume of random polytopes generated by taking the convex hull of independent identically distributed points from a given distribution. We show that for log-concave distributions supported on convex bodies, we need at…

Metric Geometry · Mathematics 2021-11-16 Debsoumya Chakraborti , Tomasz Tkocz , Beatrice-Helen Vritsiou

Mixture distributions are a workhorse model for multimodal data in information theory, signal processing, and machine learning. Yet even when each component density is simple, the differential entropy of the mixture is notoriously hard to…

Information Theory · Computer Science 2026-02-18 Namyoon Lee

In information theory, the link between continuous information and discrete information is established through well-known sampling theorems. Sampling theory explains, for example, how frequency-filtered music signals are reconstructible…

General Relativity and Quantum Cosmology · Physics 2009-11-10 Achim Kempf

This paper shows the strong converse and the dispersion of memoryless channels with cost constraints and performs refined analysis of the third order term in the asymptotic expansion of the maximum achievable channel coding rate, showing…

Information Theory · Computer Science 2015-10-09 Victoria Kostina , Sergio Verdú

Bayes' theorem incorporates distinct types of information through the likelihood and prior. Direct observations of state variables enter the likelihood and modify posterior probabilities through consistent updating. Information in terms of…

Methodology · Statistics 2024-07-19 Duncan K. Foley , Ellis Scharfenaker

Learning high-dimensional distributions is often done with explicit likelihood modeling or implicit modeling via minimizing integral probability metrics (IPMs). In this paper, we expand this learning paradigm to stochastic orders, namely,…

Machine Learning · Statistics 2022-11-11 Carles Domingo-Enrich , Yair Schiff , Youssef Mroueh

Degrading performance of indexing schemes for exact similarity search in high dimensions has long since been linked to histograms of distributions of distances and other 1-Lipschitz functions getting concentrated. We discuss this…

Data Structures and Algorithms · Computer Science 2012-04-13 Vladimir Pestov

This paper introduces a new constraint domain for reasoning about data with uncertainty. It extends convex modeling with the notion of p-box to gain additional quantifiable information on the data whereabouts. Unlike existing approaches,…

Logic in Computer Science · Computer Science 2014-06-25 Aya Saad , Thom Fruehwirth , Carmen Gervet

A central task in analyzing complex dynamics is to determine the loci of information storage and the communication topology of information flows within a system. Over the last decade and a half, diagnostics for the latter have come to be…

Statistical Mechanics · Physics 2016-06-20 Ryan G. James , Nix Barnett , James P. Crutchfield

In this work, conditional entropy is used to quantify the information loss induced by passing a continuous random variable through a memoryless nonlinear input-output system. We derive an expression for the information loss depending on the…

Information Theory · Computer Science 2012-02-03 Bernhard C. Geiger , Christian Feldbauer , Gernot Kubin

We consider high dimensional Wishart matrices $\mathbb{X} \mathbb{X}^{\top}$ where the entries of $\mathbb{X} \in {\mathbb{R}^{n \times d}}$ are i.i.d. from a log-concave distribution. We prove an information theoretic phase transition:…

Probability · Mathematics 2018-08-14 Sébastien Bubeck , Shirshendu Ganguly

This paper addresses the question of the fluctuations of the empirical entropy of a chain of infinite order. We assume that the chain takes values on a finite alphabet and loses memory exponentially fast. We consider two possible…

Statistical Mechanics · Physics 2007-05-23 D. Gabrielli , A. Galves , D. Guiol

Maximum entropy estimation is of broad interest for inferring properties of systems across many different disciplines. In this work, we significantly extend a technique we previously introduced for estimating the maximum entropy of a set of…

Data Analysis, Statistics and Probability · Physics 2016-01-05 Elliot A. Martin , Jaroslav Hlinka , Alexander Meinke , Filip Děchtěrenko , Jörn Davidsen

Recent years have witnessed the rapid advancements of large language models (LLMs) and their expanding applications, leading to soaring demands for computational resources. The widespread adoption of test-time scaling further intensifies…

Artificial Intelligence · Computer Science 2026-03-11 Cheng Yuan , Jiawei Shao , Xuelong Li

We observe that the distribution of the eigenvalues of an $N$-by-$N$ GUE random matrix is log-concave on $\mathbb{R}^N$, and that the same is true for the law of a single gap between two consecutive eigenvalues. We use this observation to…

Probability · Mathematics 2026-01-12 Samuel G. G. Johnston

Distributed representations of words encode lexical semantic information, but what type of information is encoded and how? Focusing on the skip-gram with negative-sampling method, we found that the squared norm of static word embedding…

Computation and Language · Computer Science 2023-11-03 Momose Oyama , Sho Yokoi , Hidetoshi Shimodaira

Shannon Information theory has achieved great success in not only communication technology where it was originally developed for but also many other science and engineering fields such as machine learning and artificial intelligence.…

Computation and Language · Computer Science 2023-04-26 Arthur Jun Zhang