English
Related papers

Related papers: Algorithmic randomness and the weak merging of com…

200 papers

We consider a Markov chain that iteratively generates a sequence of random finite words in such a way that the $n^{\mathrm{th}}$ word is uniformly distributed over the set of words of length $2n$ in which $n$ letters are $a$ and $n$ letters…

Probability · Mathematics 2016-12-23 Hye Soo Choi , Steven N. Evans

A theoretical framework for non-negative matrix factorization based on generalized dual Kullback-Leibler divergence, which includes members of the exponential family of models, is proposed. A family of algorithms is developed using this…

Machine Learning · Statistics 2019-05-20 Karthik Devarajan

Transfer learning, or domain adaptation, is concerned with machine learning problems in which training and testing data come from possibly different distributions (denoted as $\mu$ and $\mu'$, respectively). In this work, we give an…

Machine Learning · Computer Science 2020-05-20 Xuetong Wu , Jonathan H. Manton , Uwe Aickelin , Jingge Zhu

In this paper, we compare the performance of two methods for estimating Bayesian networks from data containing exogenous variables and random effects. The first method is fully Bayesian in which a prior distribution is placed on the…

Methodology · Statistics 2011-12-02 Jessica Kasza , Patty Solomon

This paper proposes a distributionally robust unit commitment approach for microgrids under net load and electricity market price uncertainty. The key thrust of the proposed approach is to leverage the Kullback-Leibler divergence to…

Optimization and Control · Mathematics 2020-12-15 Ogun Yurdakul , Fikret Sivrikaya , Sahin Albayrak

Kullback--Leibler (KL) divergence is a fundamental measure of the dissimilarity between two probability distributions, but it can become unstable in high-dimensional settings due to its sensitivity to mismatches in distributional support.…

Information Theory · Computer Science 2025-02-03 Yifeng Peng , Dantong Li , Xinyi Li , Zhiding Liang , Yongshan Ding , Ying Wang

$f$-divergences, which quantify discrepancy between probability distributions, are ubiquitous in information theory, machine learning, and statistics. While there are numerous methods for estimating $f$-divergences from data, a limit…

Statistics Theory · Mathematics 2023-10-13 Sreejith Sreekumar , Ziv Goldfeld , Kengo Kato

We study the interaction between polynomial space randomness and a fundamental result of analysis, the Lebesgue differentiation theorem. We generalize Ko's framework for polynomial space computability in $\mathbb{R}^n$ to define…

Computational Complexity · Computer Science 2016-04-27 Xiang Huang , D. M. Stull

This paper presents a new derivation of the variational Poisson multi-Bernoulli (V-PMB) filter for multi-target estimation proposed in [#Williams15]. The proposed derivation is based on considering an augmented space that includes the set…

Systems and Control · Electrical Eng. & Systems 2026-05-11 Ángel F. García-Fernández , Yuxuan Xia

There are three classical divergence measures in the literature on information theory and statistics, namely, Jeffryes-Kullback-Leiber's J-divergence, Sibson-Burbea-Rao's Jensen-Shannon divegernce and Taneja's arithemtic-geometric mean…

Statistics Theory · Mathematics 2007-06-13 Inder Jeet Taneja

The non-parametric version of Amari's dually affine Information Geometry provides a practical calculus to perform computations of interest in statistical machine learning. The method uses the notion of a statistical bundle, a mathematical…

Statistics Theory · Mathematics 2025-04-07 Giovanni Pistone

We develop a fully non-parametric, easy-to-use, and powerful test for the missing completely at random (MCAR) assumption on the missingness mechanism of a dataset. The test compares distributions of different missing patterns on random…

Methodology · Statistics 2022-12-01 Meta-Lina Spohn , Jeffrey Näf , Loris Michel , Nicolai Meinshausen

Meta-analytic methods tend to take all-or-nothing approaches to study-level heterogeneity, assuming all studies are heterogeneous or homogeneous, leading to inefficiency and/or bias in estimation and inference. In this paper, we develop a…

Methodology · Statistics 2026-03-12 Elizabeth M. Davis , Emily C. Hector

Estimating Kullback-Leibler divergence from identical and independently distributed samples is an important problem in various domains. One simple and effective estimator is based on the k nearest neighbor distances between these samples.…

Information Theory · Computer Science 2020-02-27 Puning Zhao , Lifeng Lai

We address the problem of Schr\"odinger potential estimation, which plays a crucial role in modern generative modelling approaches based on Schr\"odinger bridges and stochastic optimal control for SDEs. Given a simple prior diffusion…

Machine Learning · Computer Science 2025-06-04 Nikita Puchkin , Iurii Pustovalov , Yuri Sapronov , Denis Suchkov , Alexey Naumov , Denis Belomestny

Universal hypothesis testing refers to the problem of deciding whether samples come from a nominal distribution or an unknown distribution that is different from the nominal distribution. Hoeffding's test, whose test statistic is equivalent…

Information Theory · Computer Science 2017-11-15 Pengfei Yang , Biao Chen

We study optimal solutions to an abstract optimization problem for measures, which is a generalization of classical variational problems in information theory and statistical physics. In the classical problems, information and relative…

Optimization and Control · Mathematics 2021-11-23 Roman V. Belavkin

Statistical distances (SDs), which quantify the dissimilarity between probability distributions, are central to machine learning and statistics. A modern method for estimating such distances from data relies on parametrizing a variational…

Statistics Theory · Mathematics 2021-03-18 Sreejith Sreekumar , Zhengxin Zhang , Ziv Goldfeld

We study quasi-Newton methods from the viewpoint of information geometry induced associated with Bregman divergences. Fletcher has studied a variational problem which derives the approximate Hessian update formula of the quasi-Newton…

Computation · Statistics 2010-10-15 Takafumi Kanamori , Atsumi Ohara

In this paper we formulate in general terms an approach to prove strong consistency of the Empirical Risk Minimisation inductive principle applied to the prototype or distance based clustering. This approach was motivated by the Divisive…

Machine Learning · Statistics 2010-04-20 Vladimir Nikulin , Geoffrey J. McLachlan