English
Related papers

Related papers: Estimation of KL Divergence: Optimal Minimax Rate

200 papers

We present new and improved non-asymptotic deviation bounds for Dirichlet processes (DPs), formulated using the Kullback-Leibler (KL) divergence, which is known for its optimal characterization of the asymptotic behavior of DPs. Our method…

Probability · Mathematics 2025-03-24 Pierre Perrault

We establish bounds on the KL divergence between two multivariate Gaussian distributions in terms of the Hamming distance between the edge sets of the corresponding graphical models. We show that the KL divergence is bounded below by a…

Information Theory · Computer Science 2015-04-06 Varun Jog , Po-Ling Loh

The maximum likelihood method is the best-known method for estimating the probabilities behind the data. However, the conventional method obtains the probability model closest to the empirical distribution, resulting in overfitting. Then…

Machine Learning · Statistics 2023-10-03 Akihisa Ichiki

Obtaining an accurate estimate of the underlying covariance matrix from finite sample size data is challenging due to sample size noise. In recent years, sophisticated covariance-cleaning techniques based on random matrix theory have been…

Computation · Statistics 2024-11-11 Christian Bongiorno , Lamia Lamrani

This paper addresses the problem of distributed detection in fixed and switching networks. A network of agents observe partially informative signals about the unknown state of the world. Hence, they collaborate with each other to identify…

Systems and Control · Computer Science 2016-01-01 Shahin Shahrampour , Alexander Rakhlin , Ali Jadbabaie

Assume that we observe i.i.d.~points lying close to some unknown $d$-dimensional $\mathcal{C}^k$ submanifold $M$ in a possibly high-dimensional space. We study the problem of reconstructing the probability distribution generating the…

Statistics Theory · Mathematics 2022-02-15 Vincent Divol

The goal of this short note is to discuss the relation between Kullback--Leibler divergence and total variation distance, starting with the celebrated Pinsker's inequality relating the two, before switching to a simple, yet (arguably) more…

Probability · Mathematics 2023-08-03 Clément L. Canonne

In this paper, we propose some estimators for the parameters of a statistical model based on Kullback-Leibler divergence of the survival function in continuous setting. We prove that the proposed estimators are subclass of "generalized…

Statistics Theory · Mathematics 2016-07-01 Yaser Mehrali , Majid Asadi

Recent results in quantization theory show that the mean-squared expected distortion can reach a rate of convergence of $\mathcal{O}(1/n)$, where $n$ is the sample size [see, e.g., IEEE Trans. Inform. Theory 60 (2014) 7279-7292 or Electron.…

Statistics Theory · Mathematics 2015-04-02 Clément Levrard

This archiving article consists of several short reports on the discussions between the two authors over the past two years at Oxford and Madrid, and their work carried out during that period on the upper bound of the Kullback-Leibler…

Information Theory · Computer Science 2019-11-20 Min Chen , Mateu Sbert

To characterize the Kullback-Leibler divergence and Fisher information in general parametrized hidden Markov models, in this paper, we first show that the log likelihood and its derivatives can be represented as an additive functional of a…

Statistics Theory · Mathematics 2023-03-15 Cheng-Der Fuh , Chu-Lan Michael Kao , Tianxiao Pang

Complex, high-dimensional data is ubiquitous across many scientific disciplines, including machine learning, biology, and the social sciences. One of the primary methods of visualizing these datasets is with two-dimensional scatter plots…

Machine Learning · Computer Science 2025-10-13 Kiran Smelser , Kaviru Gunaratne , Jacob Miller , Stephen Kobourov

We give a general unified method that can be used for $L_1$ {\em closeness testing} of a wide range of univariate structured distribution families. More specifically, we design a sample optimal and computationally efficient algorithm for…

Data Structures and Algorithms · Computer Science 2015-08-25 Ilias Diakonikolas , Daniel M. Kane , Vladimir Nikishkin

We use the fitted Pareto law to construct an accompanying approximation of the excess distribution function. A selection rule of the location of the excess distribution function is proposed based on a stagewise lack-of-fit testing…

Statistics Theory · Mathematics 2008-08-08 Ion Grama , Vladimir Spokoiny

In this study, simultaneous predictive distributions for independent Poisson observables were considered and the performance of predictive distributions was evaluated using the Kullback-Leibler (K-L) loss. This study proposes a class of…

Statistics Theory · Mathematics 2024-02-13 Xiao Li

We introduce one-sided versions of Huber's contamination model, in which corrupted samples tend to take larger values than uncorrupted ones. Two intertwined problems are addressed: estimation of the mean of uncorrupted samples (minimum…

Statistics Theory · Mathematics 2018-09-25 Alexandra Carpentier , Sylvain Delattre , Etienne Roquain , Nicolas Verzelen

Recently, we have proposed a maximum likelihood iterative algorithm for estimation of the parameters of the Nakagami-m distribution. This technique performs better than state of art estimation techniques for this distribution. This could be…

Machine Learning · Computer Science 2014-02-04 Rangeet Mitra , Amit Kumar Mishra , Tarun Choubisa

Many real-life data sets can be analyzed using Linear Mixed Models (LMMs). Since these are ordinarily based on normality assumptions, under small deviations from the model the inference can be highly unstable when the associated parameters…

Methodology · Statistics 2024-02-06 Giovanni Saraceno , Abhik Ghosh , Ayanendranath Basu , Claudio Agostinelli

This paper studies the estimation of low-rank Markov chains from empirical trajectories. We propose a non-convex estimator based on rank-constrained likelihood maximization. Statistical upper bounds are provided for the Kullback-Leiber…

Machine Learning · Statistics 2018-07-20 Xudong Li , Mengdi Wang , Anru Zhang

In statistical classification/multiple hypothesis testing and machine learning, a model distribution estimated from the training data is usually applied to replace the unknown true distribution in the Bayes decision rule, which introduces a…

Information Theory · Computer Science 2024-09-24 Zijian Yang , Vahe Eminyan , Ralf Schlüter , Hermann Ney