English
Related papers

Related papers: Direct Estimation of Information Divergence Using …

200 papers

The problem of estimation of density functionals like entropy and mutual information has received much attention in the statistics and information theory communities. A large class of estimators of functionals of the probability density…

Statistics Theory · Mathematics 2013-03-05 Kumar Sricharan , Dennis Wei , Alfred O. Hero

Many practical problems are related to the pointwise estimation of dis- tribution functions when data contains measurement errors. Motivation for these problems comes from diverse fields such as astronomy, reliability, quality control,…

Methodology · Statistics 2012-02-21 I. Dattner , B. Reiser

Density ratio estimation (DRE) is a fundamental machine learning technique for comparing two probability distributions. However, existing methods struggle in high-dimensional settings, as it is difficult to accurately compare probability…

Machine Learning · Computer Science 2022-03-15 Kristy Choi , Chenlin Meng , Yang Song , Stefano Ermon

A new approach to $L_2$-consistent estimation of a general density functional using $k$-nearest neighbor distances is proposed, where the functional under consideration is in the form of the expectation of some function $f$ of the densities…

Statistics Theory · Mathematics 2022-03-14 J. Jon Ryu , Shouvik Ganguly , Young-Han Kim , Yung-Kyun Noh , Daniel D. Lee

Multiple regression has been the go-to method for data analysis for generations of scholars due to its transparency, interpretability, and desirable theoretical properties. However, the method's simplicity precludes the discovery of complex…

Machine Learning · Statistics 2021-02-02 Marc Ratkovic , Dustin Tingley

We introduce the Information-Estimation Metric (IEM), a novel form of distance function derived from an underlying continuous probability density over a domain of signals. The IEM is rooted in a fundamental relationship between information…

Image and Video Processing · Electrical Eng. & Systems 2026-02-09 Guy Ohayon , Pierre-Etienne H. Fiquet , Florentin Guth , Jona Ballé , Eero P. Simoncelli

k Nearest Neighbor (kNN) method is a simple and popular statistical method for classification and regression. For both classification and regression problems, existing works have shown that, if the distribution of the feature vector has…

Statistics Theory · Mathematics 2019-10-24 Puning Zhao , Lifeng Lai

Four estimators of the directed information rate between a pair of jointly stationary ergodic finite-alphabet processes are proposed, based on universal probability assignments. The first one is a Shannon--McMillan--Breiman type estimator,…

Information Theory · Computer Science 2016-11-15 Jiantao Jiao , Haim H. Permuter , Lei Zhao , Young-Han Kim , Tsachy Weissman

In this work we consider a model problem of deep neural learning, namely the learning of a given function when it is assumed that we have access to its point values on a finite set of points. The deep neural network interpolant is the the…

Machine Learning · Statistics 2023-06-27 Michail Loulakis , Charalambos G. Makridakis

We consider the regression model with errors-in-variables where we observe $n$ i.i.d. copies of $(Y,Z)$ satisfying $Y=f(X)+\xi, Z=X+\sigma\epsilon$, involving independent and unobserved random variables $X,\xi,\epsilon$. The density $g$ of…

Statistics Theory · Mathematics 2008-02-11 Fabienne Comte , Marie-Luce Taupin

This paper develops a new framework for indirect statistical inference with guaranteed necessity and sufficiency, applicable to continuous random variables. We prove that when comparing exponentially transformed order statistics from an…

Statistics Theory · Mathematics 2025-09-25 Z Zhang , X Hu , C Lu , T Liu

Directed acyclic graphs provide a fundamental tool for representing directed dependence structures in multivariate network data, and are widely used to model financial and economic networks. However, accurate and interpretable estimation…

Methodology · Statistics 2026-05-26 Huihang Liu , Wenhui Li , Xinyu Zhang

Big data mining is well known to be an important task for data science, because it can provide useful observations and new knowledge hidden in given large datasets. Proximity-based data analysis is particularly utilized in many real-life…

Databases · Computer Science 2022-11-29 Daichi Amagata , Yusuke Arai , Sumio Fujita , Takahiro Hara

The association between a continuous and an ordinal variable is commonly modeled through the polyserial correlation model. However, this model, which is based on a partially-latent normality assumption, may be misspecified in practice, due…

Methodology · Statistics 2026-02-11 Max Welz

We demonstrate that a popular class of nonparametric mutual information (MI) estimators based on k-nearest-neighbor graphs requires number of samples that scales exponentially with the true MI. Consequently, accurate estimation of MI…

Information Theory · Computer Science 2015-03-09 Shuyang Gao , Greg Ver Steeg , Aram Galstyan

This work investigates the electrical impedance tomography (EIT) problem when only limited boundary measurements are available, which is known to be challenging due to the extreme ill-posedness. Based on the direct sampling method (DSM), we…

Numerical Analysis · Mathematics 2020-09-18 Ruchi Guo , Jiahua Jiang

In this work, we propose new objective functions to train deep neural network based density ratio estimators and apply it to a change point detection problem. Existing methods use linear combinations of kernels to approximate the density…

Machine Learning · Computer Science 2019-05-27 Haidar Khan , Lara Marcuse , Bülent Yener

$k$-nearest neighbour ($k$-NN) is one of the simplest and most widely-used methods for supervised classification, that predicts a query's label by taking weighted ratio of observed labels of $k$ objects nearest to the query. The weights and…

Machine Learning · Statistics 2020-11-12 Akifumi Okuno , Hidetoshi Shimodaira

Change-point analysis is thriving in this big data era to address problems arising in many fields where massive data sequences are collected to study complicated phenomena over time. It plays an important role in processing these data by…

Methodology · Statistics 2022-03-23 Yi-Wei Liu , Hao Chen

Expected values weighted by the inverse of a multivariate density or, equivalently, Lebesgue integrals of regression functions with multivariate regressors occur in various areas of applications, including estimating average treatment…

Statistics Theory · Mathematics 2025-02-17 Hajo Holzmann , Alexander Meister