English
Related papers

Related papers: Model selection by minimum description length: Low…

200 papers

Information divergence functions play a critical role in statistics and information theory. In this paper we show that a non-parametric f-divergence measure can be used to provide improved bounds on the minimum binary classification…

Information Theory · Computer Science 2015-02-11 Visar Berisha , Alan Wisler , Alfred O. Hero , Andreas Spanias

Motivated by the growing interest in quantum machine learning, in particular quantum neural networks (QNNs), we study how recently introduced evaluation metrics based on the Fisher information matrix (FIM) are effective for predicting their…

Machine Learning · Computer Science 2025-10-09 Lorenzo Pastori , Veronika Eyring , Mierk Schwabe

When developing a clinical prediction model, the sample size of the development dataset is a key consideration. Small sample sizes lead to greater concerns of overfitting, instability, poor performance and lack of fairness. Previous…

We consider the problem of learning high-dimensional, nonparametric and structured (e.g. Gaussian) distributions in distributed networks, where each node in the network observes an independent sample from the underlying distribution and can…

Information Theory · Computer Science 2019-06-04 Leighton Pate Barnes , Yanjun Han , Ayfer Ozgur

The Fisher information matrix provides a way to measure the amount of information given observed data based on parameters of interest. Many applications of the FIM exist in statistical modeling, system identification, and parameter…

Computation · Statistics 2021-04-16 Xuan Wu

Inference from limited data requires a notion of measure on parameter space, most explicit in the Bayesian framework as a prior. Here we demonstrate that Jeffreys prior, the best-known uninformative choice, introduces enormous bias when…

Other Statistics · Statistics 2023-04-03 Michael C. Abbott , Benjamin B. Machta

The Fisher information matrix summarizes the amount of information in a set of data relative to the quantities of interest. There are many applications of the information matrix in statistical modeling, system identification and parameter…

Computation · Statistics 2014-05-08 Xumeng Cao

The Fisher information matrix (FIM) is a foundational concept in statistical signal processing. The FIM depends on the probability distribution, assumed to belong to a smooth parametric family. Traditional approaches to estimating the FIM…

Computation · Statistics 2015-06-22 Visar Berisha , Alfred O. Hero

A bias correction to Akaike's information criterion (AIC) is derived for seemingly unrelated regressions models. The correction is of particular use when the sample size is not much larger than the number of fitted parameters. A…

Methodology · Statistics 2009-06-05 J. L. van Velsen

Information criteria are an appropriate and widely used tool for solving model selection problems. However, different ways to use them exist, each leading to a more or less precise approximation of the sought model. In this paper, we mainly…

Statistics Theory · Mathematics 2007-06-13 Guilhem Coq , Olivier Alata , Marc Arnaudon , Christian Olivier

The Akaike information criterion (AIC) is a model selection criterion widely used in practical applications. The AIC is an estimator of the log-likelihood expected value, and measures the discrepancy between the true model and the estimated…

Computation · Statistics 2017-02-03 Fábio M. Bayer , Francisco Cribari-Neto

In this paper the stochastic complexity criterion is applied to estimation of the order in AR and ARMA models. The power of the criterion for short strings is illustrated by simulations. It requires an integral of the square root of Fisher…

Statistics Theory · Mathematics 2007-06-13 Ciprian Doru Giurcăneanu , Jorma Rissanen

The main purpose of this paper is to introduce and study the behavior of minimum {\phi}-divergence estimators as an alternative to the maximum likelihood estimator in latent class models for binary items. As it will become clear below,…

Methodology · Statistics 2014-06-03 Ángel Felipe , Pedro Miranda , Leandro Pardo

Clinical prediction models enable healthcare professionals to estimate individual outcomes using patient characteristics. Current sample size guidelines for developing or updating models with continuous outcomes aim to minimise overfitting…

We consider the processing of statistical samples $X\sim P_\theta$ by a channel $p(y|x)$, and characterize how the statistical information from the samples for estimating the parameter $\theta\in\mathbb{R}^d$ can scale with the mutual…

Information Theory · Computer Science 2021-07-12 Leighton Pate Barnes , Ayfer Ozgur

Information theoretic criteria (ITC) have been widely adopted in engineering and statistics for selecting, among an ordered set of candidate models, the one that better fits the observed sample data. The selected model minimizes a penalized…

Machine Learning · Statistics 2019-10-10 Andrea Mariani , Andrea Giorgetti , Marco Chiani

Factorized information criterion (FIC) is a recently developed approximation technique for the marginal log-likelihood, which provides an automatic model selection framework for a few latent variable models (LVMs) with tractable inference…

Machine Learning · Computer Science 2015-04-23 Kohei Hayashi , Shin-ichi Maeda , Ryohei Fujimaki

This paper considers the problem of estimation of the Fisher information for location from a random sample of size $n$. First, an estimator proposed by Bhattacharya is revisited and improved convergence rates are derived. Second, a new…

Information Theory · Computer Science 2020-05-08 Wei Cao , Alex Dytso , Michael Fauß , H. Vincent Poor , Gang Feng

We study a scenario where a group of agents, each with multiple heterogeneous sensors are collecting measurements of a vehicle and the measurements are transmitted over a communication channel to a centralized node for processing. The…

Systems and Control · Electrical Eng. & Systems 2021-04-21 Matthew R. Kirchner , João P. Hespanha , Denis Garagić

Machine learning models trained on uncurated datasets can often end up adversely affecting inputs belonging to underrepresented groups. To address this issue, we consider the problem of adaptively constructing training sets which allow us…

Machine Learning · Computer Science 2021-07-21 Shubhanshu Shekhar , Greg Fields , Mohammad Ghavamzadeh , Tara Javidi