English
Related papers

Related papers: Model selection by minimum description length: Low…

200 papers

The paper considers model selection in regression under the additional structural constraints on admissible models where the number of potential predictors might be even larger than the available sample size. We develop a Bayesian formalism…

Statistics Theory · Mathematics 2013-02-19 Felix Abramovich , Vadim Grinshtein

We consider Bayesian model selection in generalized linear models that are high-dimensional, with the number of covariates p being large relative to the sample size n, but sparse in that the number of active covariates is small compared to…

Statistics Theory · Mathematics 2011-12-26 Rina Foygel , Mathias Drton

Mixture-of-Experts models are commonly used when there exist distinct clusters with different relationships between the independent and dependent variables. Fitting such models for large datasets, however, is computationally virtually…

Methodology · Statistics 2023-09-06 Yanxi Liu , John Stufken , Min Yang

We develop a general method to study the Fisher information distance in central limit theorem for nonlinear statistics. We first construct completely new representations for the score function. We then use these representations to derive…

Probability · Mathematics 2024-09-23 Nguyen Tien Dung

The expectation-maximization (EM) algorithm is an iterative computational method to calculate the maximum likelihood estimators (MLEs) from the sample data. It converts a complicated one-time calculation for the MLE of the incomplete data…

Computation · Statistics 2016-08-08 Lingyao Meng

For linear models with a diverging number of parameters, it has recently been shown that modified versions of Bayesian information criterion (BIC) can identify the true model consistently. However, in many cases there is little…

Methodology · Statistics 2011-07-26 Heng Lian

Importance sampling is used to approximate Bayes' rule in many computational approaches to Bayesian inverse problems, data assimilation and machine learning. This paper reviews and further investigates the required sample size for…

Computation · Statistics 2021-02-03 Daniel Sanz-Alonso , Zijian Wang

Feature selection is frequently used as a pre-processing step to machine learning. It is a process of choosing a subset of original features so that the feature space is optimally reduced according to a certain evaluation criterion. The…

Computer Vision and Pattern Recognition · Computer Science 2014-01-07 Vijendra Singh , Shivani Pathak

Model selection is central to statistics, and many learning problems can be formulated as model selection problems. In this paper, we treat the problem of selecting a maximum entropy model given various feature subsets and their moments, as…

Information Theory · Computer Science 2013-11-28 Gaurav Pandey , Ambedkar Dukkipati

Many researches proposed the use of the noon state as the input state for phase estimation, which is one topic of quantum metrology. This is because the input noon state provides the maximum Fisher information at the specific point.…

Quantum Physics · Physics 2011-06-24 Masahito Hayashi

We proposed a Least Information theory (LIT) to quantify meaning of information in probability distribution changes, from which a new information retrieval model was developed. We observed several important characteristics of the proposed…

Information Retrieval · Computer Science 2012-05-03 Weimao Ke

This paper applies the minimum message length principle to inference of linear regression models with Student-t errors. A new criterion for variable selection and parameter estimation in Student-t regression is proposed. By exploiting…

Methodology · Statistics 2018-02-21 Chi Kuen Wong , Enes Makalic , Daniel F. Schmidt

This paper compares three approaches to the problem of selecting among probability models to fit data (1) use of statistical criteria such as Akaike's information criterion and Schwarz's "Bayesian information criterion," (2) maximization of…

Methodology · Statistics 2016-11-04 William B. Poland , Ross D. Shachter

Information theory provides a useful tool to understand the evolution of complex nonlinear systems and their sustainability. In particular, Fisher Information (FI) has been evoked as a useful measure of sustainability and the variability of…

Dynamical Systems · Mathematics 2016-08-18 Avan Al-Saffar , Eun-jin Kim

We study model selection by the Bayesian information criterion (BIC) in fixed-dimensional exploratory factor analysis over a fixed finite family of compact covariance classes. Our main result shows that the BIC is strongly consistent for…

Statistics Theory · Mathematics 2026-04-10 Hien Duy Nguyen , Kei Hirose

We use the language of uninformative Bayesian prior choice to study the selection of appropriately simple effective models. We advocate for the prior which maximizes the mutual information between parameters and predictions, learning as…

Data Analysis, Statistics and Probability · Physics 2018-02-16 Henry H. Mattingly , Mark K. Transtrum , Michael C. Abbott , Benjamin B. Machta

Central to several objective approaches to Bayesian model selection is the use of training samples (subsets of the data), so as to allow utilization of improper objective priors. The most common prescription for choosing training samples is…

Statistics Theory · Mathematics 2007-06-13 James O. Berger , Luis R. Pericchi

Widespread use of artificial intelligence (AI) algorithms and machine learning (ML) models on the one hand and a number of crucial issues pertaining to them warrant the need for explainable artificial intelligence (XAI). A key…

Artificial Intelligence · Computer Science 2023-12-13 Jinqiang Yu , Graham Farr , Alexey Ignatiev , Peter J. Stuckey

The quality of numerical reconstructions for unknown parameters in inverse problems depends fundamentally on the selection of experimental data. To ensure a robust reconstruction, it is crucial to select data that are sensitive to the…

Numerical Analysis · Mathematics 2026-04-14 Kathrin Hellmuth , Christian Klingenberg , Qin Li

Scientific computer simulations cannot represent all scales in realistic applications. To bridge this model-data gap, parameters are injected into models and constrained with noisy data using Bayesian inversion. To reduce the number of…

Computation · Statistics 2026-05-22 Arne Bouillon , Oliver R. A. Dunbar