English
Related papers

Related papers: Model selection by minimum description length: Low…

200 papers

Over the past decade, interval arithmetic (IA) has been utilized to determine tolerance bounds of phased array beampatterns. IA only requires that the errors of the array elements are bounded, and can provide reliable beampattern bounds…

Signal Processing · Electrical Eng. & Systems 2023-06-26 Håvard Kjellmo Arnestad , Gábor Geréb , Tor Inge Birkenes Lønmo , Jan Egil Kirkebø , Andreas Austeng , Sven Peter Näsholm

The problem how to approximately determine the absolute value of the Fisher information measure for a general parametric probabilistic system is considered. Having available the first and second moment of the system output in a parametric…

Information Theory · Computer Science 2015-06-16 Manuel Stein , Amine Mezghani , Josef A. Nossek

Completely random measures (CRMs) and their normalizations (NCRMs) offer flexible models in Bayesian nonparametrics. But their infinite dimensionality presents challenges for inference. Two popular finite approximations are truncated finite…

Methodology · Statistics 2023-11-07 Tin D. Nguyen , Jonathan Huggins , Lorenzo Masoero , Lester Mackey , Tamara Broderick

The information criterion for determining the number of explanatory variables in a subset regression modeling is discussed. Information criterion such as AIC is effective and frequently used in model selection for ordinary regression models…

Methodology · Statistics 2023-09-18 Genshiro Kitagawa

Exponential models of distributions are widely used in machine learning for classiffication and modelling. It is well known that they can be interpreted as maximum entropy models under empirical expectation constraints. In this work, we…

Machine Learning · Computer Science 2012-07-19 Amir Globerson , Naftali Tishby

When constructing models of the world, we aim for optimal compressions: models that include as few details as possible while remaining as accurate as possible. But which details -- or features measured in data -- should we choose to include…

Quantitative Methods · Quantitative Biology 2025-05-06 David P. Carcamo , Nicholas J. Weaver , Purushottam D. Dixit , Christopher W. Lynn

Minimum message length is a general Bayesian principle for model selection and parameter estimation that is based on information theory. This paper applies the minimum message length principle to a small-sample model selection problem…

Methodology · Statistics 2018-02-13 Chi Kuen Wong , Enes Makalic , Daniel F. Schmidt

Modern applications of Bayesian inference involve models that are sufficiently complex that the corresponding posterior distributions are intractable and must be approximated. The most common approximation is based on Markov chain Monte…

Machine Learning · Statistics 2019-05-15 Yue Yang , Ryan Martin , Howard Bondell

Definition bias is a negative phenomenon that can mislead models. Definition bias in information extraction appears not only across datasets from different domains but also within datasets sharing the same domain. We identify two types of…

Computation and Language · Computer Science 2024-03-26 Wenhao Huang , Qianyu He , Zhixu Li , Jiaqing Liang , Yanghua Xiao

We consider the free non-commutative analogue Phi^*, introduced by D. Voiculescu, of the concept of Fisher information for random variables. We determine the minimal possible value of Phi^*(a,a^*), if a is a non-commutative random variable…

Operator Algebras · Mathematics 2007-05-23 A. Nica , D. Shlyakhtenko , R. Speicher

My dissertation revolves around Bayesian approaches towards constrained statistical inference in the factor analysis (FA) model. Two interconnected types of restricted-model selection are considered. These types have a natural connection to…

Applications · Statistics 2016-04-13 Carel F. W. Peeters

Deep neural networks often fail to adapt representations to novel tasks under distribution shifts, especially when only a few examples are available. This paper identifies a core obstacle behind this failure: channel bias, where networks…

Computer Vision and Pattern Recognition · Computer Science 2025-09-11 Ji Zhang , Xu Luo , Lianli Gao , Difan Zou , Hengtao Shen , Jingkuan Song

Fisher discriminant analysis (FDA) is a widely used method for classification and dimensionality reduction. When the number of predictor variables greatly exceeds the number of observations, one of the alternatives for conventional FDA is…

Machine Learning · Statistics 2018-11-30 Agniva Chowdhury , Jiasen Yang , Petros Drineas

Importance sampling approximates expectations with respect to a target measure by using samples from a proposal measure. The performance of the method over large classes of test functions depends heavily on the closeness between both…

Computation · Statistics 2016-09-01 Daniel Sanz-Alonso

We propose the approximate Laplace approximation (ALA) to evaluate integrated likelihoods, a bottleneck in Bayesian model selection. The Laplace approximation (LA) is a popular tool that speeds up such computation and equips strong model…

Computation · Statistics 2021-10-07 David Rossell , Oriol Abril , Anirban Bhattacharya

In classification problems, the purpose of feature selection is to identify a small, highly discriminative subset of the original feature set. In many applications, the dataset may have thousands of features and only a few dozens of samples…

Machine Learning · Computer Science 2020-08-28 Ludmila I. Kuncheva , Clare E. Matthews , Álvar Arnaiz-González , Juan J. Rodríguez

Model selection and order selection problems frequently arise in statistical practice. A popular approach to addressing these problems in the frequentist setting involves information criteria based on penalised maxima of log-likelihoods for…

Statistics Theory · Mathematics 2025-10-29 Hien Duy Nguyen , Mayetri Gupta , Jacob Westerhout , TrungTin Nguyen

Naively trained AI models can be heavily biased. This can be particularly problematic when the biases involve legally or morally protected attributes such as ethnic background, age or gender. Existing solutions to this problem come at the…

Computer Vision and Pattern Recognition · Computer Science 2022-10-11 Nicholas Rosa , Tom Drummond , Mehrtash Harandi

We emphasize that it is possible to improve the principle of unbiased risk estimation for model selection by addressing excess risk deviations in the design of penalization procedures. Indeed, we propose a modification of Akaike's…

Statistics Theory · Mathematics 2018-07-23 Adrien Saumard , Fabien Navarro

Granger causality analysis (GCA) provides a powerful tool for uncovering the patterns of brain connectivity mechanism using neuroimaging techniques. Conventional GCA applies two different mathematical theories in a two-stage scheme: (1) the…

Methodology · Statistics 2020-04-28 Fei Li , Xuewei Wang , Qiang Lin , Zhenghui Hu