中文
相关论文

相关论文: An improved method for model selection based on In…

200 篇论文

The Akaike information criterion (AIC) is a model selection criterion widely used in practical applications. The AIC is an estimator of the log-likelihood expected value, and measures the discrepancy between the true model and the estimated…

统计计算 · 统计学 2017-02-03 Fábio M. Bayer , Francisco Cribari-Neto

Claeskens and Hjort (2003) constructed the focused information criterion (FIC) and developed frequentist model averaging methods using maximum likelihood estimators assuming the observations to be independent and identically distributed.…

统计理论 · 数学 2018-07-24 S. C. Pandhare , T. V. Ramanathan

In the era of big data, analysts usually explore various statistical models or machine learning methods for observed data in order to facilitate scientific discoveries or gain predictive power. Whatever data and fitting procedures are…

机器学习 · 统计学 2018-10-24 Jie Ding , Vahid Tarokh , Yuhong Yang

Automatic image description systems are commonly trained and evaluated using crowdsourced, human-generated image descriptions. The best-performing system is then determined using some measure of similarity to the reference data (BLEU,…

计算与语言 · 计算机科学 2020-06-17 Emiel van Miltenburg

The Fisher information approximation (FIA) is an implementation of the minimum description length principle for model selection. Unlike information criteria such as AIC or BIC, it has the advantage of taking the functional form of a model…

统计方法学 · 统计学 2018-08-02 Daniel W. Heck , Morten Moshagen , Edgar Erdfelder

We consider a generalization of an important class of high-dimensional inference problems, namely spiked symmetric matrix models, often used as probabilistic models for principal component analysis. Such paradigmatic models have recently…

信息论 · 计算机科学 2020-05-19 Jean Barbier , Galen Reeves

We provide a brief overview of both Bayes and classical model selection. We argue tentatively that model selection has at least two major goals, that of finding the correct model or predicting well, and that in general both these goals may…

统计理论 · 数学 2015-10-05 Ritabrata Dutta , Malgortaza Bogdan , Jayanta K. Ghosh

From a machine learning point of view, identifying a subset of relevant features from a real data set can be useful to improve the results achieved by classification methods and to reduce their time and space complexity. To achieve this…

机器学习 · 计算机科学 2017-05-23 Pietro Cassara , Alessandro Rozza , Mirco Nanni

We propose a new model selection method, the posterior averaging information criterion, for Bayesian model assessment from a predictive perspective. The theoretical foundation is built on the Kullback-Leibler divergence to quantify the…

统计方法学 · 统计学 2020-09-22 Shouhao Zhou

Feature selection is one of the most fundamental problems in machine learning. An extensive body of work on information-theoretic feature selection exists which is based on maximizing mutual information between subsets of features and class…

机器学习 · 统计学 2016-06-10 Shuyang Gao , Greg Ver Steeg , Aram Galstyan

Recommender systems are significant to help people deal with the world of information explosion and overload. In this Letter, we develop a general framework named self-consistent refinement and implement it be embedding two representative…

数据分析、统计与概率 · 物理学 2008-06-10 Jie Ren , Tao Zhou , Yi-Cheng Zhang

In this work, we propose a modified Bayesian Information Criterion (BIC) specifically designed for mixture models and hierarchical structures. This criterion incorporates the determinant of the Hessian matrix of the log-likelihood function,…

Model selection criteria are one of the most important tools in statistics. Proofs showing a model selection criterion is asymptotically optimal are tailored to the type of model (linear regression, quantile regression, penalized…

统计理论 · 数学 2025-10-17 Amaze Lusompa

In this paper we present a method for reformulating the Recommender Systems problem in an Information Retrieval one. In our tests we have a dataset of users who give ratings for some movies; we hide some values from the dataset, and we try…

信息检索 · 计算机科学 2011-06-03 Alberto Costa , Fabio Roda

When performing regression or classification, we are interested in the conditional probability distribution for an outcome or class variable Y given a set of explanatoryor input variables X. We consider Bayesian models for this task. In…

机器学习 · 计算机科学 2013-02-08 David Heckerman , Christopher Meek

Regression models fitted to data can be assessed on their goodness of fit, though models with many parameters should be disfavored to prevent over-fitting. Statisticians' tools for this are little known to physical scientists. These include…

统计方法学 · 统计学 2013-05-28 Robert S. Maier

Selecting an optimal subset of features or instances under an information theoretic criterion has become an effective preprocessing strategy for reducing data complexity while preserving essential information. This study investigates two…

最优化与控制 · 数学 2025-08-25 Taotao He , Jun Luo , Junkai Zhao

If the assumed model does not accurately capture the underlying structure of the data, a statistical method is likely to yield sub-optimal results, and so model selection is crucial in order to conduct any statistical analysis. However, in…

统计方法学 · 统计学 2023-06-21 Vasilis Chasiotis , Dimitris Karlis

Performing model selection between Gibbs random fields is a very challenging task. Indeed, due to the Markovian dependence structure, the normalizing constant of the fields cannot be computed using standard analytical or numerical methods.…

统计计算 · 统计学 2019-09-04 Julien Stoehr , Jean-Michel Marin , Pierre Pudlo

Comparisons are made for the amount of agreement of the composite likelihood information criteria and their full likelihood counterparts when making decisions among the fits of different models, and some properties of penalty term for…

统计理论 · 数学 2014-10-17 Chi Tim Ng , Harry Joe