English
Related papers

Related papers: Information Criteria Fail for Dynamical Systems: S…

200 papers

In data-driven optimization, the sample performance of the obtained decision typically incurs an optimistic bias against the true performance, a phenomenon commonly known as the Optimizer's Curse and intimately related to overfitting in…

Machine Learning · Computer Science 2025-07-22 Garud Iyengar , Henry Lam , Tianyu Wang

Deep learning is renowned for its theory-practice gap, whereby principled theory typically fails to provide much beneficial guidance for implementation in practice. This has been highlighted recently by the benign overfitting phenomenon:…

Machine Learning · Statistics 2023-11-14 Liam Hodgkinson , Chris van der Heide , Robert Salomone , Fred Roosta , Michael W. Mahoney

We introduce a generalized information criterion that contains other well-known information criteria, such as Bayesian information Criterion (BIC) and Akaike information criterion (AIC), as special cases. Furthermore, the proposed spectral…

Methodology · Statistics 2023-08-21 L. Martino , R. San Millan-Castillo , E. Morgado

We study model selection by the Bayesian information criterion (BIC) in fixed-dimensional exploratory factor analysis over a fixed finite family of compact covariance classes. Our main result shows that the BIC is strongly consistent for…

Statistics Theory · Mathematics 2026-04-10 Hien Duy Nguyen , Kei Hirose

Vine copulas allow to build flexible dependence models for an arbitrary number of variables using only bivariate building blocks. The number of parameters in a vine copula model increases quadratically with the dimension, which poses new…

Methodology · Statistics 2018-11-20 Thomas Nagler , Christian Bumann , Claudia Czado

Information coefficient (IC) is a widely used metric for measuring investment managers' skills in selecting stocks. However, its adequacy and effectiveness for evaluating stock selection models has not been clearly understood, as IC from a…

Computational Finance · Quantitative Finance 2020-10-20 Feng Zhang , Ruite Guo , Honggao Cao

We provide a brief overview of both Bayes and classical model selection. We argue tentatively that model selection has at least two major goals, that of finding the correct model or predicting well, and that in general both these goals may…

Statistics Theory · Mathematics 2015-10-05 Ritabrata Dutta , Malgortaza Bogdan , Jayanta K. Ghosh

Longitudinal data are common in clinical trials and observational studies, where missing outcomes due to dropouts are always encountered. Under such context with the assumption of missing at random, the weighted generalized estimating…

Methodology · Statistics 2019-04-30 Chixiang Chen , Biyi Shen , Lijun Zhang , Yuan Xue , Ming Wang

The widely applicable information criterion (WAIC) has been used as a model selection criterion for Bayesian statistics in recent years. It is an asymptotically unbiased estimator of the Kullback-Leibler divergence between a Bayesian…

Methodology · Statistics 2022-08-09 Yoshiyuki Ninomiya

Performing model selection between Gibbs random fields is a very challenging task. Indeed, due to the Markovian dependence structure, the normalizing constant of the fields cannot be computed using standard analytical or numerical methods.…

Computation · Statistics 2019-09-04 Julien Stoehr , Jean-Michel Marin , Pierre Pudlo

We propose a new parameter-adaptive uncertainty-penalized Bayesian information criterion (UBIC) to prioritize the parsimonious partial differential equation (PDE) that sufficiently governs noisy spatial-temporal observed data with few…

Machine Learning · Computer Science 2024-01-31 Pongpisit Thanasutives , Takashi Morita , Masayuki Numao , Ken-ichi Fukui

We consider a sparse linear regression model, when the number of available predictors, $p$, is much larger than the sample size, $n$, and the number of non-zero coefficients, $p_0$, is small. To choose the regression model in this…

Statistics Theory · Mathematics 2018-05-31 Piotr Szulc

This thesis focuses on the discovery of stochastic differential equations (SDEs) and stochastic partial differential equations (SPDEs) from noisy and discrete time series. A major challenge is selecting the simplest possible correct model…

Machine Learning · Statistics 2025-07-08 Andonis Gerardos

In segmented regression, when the regression function is continuous at the change-points that are the boundaries of the segments, it is also called joinpoint regression, and the analysis package developed by \cite{KimFFM00} has become a…

Methodology · Statistics 2025-06-11 Kazuki Nakajima , Yoshiyuki Ninomiya

We have recently proposed a new information-based approach to model selection, the Frequentist Information Criterion (FIC), that reconciles information-based and frequentist inference. The purpose of this current paper is to provide a…

Data Analysis, Statistics and Probability · Physics 2015-06-23 Paul A. Wiggins

We consider the use of Bayesian information criteria for selection of the graph underlying an Ising model. In an Ising model, the full conditional distributions of each variable form logistic regression models, and variable selection…

Statistics Theory · Mathematics 2015-03-09 Rina Foygel Barber , Mathias Drton

Occupancy models are typically used to determine the probability of a species being present at a given site while accounting for imperfect detection. The survey data underlying these models often include information on several predictors…

Methodology · Statistics 2016-05-09 Daniel Taylor-Rodriguez , Andrew Womack , Claudio Fuentes , Nikolay Bliznyuk

Model selection in mixed models based on the conditional distribution is appropriate for many practical applications and has been a focus of recent statistical research. In this paper we introduce the R-package cAIC4 that allows for the…

Computation · Statistics 2018-03-20 Benjamin Säfken , David Rügamer , Thomas Kneib , Sonja Greven

We present a strategy for selecting the values of elasticity parameters by comparing walk-away vertical seismic profiling data with a multilayered model in the context of Bayesian Information Criterion. We consider $P$-wave traveltimes and…

In this article we propose a general class of risk measures which can be used for data based evaluation of parametric models. The loss function is defined as generalized quadratic distance between the true density and the proposed model.…

Statistics Theory · Mathematics 2007-10-02 Surajit Ray , Bruce G. Lindsay