English
Related papers

Related papers: Parsimony in Model Selection: Tools for Assessing …

200 papers

Fitting models for non-Poisson point processes is complicated by the lack of tractable models for much of the data. By using large samples of independent and identically distributed realizations and statistical learning, it is possible to…

Methodology · Statistics 2007-12-04 Jeffrey Picka , Mingxia Deng

Multi-model mimicry (MMM) is a flexible model selection technique for comparison of multiple, non-nested models on any desired goodness-of-fit criteria. Applicable to any set of candidate models that are 1) able to be fit to observed data,…

Methodology · Statistics 2019-12-17 Lachlann McArthur , Melissa A. Humphries

We introduce an enumeration-free method based on mathematical programming to precisely characterize various properties such as fairness or sparsity within the set of "good models", known as Rashomon set. This approach is generically…

Machine Learning · Computer Science 2025-07-08 Lucas Langlade , Julien Ferry , Gabriel Laberge , Thibaut Vidal

Additive models enjoy the flexibility of nonlinear models while still being readily understandable to humans. By contrast, other nonlinear models, which involve interactions between features, are not only harder to fit but also…

Methodology · Statistics 2025-06-03 Yiling Huang , Snigdha Panigrahi , Guo Yu , Jacob Bien

A formal theory of simplicity is introduced, in the context of a "combinational" computation model that views computation as comprising the iterated transformational and compositional activity of a population of agents upon each other.…

Artificial Intelligence · Computer Science 2020-09-04 Ben Goertzel

In this note we introduce linear regression with basis functions in order to apply Bayesian model selection. The goal is to incorporate Occam's razor as provided by Bayes analysis in order to automatically pick the model optimally able to…

Statistics Theory · Mathematics 2015-12-16 Miguel de Benito Delgado , Philipp Wacker

This article applies the principle of Occam's Razor to non-parametric model building of statistical data, by finding a model with the minimal number of bits, leading to an exceptionally effective regularization method for probability…

Machine Learning · Statistics 2020-06-18 Peter Kövesarki

We investigate the evidence/flexibility (i.e., "Occam") paradigm and demonstrate the theoretical and empirical consistency of Bayesian evidence for the task of determining an appropriate generative model for network data. This model…

Methodology · Statistics 2024-05-09 Tianyu Wang , Zachary M. Pisano , Carey E. Priebe

Linear regression models are among the models most used in practice, although the practitioners are often not sure whether their assumed linear regression model is at least approximately true. In such situations, only designs for which the…

Statistics Theory · Mathematics 2007-06-13 Wolfgang Bischoff , Frank Miller

The Rashomon Effect describes the following phenomenon: for a given dataset there may exist many models with equally good performance but with different solution strategies. The Rashomon Effect has implications for Explainable Machine…

Machine Learning · Computer Science 2023-06-30 Sebastian Müller , Vanessa Toborek , Katharina Beckh , Matthias Jakobs , Christian Bauckhage , Pascal Welke

Probabilistic models analyze data by relying on a set of assumptions. Data that exhibit deviations from these assumptions can undermine inference and prediction quality. Robust models offer protection against mismatch between a model's…

Machine Learning · Statistics 2018-06-20 Yixin Wang , Alp Kucukelbir , David M. Blei

The paper proposes a novel model assessment paradigm aiming to address shortcoming of posterior predictive $p-$values, which provide the default metric of fit for Bayesian structural equation modelling (BSEM). The model framework of the…

Methodology · Statistics 2022-06-30 Konstantinos Vamvourellis , Konstantinos Kalogeropoulos , Irini Moustaki

Prompted by misconceptions in the recent literature, we review the justifications for naturalness arguments and Occam's razor found in Bayesian statistics. We discuss the automatic Occam's razor that emerges in Bayesian formalism, bringing…

History and Philosophy of Physics · Physics 2026-04-22 Andrew Fowlie

Here we argue that the notion of falsifiability, a key concept in defining a valid scientific theory, can be quantified using Bayesian Model Selection, which is a standard tool in modern statistics. This relates falsifiability to the…

History and Philosophy of Physics · Physics 2015-06-03 Ilya Nemenman

Behavioral theories rest on parsimony: a small number of mechanisms organizing many decisions. We define a Maximum Rule Concentration Index that measures how parsimoniously a dataset of risky choices can be organized through a library of…

General Economics · Economics 2026-05-29 Avner Seror

When selecting a model from a set of equally performant models, how much unfairness can you really reduce? Is it important to be intentional about fairness when choosing among this set, or is arbitrarily choosing among the set of ''good''…

Computers and Society · Computer Science 2025-01-28 Gordon Dai , Pavan Ravishankar , Rachel Yuan , Daniel B. Neill , Emily Black

Statistical learning theory is often associated with the principle of Occam's razor, which recommends a simplicity preference in inductive inference. This paper distills the core argument for simplicity obtainable from statistical learning…

Machine Learning · Computer Science 2024-12-02 Tom F. Sterkenburg

This paper uses techniques from Random Matrix Theory to find the ideal training-testing data split for a simple linear regression with m data points, each an independent n-dimensional multivariate Gaussian. It defines "ideal" as satisfying…

Machine Learning · Statistics 2022-07-26 Alexander Dubbs

An earlier introduced characterization of nonuniform learnability that allows the sample size to depend on the hypothesis to which the learner is compared has been redefined using the measure theoretic approach. Where nonuniform…

Machine Learning · Computer Science 2020-11-03 Ankit Bandyopadhyay

We report on a series of experiments in which all decision trees consistent with the training data are constructed. These experiments were run to gain an understanding of the properties of the set of consistent decision trees and the…

Artificial Intelligence · Computer Science 2008-02-03 P. M. Murphy , M. J. Pazzani