English
Related papers

Related papers: Revisiting Rashomon: A Comment on "The Two Culture…

200 papers

Two decades ago, Leo Breiman identified two cultures for statistical modeling. The data modeling culture (DMC) refers to practices aiming to conduct statistical inference on one or several quantities of interest. The algorithmic modeling…

Methodology · Statistics 2020-12-09 Adel Daoud , Devdatt Dubhashi

Breiman (2001) proposed to statisticians awareness of two cultures: 1. Parametric modeling culture, pioneered by R.A.Fisher and Jerzy Neyman; 2. Algorithmic predictive culture, pioneered by machine learning research. Parzen (2001), as a…

Statistics Theory · Mathematics 2012-04-25 Emanuel Parzen , Subhadeep Mukhopadhyay

This paper considers the two-dataset problem, where data are collected from two potentially different populations sharing common aspects. This problem arises when data are collected by two different types of researchers or from two…

Methodology · Statistics 2022-09-27 Steven N. MacEachern , Koji Miyawaki

The machine learning modeling process conventionally culminates in selecting a single model that maximizes a selected performance metric. However, this approach leads to abandoning a more profound analysis of slightly inferior models.…

Machine Learning · Computer Science 2024-10-28 Katarzyna Kobylińska , Mateusz Krzyziński , Rafał Machowicz , Mariusz Adamek , Przemysław Biecek

It is almost always easier to find an accurate-but-complex model than an accurate-yet-simple model. Finding optimal, sparse, accurate models of various forms (linear models with integer coefficients, decision sets, rule lists, decision…

Machine Learning · Computer Science 2022-05-16 Lesia Semenova , Cynthia Rudin , Ronald Parr

When selecting a model from a set of equally performant models, how much unfairness can you really reduce? Is it important to be intentional about fairness when choosing among this set, or is arbitrarily choosing among the set of ''good''…

Computers and Society · Computer Science 2025-01-28 Gordon Dai , Pavan Ravishankar , Rachel Yuan , Daniel B. Neill , Emily Black

The existence of multiple, equally accurate models for a given predictive task leads to predictive multiplicity, where a ``Rashomon set'' of models achieve similar accuracy but diverges in their individual predictions. This inconsistency…

Machine Learning · Computer Science 2026-05-19 Parian Haghighat , Hadis Anahideh , Cynthia Rudin

Real-world machine learning (ML) pipelines rarely produce a single model; instead, they produce a Rashomon set of many near-optimal ones. We show that this multiplicity reshapes key aspects of trustworthiness. At the individual-model level,…

Machine Learning · Computer Science 2025-12-01 Ethan Hsu , Harry Chen , Chudi Zhong , Lesia Semenova

The usual goal of supervised learning is to find the best model, the one that optimizes a particular performance measure. However, what if the explanation provided by this model is completely different from another model and different again…

Machine Learning · Statistics 2024-09-11 Przemyslaw Biecek , Hubert Baniecki , Mateusz Krzyzinski , Dianne Cook

The Rash\=omon effect poses challenges for deriving reliable knowledge from machine learning models. This study examined the influence of sample size on explanations from models in a Rash\=omon set using SHAP. Experiments on 5 public…

Machine Learning · Computer Science 2023-08-15 Clement Poiret , Antoine Grigis , Justin Thomas , Marion Noulhiane

Different prediction models might perform equally well (Rashomon set) in the same task, but offer conflicting interpretations and conclusions about the data. The Rashomon effect in the context of Explainable AI (XAI) has been recognized as…

Machine Learning · Computer Science 2024-07-29 Sichao Li , Amanda S. Barnard , Quanling Deng

The growing need for in-depth analysis of predictive models leads to a series of new methods for explaining their local and global properties. Which of these methods is the best? It turns out that this is an ill-posed question. One cannot…

Machine Learning · Computer Science 2024-09-09 Hubert Baniecki , Dariusz Parzych , Przemyslaw Biecek

Predictive multiplicity occurs when classification models with statistically indistinguishable performances assign conflicting predictions to individual samples. When used for decision-making in applications of consequence (e.g., lending,…

Machine Learning · Computer Science 2022-10-21 Hsiang Hsu , Flavio du Pin Calmon

Econometrics and machine learning seem to have one common goal: to construct a predictive model, for a variable of interest, using explanatory variables (or features). However, these two fields developed in parallel, thus creating two…

Other Statistics · Statistics 2020-06-26 Arthur Charpentier , Emmanuel Flachaire , Antoine Ly

This paper proposes a rigorous framework to examine the two-way relationship between artificial intelligence (AI), human cognition, problem-solving, and cultural adaptation across academic and business settings. It addresses a key gap by…

Human-Computer Interaction · Computer Science 2025-10-14 Matthias Huemmer , Theophile Shyiramunda , Michelle J. Cummings-Koether

In any given machine learning problem, there may be many models that could explain the data almost equally well. However, most learning algorithms return only one of these models, leaving practitioners with no practical way to explore…

Machine Learning · Computer Science 2022-10-27 Rui Xin , Chudi Zhong , Zhi Chen , Takuya Takagi , Margo Seltzer , Cynthia Rudin

The widespread acceptance of empirically derived codal provisions and equations in civil engineering stands in stark contrast to the skepticism facing machine learning (ML) models, despite their shared statistical foundations. This paper…

Machine Learning · Computer Science 2025-01-10 MZ Naser

Impact of academic research onto the non-academic world is of increasing importance as authorities seek return on public investment. Impact opens new opportunities for what are known as "professional services": as scientometrical tools…

History and Philosophy of Physics · Physics 2020-05-26 R. Kenna

Statistical wisdom suggests that very complex models, interpolating training data, will be poor at predicting unseen examples.Yet, this aphorism has been recently challenged by the identification of benign overfitting regimes, specially…

Statistics Theory · Mathematics 2023-02-10 Ludovic Arnould , Claire Boyer , Erwan Scornet

Additive two-tower models are popular learning-to-rank methods for handling biased user feedback in industry settings. Recent studies, however, report a concerning phenomenon: training two-tower models on clicks collected by well-performing…

Information Retrieval · Computer Science 2025-09-01 Philipp Hager , Onno Zoeter , Maarten de Rijke