English
Related papers

Related papers: On the Statistical Differences between Binary Fore…

200 papers

Probabilistic forecasts are becoming more and more available. How should they be used and communicated? What are the obstacles to their use in practice? I review experience with five problems where probabilistic forecasting played an…

Applications · Statistics 2014-08-22 Adrian E. Raftery

Binary classification models which can assign probabilities to categories such as "the tissue is 75% likely to be tumorous" or "the chemical is 25% likely to be toxic" are well understood statistically, but their utility as an input to…

Applications · Statistics 2017-12-05 Damjan Krstajic , Ljubomir Buturovic , Simon Thomas , David E Leahy

We revisit the foundations of fairness and its interplay with utility and efficiency in settings where the training data contain richer labels, such as individual types, rankings, or risk estimates, rather than just binary outcomes. In this…

Machine Learning · Computer Science 2025-05-23 Noga Amit , Omer Reingold , Guy N. Rothblum

We consider the weakly supervised binary classification problem where the labels are randomly flipped with probability $1- {\alpha}$. Although there exist numerous algorithms for this problem, it remains theoretically unexplored how the…

Machine Learning · Computer Science 2019-07-16 Xinyang Yi , Zhaoran Wang , Zhuoran Yang , Constantine Caramanis , Han Liu

Deep Bayesian neural networks (BNNs) are a powerful tool, though computationally demanding, to perform parameter estimation while jointly estimating uncertainty around predictions. BNNs are typically implemented using arbitrary…

Machine Learning · Computer Science 2020-05-12 Daniele Silvestro , Tobias Andermann

In this work, we study the effects of feature-based explanations on distributive fairness of AI-assisted decisions, specifically focusing on the task of predicting occupations from short textual bios. We also investigate how any effects are…

Human-Computer Interaction · Computer Science 2024-03-20 Jakob Schoeffer , Maria De-Arteaga , Niklas Kuehl

Performative predictions are forecasts which influence the outcomes they aim to predict, undermining the existence of correct forecasts and standard methods of elicitation and estimation. We show that conditioning forecasts on covariates…

Statistics Theory · Mathematics 2025-10-27 Philip Boeken , Onno Zoeter , Joris M. Mooij

We study the asymptotic behaviour of widely used tests for evaluating and comparing predictive accuracy when forecast errors exhibit heavy tails. In particular, when loss differentials have infinite variance, the Diebold-Mariano test…

Methodology · Statistics 2026-05-20 Jonas F. Frederiksen , Muneya Matsui , Rasmus S. Pedersen

How reliably can we trust the scores obtained from social bias benchmarks as faithful indicators of problematic social biases in a given language model? In this work, we study this question by contrasting social biases with non-social…

Computation and Language · Computer Science 2023-06-21 Nikil Roashan Selvam , Sunipa Dev , Daniel Khashabi , Tushar Khot , Kai-Wei Chang

This paper discusses the use of fat-tailed distributions in catastrophe prediction as opposed to the more common use of the Normal Distribution.

Other Computer Science · Computer Science 2011-11-09 Louis Mello

As machine learning (ML) models are increasingly being employed to assist human decision makers, it becomes critical to provide these decision makers with relevant inputs which can help them decide if and how to incorporate model…

Machine Learning · Computer Science 2023-06-14 Sean McGrath , Parth Mehta , Alexandra Zytek , Isaac Lage , Himabindu Lakkaraju

In many real-world applications of machine learning such as recommendations, hiring, and lending, deployed models influence the data they are trained on, leading to feedback loops between predictions and data distribution. The performative…

Machine Learning · Computer Science 2025-11-18 Kun Jin , Tian Xie , Yang Liu , Xueru Zhang

Although living organisms are affected by many interrelated and unidentified variables, this complexity does not automatically impose a fundamental limitation on statistical inference. Nor need one invoke such complexity as an explanation…

Data Analysis, Statistics and Probability · Physics 2013-02-06 Drew M. Thomas

With the rise of increasingly powerful and user-facing NLP systems, there is growing interest in assessing whether they have a good representation of uncertainty by evaluating the quality of their predictive distribution over outcomes. We…

Computation and Language · Computer Science 2024-02-27 Joris Baan , Raquel Fernández , Barbara Plank , Wilker Aziz

The vast majority of statistical theory on binary classification characterizes performance in terms of accuracy. However, accuracy is known in many cases to poorly reflect the practical consequences of classification error, most famously in…

Statistics Theory · Mathematics 2022-09-27 Shashank Singh , Justin Khim

Reliable probability estimation is of crucial importance in many real-world applications where there is inherent (aleatoric) uncertainty. Probability-estimation models are trained on observed outcomes (e.g. whether it has rained or not, or…

Mitigating bias in training on biased datasets is an important open problem. Several techniques have been proposed, however the typical evaluation regime is very limited, considering very narrow data conditions. For instance, the effect of…

Machine Learning · Computer Science 2022-10-18 Xudong Han , Aili Shen , Trevor Cohn , Timothy Baldwin , Lea Frermann

In forecasting competitions, the traditional mechanism scores the predictions of each contestant against the outcome of each event, and the contestant with the highest total score wins. While it is well-known that this traditional mechanism…

Machine Learning · Computer Science 2024-10-14 Mary Monroe , Anish Thilagar , Melody Hsu , Rafael Frongillo

A decision must often be made between heavy-tailed and Gaussian errors for a regression or a time series model, and the t-distribution is frequently used when it is assumed that the errors are heavy-tailed distributed. The performance of…

Computation · Statistics 2015-05-11 J. Martin van Zyl

We respond to Tetlock et al. (2022) showing 1) how expert judgment fails to reflect tail risk, 2) the lack of compatibility between forecasting tournaments and tail risk assessment methods (such as extreme value theory). More importantly,…

Risk Management · Quantitative Finance 2023-01-27 Nassim Nicholas Taleb , Ron Richman , Marcos Carreira , James Sharpe
‹ Prev 1 3 4 5 6 7 10 Next ›