English
Related papers

Related papers: Cross-calibration of probabilistic forecasts

200 papers

The wide-spread adoption of representation learning technologies in clinical decision making strongly emphasizes the need for characterizing model reliability and enabling rigorous introspection of model behavior. While the former need is…

Machine Learning · Computer Science 2020-05-01 Jayaraman J. Thiagarajan , Prasanna Sattigeri , Deepta Rajan , Bindya Venkatesh

We present a scheme by which a probabilistic forecasting system whose predictions have poor probabilistic calibration may be recalibrated by incorporating past performance information to produce a new forecasting system that is demonstrably…

Methodology · Statistics 2019-04-08 Carlo Graziani , Robert Rosner , Jennifer M. Adams , Reason L. Machete

We show how to achieve the notion of "multicalibration" from H\'ebert-Johnson et al. [2018] not just for means, but also for variances and other higher moments. Informally, it means that we can find regression functions which, given a data…

Machine Learning · Computer Science 2020-08-19 Christopher Jung , Changhwa Lee , Mallesh M. Pai , Aaron Roth , Rakesh Vohra

Applications such as weather forecasting and personalized medicine demand models that output calibrated probability estimates---those representative of the true likelihood of a prediction. Most models are not calibrated out of the box but…

Machine Learning · Computer Science 2020-02-03 Ananya Kumar , Percy Liang , Tengyu Ma

Selective classification allows models to abstain from making predictions (e.g., say "I don't know") when in doubt in order to obtain better effective accuracy. While typical selective models can be effective at producing more accurate…

Machine Learning · Computer Science 2024-06-24 Adam Fisch , Tommi Jaakkola , Regina Barzilay

Reliable uncertainty estimates are an important tool for helping autonomous agents or human decision makers understand and leverage predictive models. However, existing approaches to estimating uncertainty largely ignore the possibility of…

Machine Learning · Computer Science 2020-05-22 Sangdon Park , Osbert Bastani , James Weimer , Insup Lee

We consider the problem of model multiplicity in downstream decision-making, a setting where two predictive models of equivalent accuracy cannot agree on the best-response action for a downstream loss function. We show that even when the…

Machine Learning · Computer Science 2024-05-31 Ally Yalei Du , Dung Daniel Ngo , Zhiwei Steven Wu

Post-hoc calibration methods are widely used to improve the reliability of probabilistic predictions from machine learning models. Despite their prevalence, a comprehensive theoretical understanding of these methods remains elusive,…

Machine Learning · Computer Science 2025-09-30 Kristina P. Sinaga , Arjun S. Nair

Probabilistic predictions from neural networks which account for predictive uncertainty during classification is crucial in many real-world and high-impact decision making settings. However, in practice most datasets are trained on…

Machine Learning · Computer Science 2022-09-30 Satya Borgohain , Klaus Ackermann , Ruben Loaiza-Maya

Calibrating blackbox machine learning models to achieve risk control is crucial to ensure reliable decision-making. A rich line of literature has been studying how to calibrate a model so that its predictions satisfy explicit finite-sample…

Machine Learning · Statistics 2025-06-02 Victor Li , Baiting Chen , Yuzhen Mao , Qi Lei , Zhun Deng

Forecasting has always been at the forefront of decision making and planning. The uncertainty that surrounds the future is both exciting and challenging, with individuals and organisations seeking to minimise risks and maximise utilities.…

Applications · Statistics 2022-02-09 Fotios Petropoulos , Daniele Apiletti , Vassilios Assimakopoulos , Mohamed Zied Babai , Devon K. Barrow , Souhaib Ben Taieb , Christoph Bergmeir , Ricardo J. Bessa , Jakub Bijak , John E. Boylan , Jethro Browell , Claudio Carnevale , Jennifer L. Castle , Pasquale Cirillo , Michael P. Clements , Clara Cordeiro , Fernando Luiz Cyrino Oliveira , Shari De Baets , Alexander Dokumentov , Joanne Ellison , Piotr Fiszeder , Philip Hans Franses , David T. Frazier , Michael Gilliland , M. Sinan Gönül , Paul Goodwin , Luigi Grossi , Yael Grushka-Cockayne , Mariangela Guidolin , Massimo Guidolin , Ulrich Gunter , Xiaojia Guo , Renato Guseo , Nigel Harvey , David F. Hendry , Ross Hollyman , Tim Januschowski , Jooyoung Jeon , Victor Richmond R. Jose , Yanfei Kang , Anne B. Koehler , Stephan Kolassa , Nikolaos Kourentzes , Sonia Leva , Feng Li , Konstantia Litsiou , Spyros Makridakis , Gael M. Martin , Andrew B. Martinez , Sheik Meeran , Theodore Modis , Konstantinos Nikolopoulos , Dilek Önkal , Alessia Paccagnini , Anastasios Panagiotelis , Ioannis Panapakidis , Jose M. Pavía , Manuela Pedio , Diego J. Pedregal , Pierre Pinson , Patrícia Ramos , David E. Rapach , J. James Reade , Bahman Rostami-Tabar , Michał Rubaszek , Georgios Sermpinis , Han Lin Shang , Evangelos Spiliotis , Aris A. Syntetos , Priyanga Dilini Talagala , Thiyanga S. Talagala , Len Tashman , Dimitrios Thomakos , Thordis Thorarinsdottir , Ezio Todini , Juan Ramón Trapero Arenas , Xiaoqian Wang , Robert L. Winkler , Alisa Yusupova , Florian Ziel

Calibrated probability outputs of trained classifiers are increasingly used as inputs to downstream regression estimands such as effects, prevalences, or disparities for a latent group observed only on a small labelled subset. A standard…

Methodology · Statistics 2026-05-14 Marcell T. Kurbucz

Probabilistic forecasting estimates the likelihood of uncertain future events. To improve LLM forecasting, existing methods typically learn from binary outcomes to output verbalized forecasts. However, while aggregated human forecasts…

Machine Learning · Computer Science 2026-05-28 Hui Dai , Ryan Teehan , Parsa Torabian , Mengye Ren

Calibration is a critical property for establishing the trustworthiness of predictors that provide uncertainty estimates. Multicalibration is a strengthening of calibration which requires that predictors be calibrated on a potentially…

Machine Learning · Computer Science 2025-09-23 Nathan Derhake , Siddartha Devic , Dutch Hansen , Kuan Liu , Vatsal Sharan

Probability predictions from binary regressions or machine learning methods ought to be calibrated: If an event is predicted to occur with probability $x$, it should materialize with approximately that frequency, which means that the…

Statistics Theory · Mathematics 2023-01-11 Timo Dimitriadis , Lutz Duembgen , Alexander Henzi , Marius Puke , Johanna Ziegel

We provide another look at the statistical calibration problem in computer models. This viewpoint is inspired by two overarching practical considerations of computer models: (i) many computer models are inadequate for perfectly modeling…

Methodology · Statistics 2018-09-26 Xiaowu Dai , Peter Chien

Forecast systems in science and technology are increasingly moving beyond point prediction toward methods that produce full predictive distributions of future outcomes y, conditional on high-dimensional and complex sequences of inputs x.…

Machine Learning · Statistics 2026-03-13 Elizabeth Cucuzzella , Rafael Izbicki , Ann B. Lee

Long-range ensemble forecasts are typically verified as anomalies with respect to a lead-time dependent climatological mean to remove the influence of systematic biases. However, common methods for calculating anomalies result in…

Atmospheric and Oceanic Physics · Physics 2025-06-11 Christopher D. Roberts , Martin Leutbecher

The machine learning community has become increasingly concerned with the potential for bias and discrimination in predictive models. This has motivated a growing line of work on what it means for a classification procedure to be "fair." In…

Machine Learning · Computer Science 2017-11-07 Geoff Pleiss , Manish Raghavan , Felix Wu , Jon Kleinberg , Kilian Q. Weinberger

Reliable confidence estimation for the predictions is important in many safety-critical applications. However, modern deep neural networks are often overconfident for their incorrect predictions. Recently, many calibration methods have been…

Machine Learning · Computer Science 2023-03-07 Fei Zhu , Zhen Cheng , Xu-Yao Zhang , Cheng-Lin Liu