English
Related papers

Related papers: Evaluating the Discrimination Ability of Proper Mu…

200 papers

We demonstrate that, for a range of state-of-the-art machine learning algorithms, the differences in generalisation performance obtained using default parameter settings and using parameters tuned via cross-validation can be similar in…

Machine Learning · Computer Science 2017-03-21 Anthony Bagnall , Gavin C. Cawley

Diffusion models are typically trained using score matching, a learning objective agnostic to the underlying noising process that guides the model. This paper argues that Markov noising processes enjoy an advantage over alternatives, as the…

Machine Learning · Statistics 2025-05-27 Zheyang Shen , Huihui Wang , Marina Riabiz , Chris J. Oates

Statistical analysis of extremes can be used to predict the probability of future extreme events, such as large rainfalls or devastating windstorms. The quality of these forecasts can be measured through scoring rules. Locally scale…

Methodology · Statistics 2024-02-22 Helga Kristin Olafsdottir , Holger Rootzén , David Bolin

Proper scoring rules elicit truth-telling when making predictions, or otherwise revealing information. However, when multiple predictions are made of the same event, telling the truth is in general no longer optimal, as agents are motivated…

Computer Science and Game Theory · Computer Science 2017-07-04 Amir Ban

We investigate differences between a simple Dominance Principle applied to sums of fair prices for variables and dominance applied to sums of forecasts for variables scored by proper scoring rules. In particular, we consider differences…

Statistics Theory · Mathematics 2014-05-27 M. J. Schervish , Teddy Seidenfeld , J. B. Kadane

A voting rule decides on a probability distribution over a set of m alternatives, based on rankings of those alternatives provided by agents. We assume that agents have cardinal utility functions over the alternatives, but voting rules have…

Computer Science and Game Theory · Computer Science 2024-01-23 Soroush Ebadian , Anson Kahng , Dominik Peters , Nisarg Shah

We investigate the properties of two standard energy estimators used in path-integral Monte Carlo simulations. By disentangling the variance of the estimators and their autocorrelation times we analyse the dependence of the performance on…

Condensed Matter · Physics 2009-10-30 Wolfhard Janke , Tilman Sauer

Committee-selection problems arise in many contexts and applications, and there has been increasing interest within the social choice research community on identifying which properties are satisfied by different multi-winner voting rules.…

Artificial Intelligence · Computer Science 2025-08-11 Joshua Caiata , Ben Armstrong , Kate Larson

The increasing demand for personalized health care has led to the expectation that individualized quantitative evaluation of human disease states is possible. However, this has not yet been achieved at a sufficiently low cost. Our ultimate…

Applications · Statistics 2022-06-17 Rina Kagawa , Masanori Shiro

Compositional data and multivariate count data with known totals are challenging to analyse due to the non-negativity and sum-to-one constraints on the sample space. It is often the case that many of the compositional components are highly…

Methodology · Statistics 2020-12-24 Janice L. Scealy , Andrew T. A. Wood

Recent and influential critiques of standardized testing have noted the existence of non-trivial numbers of successful scientists who received low scores on the GRE. Here we use computer simulations to show that the prevalence of such…

Physics Education · Physics 2017-09-12 Alexander R. Small

Humans are accustomed to environments that contain both regularities and exceptions. For example, at most gas stations, one pays prior to pumping, but the occasional rural station does not accept payment in advance. Likewise, deep neural…

Machine Learning · Computer Science 2021-06-16 Ziheng Jiang , Chiyuan Zhang , Kunal Talwar , Michael C. Mozer

Energy price forecasting is a relevant yet hard task in the field of multi-step time series forecasting. In this paper we compare a well-known and established method, ARMA with exogenous variables with a relatively new technique Gradient…

Machine Learning · Statistics 2015-06-24 Gergo Barta , Gyula Borbely , Gabor Nagy , Sandor Kazi , Tamas Henk

Rare properties remain a challenge for statistical model checking (SMC) due to the quadratic scaling of variance with rarity. We address this with a variance reduction framework based on lightweight importance splitting observers. These…

Logic in Computer Science · Computer Science 2015-04-29 Cyrille Jegourel , Axel Legay , Sean Sedwards , Louis-Marie Traonouez

A multi-label classifier estimates the binary label state (relevant vs irrelevant) for each of a set of concept labels, for any given instance. Probabilistic multi-label classifiers provide a predictive posterior distribution over all…

Machine Learning · Computer Science 2022-09-12 Laurence A. F. Park , Jesse Read

We address the problem of uncertainty quantification and propose measures of total, aleatoric, and epistemic uncertainty based on a known decomposition of (strictly) proper scoring rules, a specific type of loss function, into a divergence…

Machine Learning · Computer Science 2025-05-29 Paul Hofman , Yusuf Sale , Eyke Hüllermeier

Prior-weighted logistic regression has become a standard tool for calibration in speaker recognition. Logistic regression is the optimization of the expected value of the logarithmic scoring rule. We generalize this via a parametric family…

Machine Learning · Statistics 2013-07-31 Niko Brümmer , George Doddington

Neural Posterior Estimation methods for simulation-based inference can be ill-suited for dealing with posterior distributions obtained by conditioning on multiple observations, as they tend to require a large number of simulator calls to…

Machine Learning · Computer Science 2023-07-11 Tomas Geffner , George Papamakarios , Andriy Mnih

The estimated accuracy of a classifier is a random quantity with variability. A common practice in supervised machine learning, is thus to test if the estimated accuracy is significantly better than chance level. This method of signal…

Methodology · Statistics 2020-01-28 Jonathan D. Rosenblatt , Yuval Benjamini , Roee Gilron , Roy Mukamel , Jelle J. Goeman

Diffusion models offer a robust framework for sampling from unnormalized probability densities, which requires accurately estimating the score of the noise-perturbed target distribution. While the standard Denoising Score Identity (DSI)…

Machine Learning · Computer Science 2025-12-24 Khaled Kahouli , Romuald Elie , Klaus-Robert Müller , Quentin Berthet , Oliver T. Unke , Arnaud Doucet