English
Related papers

Related papers: A Cardinal Comparison of Experts

200 papers

Societies often rely on human experts to take a wide variety of decisions affecting their members, from jail-or-release decisions taken by judges and stop-and-frisk decisions taken by police officers to accept-or-reject decisions taken by…

Machine Learning · Statistics 2018-05-29 Isabel Valera , Adish Singla , Manuel Gomez Rodriguez

Forecasting is a task that is difficult to evaluate: the ground truth can only be known in the future. Recent work showing LLM forecasters rapidly approaching human-level performance begs the question: how can we benchmark and evaluate…

Machine Learning · Computer Science 2025-01-13 Daniel Paleka , Abhimanyu Pallavi Sudhir , Alejandro Alvarez , Vineeth Bhat , Adam Shen , Evan Wang , Florian Tramèr

We study the problem of prediction with expert advice with adversarial corruption where the adversary can at most corrupt one expert. Using tools from viscosity theory, we characterize the long-time behavior of the value function of the…

Machine Learning · Computer Science 2021-03-02 Erhan Bayraktar , Ibrahim Ekren , Xin Zhang

Cardinality Estimation is to estimate the size of the output of a query without computing it, by using only statistics on the input relations. Existing estimators try to return an unbiased estimate of the cardinality: this is notoriously…

Databases · Computer Science 2024-12-03 Mahmoud Abo Khamis , Kyle Deeds , Dan Olteanu , Dan Suciu

Computing reachability probabilities is a fundamental problem in the analysis of probabilistic programs. This paper aims at a comprehensive and comparative account on various martingale-based methods for over- and under-approximating…

Programming Languages · Computer Science 2018-11-16 Toru Takisaka , Yuichiro Oyabu , Natsuki Urabe , Ichiro Hasuo

Survival analysis deals with modeling the time until an event occurs, and accurate probability estimates are crucial for decision-making, particularly in the competing-risks setting where multiple events are possible. While recent work has…

Methodology · Statistics 2026-02-03 Julie Alberge , Tristan Haugomat , Gaël Varoquaux , Judith Abécassis

Comparison data arises in many important contexts, e.g. shopping, web clicks, or sports competitions. Typically we are given a dataset of comparisons and wish to train a model to make predictions about the outcome of unseen comparisons. In…

Machine Learning · Statistics 2018-07-25 Stephen Ragain , Alexander Peysakhovich , Johan Ugander

We investigate whether experts possess differential expertise when making predictions. We note that this would make it possible to aggregate multiple predictions into a result that is more accurate than their consensus average, and that the…

Computers and Society · Computer Science 2018-06-19 Amir Ban , Yishay Mansour

Providing explanations about how machine learning algorithms work and/or make particular predictions is one of the main tools that can be used to improve their trusworthiness, fairness and robustness. Among the most intuitive type of…

Machine Learning · Computer Science 2024-04-12 Rubén Ruiz-Torrubiano

Operational earthquake forecasting for risk management and communication during seismic sequences depends on our ability to select an optimal forecasting model. To do this, we need to compare the performance of competing models with each…

Applications · Statistics 2022-04-20 Francesco Serafini , Mark Naylor , Finn Lindgren , Maximilian Werner , Ian Main

Proposed is a new formal approach for solution of extreme multi-criteria problems transforming them into single-criterion mathematical models, without any additional information. Transforming rules are based on comparison standards and…

Optimization and Control · Mathematics 2007-05-23 V. O. Groppen

We prove the sharp bound for the probability that two experts who have access to different information, represented by different $\sigma$-fields, will give radically different estimates of the probability of an event. This is relevant when…

Probability · Mathematics 2019-12-03 Krzysztof Burdzy , Soumik Pal

This text is a survey on cross-validation. We define all classical cross-validation procedures, and we study their properties for two different goals: estimating the risk of a given estimator, and selecting the best estimator among a given…

Statistics Theory · Mathematics 2017-03-10 Sylvain Arlot

Comparing the top $k$ elements between two or more ranked results is a common task in many contexts and settings. A few measures have been proposed to compare top $k$ lists with attractive mathematical properties, but they face a number of…

Information Theory · Computer Science 2013-10-02 Arun Konagurthu , James Collier

Predicting the answer to a product-related question is an emerging field of research that recently attracted a lot of attention. Answering subjective and opinion-based questions is most challenging due to the dependency on…

Computation and Language · Computer Science 2021-05-20 Ohad Rozen , David Carmel , Avihai Mejer , Vitaly Mirkis , Yftah Ziser

This paper provides a statistical method to test whether a system that performs a binary sequential hypothesis test is optimal in the sense of minimizing the average decision times while taking decisions with given reliabilities. The…

Information Theory · Computer Science 2018-01-08 Meik Dörpinghaus , Izaak Neri , Édgar Roldán , Heinrich Meyr , Frank Jülicher

We study methods to obtain the consistency of forcing axioms, and particularly higher forcing axioms. We first force over a model with a supercompact cardinal $\theta>\kappa$ to get the consistency of the forcing axiom for $\kappa$-strongly…

Logic · Mathematics 2024-03-19 David Asperó , Sean Cox , Asaf Karagila , Christoph Weiss

Aligning large language models with expert judgment is especially difficult in subjective evaluation tasks, where experts may disagree, rely on tacit criteria, and change their judgments over time. In this paper, we study expert alignment…

Computation and Language · Computer Science 2026-05-07 Tzu-Mi Lin , Wataru Hirota , Tatsuya Ishigaki , Lung-Hao Lee , Chung-Chi Chen

Decision makers often need to rely on imperfect probabilistic forecasts. While average performance metrics are typically available, it is difficult to assess the quality of individual forecasts and the corresponding utilities. To convey…

Machine Learning · Statistics 2021-03-03 Shengjia Zhao , Stefano Ermon

We study committee elections from a perspective of finding the most conflicting candidates, that is, candidates that imply the largest amount of conflict, as per voter preferences. By proposing basic axioms to capture this objective, we…

Computer Science and Game Theory · Computer Science 2024-05-10 Théo Delemazure , Łukasz Janeczko , Andrzej Kaczmarczyk , Stanisław Szufa
‹ Prev 1 3 4 5 6 7 10 Next ›