English
Related papers

Related papers: Aggregating Elo Ratings: An Axiomatization

200 papers

This paper presents an original model for assessing scientific productivity, research power ranking, RPR, which is based on the adaptation of the Elo rating system to the context of scientific activity. Unlike traditional scientometric…

Physics and Society · Physics 2025-04-30 Eldar Knar

Arena-based evaluation is a fundamental yet significant evaluation paradigm for modern AI models, especially large language models (LLMs). Existing framework based on ELO rating system suffers from the inevitable instability problem due to…

Artificial Intelligence · Computer Science 2025-05-30 Zirui Liu , Jiatong Li , Yan Zhuang , Qi Liu , Shuanghong Shen , Jie Ouyang , Mingyue Cheng , Shijin Wang

This paper investigates the evaluation of learned multiagent strategies in the incomplete information setting, which plays a critical role in ranking and training of agents. Traditionally, researchers have relied on Elo ratings for this…

Multiagent Systems · Computer Science 2020-01-13 Mark Rowland , Shayegan Omidshafiei , Karl Tuyls , Julien Perolat , Michal Valko , Georgios Piliouras , Remi Munos

We argue that many general evaluation problems can be viewed through the lens of voting theory. Each task is interpreted as a separate voter, which requires only ordinal rankings or pairwise comparisons of agents to produce an overall…

Artificial Intelligence · Computer Science 2025-07-01 Marc Lanctot , Kate Larson , Yoram Bachrach , Luke Marris , Zun Li , Avishkar Bhoopchand , Thomas Anthony , Brian Tanner , Anna Koop

The Elo rating system has been recognised as an effective method for modelling students and items within adaptive educational systems. The existing Elo-based models have the limiting assumption that items are only tagged with a single…

Computers and Society · Computer Science 2019-10-29 Solmaz Abdi , Hassan Khosravi , Shazia Sadiq , Dragan Gasevic

Reasoning about agent preferences on a set of alternatives, and the aggregation of such preferences into some social ranking is a fundamental issue in reasoning about uncertainty and multi-agent systems. When the set of agents and the set…

Computer Science and Game Theory · Computer Science 2012-07-19 Moshe Tennenholtz

In many real life situations, including job and loan applications, gatekeepers must make justified and fair real-time decisions about a person's fitness for a particular opportunity. In this paper, we aim to accomplish approximate group…

Machine Learning · Computer Science 2021-05-26 Yi Sun , Ivan Ramirez , Alfredo Cuesta-Infante , Kalyan Veeramachaneni

Rating aggregation plays a crucial role in various fields, such as product recommendations, hotel rankings, and teaching evaluations. However, traditional averaging methods can be affected by participation bias, where some raters do not…

Machine Learning · Computer Science 2025-02-07 Yongkang Guo , Yuqing Kong , Jialiang Liu

To aggregate rankings into a social ranking, one can use scoring systems such as Plurality, Veto, and Borda. We distinguish three types of methods: ranking by score, ranking by repeatedly choosing a winner that we delete and rank at the…

Computer Science and Game Theory · Computer Science 2022-09-20 Niclas Boehmer , Robert Bredereck , Dominik Peters

It is common that a jury must grade a set of candidates in a cardinal scale such as {1,2,3,4,5} or an ordinal scale such as {Great, Good, Average, Bad }. When the number of candidates is very large such as hotels (BOOKING), restaurants…

Computer Science and Game Theory · Computer Science 2023-02-24 Rida Laraki , Estelle Varloot

Benchmarking is a fundamental practice in machine learning (ML) for comparing the performance of classification algorithms. However, traditional evaluation methods often overlook a critical aspect: the joint consideration of dataset…

Machine Learning · Computer Science 2025-04-15 Lucas Cardoso , Vitor Santos , José Ribeiro , Regiane Kawasaki , Ricardo Prudêncio , Ronnie Alves

Sequential estimators are proposed for the relative risk, odds ratio, log relative risk or log odds ratio of a dichotomous attribute in two populations. The estimators take the same number of observations from each population, and guarantee…

Methodology · Statistics 2026-04-07 Luis Mendo

Equitability is a fundamental notion in fair division which requires that all agents derive equal value from their allocated bundles. We study, for general (possibly non-monotone) valuations, a popular relaxation of equitability known as…

Computer Science and Game Theory · Computer Science 2025-11-11 Hadi Hosseini , Vishwa Prakash HV , Aditi Sethia , Jatin Yadav

This paper develops an axiomatic framework for ranking metrics, a general class of functionals for evaluating and ordering financial or insurance positions. Unlike traditional risk-adjusted performance measures-such as the Sharpe ratio,…

Risk Management · Quantitative Finance 2026-04-21 Asmerilda Hitaj , Elisa Mastrogiacomo , Ilaria Peri , Marcelo Righi

Competitor rating systems for head-to-head games are typically used to measure playing strength from game outcomes. Ratings computed from these systems are often used to select top competitors for elite events, for pairing players of…

Methodology · Statistics 2025-07-14 Mark E. Glickman

In this work, we deal with the problem of rating in sports, where the skills of the players/teams are inferred from the observed outcomes of the games. Our focus is on the online rating algorithms which estimate the skills after each new…

Machine Learning · Statistics 2021-04-30 Leszek Szczecinski , Raphaëlle Tihon

We study efficient, linear, and symmetric (ELS) values, a central family of allocation rules for cooperative games with transferable-utility (TU-games) that includes the Shapley value, the CIS value, and the ENSC value. We first show that…

Theoretical Economics · Economics 2026-05-06 Yukihiko Funaki , Yukio Koriyama , Satoshi Nakada , Yuki Tamura

Accurate estimation of question difficulty and prediction of student performance play key roles in optimizing educational instruction and enhancing learning outcomes within digital learning platforms. The Elo rating system is widely…

Computers and Society · Computer Science 2024-03-14 Erva Nihan Kandemir , Jill-Jenn Vie , Adam Sanchez-Ayte , Olivier Palombi , Franck Ramus

Approval-based committee (ABC) voting rules elect a fixed size subset of the candidates, a so-called committee, based on the voters' approval ballots over the candidates. While these rules have recently attracted significant attention,…

Computer Science and Game Theory · Computer Science 2023-02-24 Chris Dong , Patrick Lederer

Variable selection for models including interactions between explanatory variables often needs to obey certain hierarchical constraints. The weak or strong structural hierarchy requires that the existence of an interaction term implies at…

Statistics Theory · Mathematics 2016-11-10 Yiyuan She , Zhifeng Wang , He Jiang