English
Related papers

Related papers: Measuring intergroup agreement and disagreement

200 papers

Consistency, defined as the requirement that a series of measurements of the same project carried out by different raters using the same method should produce similar results, is one of the most important aspects to be taken into account in…

Software Engineering · Computer Science 2007-05-23 R. Asensio Monge , F. Sanchis Marco , F. Torre Cervigon

Majority voting and averaging are common approaches employed to resolve annotator disagreements and derive single ground truth labels from multiple annotations. However, annotators may systematically disagree with one another, often…

Computation and Language · Computer Science 2021-10-13 Aida Mostafazadeh Davani , Mark Díaz , Vinodkumar Prabhakaran

Discrimination via algorithmic decision making has received considerable attention. Prior work largely focuses on defining conditions for fairness, but does not define satisfactory measures of algorithmic unfairness. In this paper, we focus…

Calibration is a popular framework to evaluate whether a classifier knows when it does not know - i.e., its predictive probabilities are a good indication of how likely a prediction is to be correct. Correctness is commonly estimated…

Computation and Language · Computer Science 2022-12-01 Joris Baan , Wilker Aziz , Barbara Plank , Raquel Fernández

Peer assessment is an efficient and effective learning assessment method that has been used widely in diverse fields in higher education. Despite its many benefits, a fundamental problem in peer assessment is that participants lack the…

Computers and Society · Computer Science 2015-06-19 Yanqing Wang , Yaowen Liang , Luning Liu , Ying Liu

Progress in NLP is increasingly measured through benchmarks; hence, contextualizing progress requires understanding when and why practitioners may disagree about the validity of benchmarks. We develop a taxonomy of disagreement, drawing on…

Computation and Language · Computer Science 2023-05-22 Arjun Subramonian , Xingdi Yuan , Hal Daumé , Su Lin Blodgett

We study a model of consensus decision making, in which a finite group of Bayesian agents has to choose between one of two courses of action. Each member of the group has a private and independent signal at his or her disposal, giving some…

Statistics Theory · Mathematics 2018-04-24 Elchanan Mossel , Omer Tamuz

In this paper, we explore how we should aggregate the degrees of belief of of a group of agents to give a single coherent set of degrees of belief, when at least some of those agents might be probabilistically incoherent. There are a number…

Other Statistics · Statistics 2017-09-14 Richard Pettigrew

Instead of studying the properties of social relationship from an objective view, in this paper, we focus on individuals' subjective and asymmetric opinions on their interrelationships. Inspired by the theories from sociolinguistics, we…

Social and Information Networks · Computer Science 2016-11-16 Bo Wang , Yanshu Yu , Yuan Wang

In modern interconnected societies, opinions and beliefs can quickly spread across large populations, giving rise to collective behaviors such as the adoption of social norms or polarization. These phenomena have motivated many models aimed…

Physics and Society · Physics 2026-05-27 Cosimo Agostinelli , Marco Mancastroppa , Alain Barrat

A survey can be represented by a bipartite network as it has two types of nodes, participants and items in which participants can only interact with items. We introduce an agreement threshold to take a minimal projection of the participants…

Social and Information Networks · Computer Science 2020-12-22 Pádraig MacCarron , Paul J. Maher , Michael Quayle

A measure of interrater absolute agreement for ordinal scales is proposed capitalizing on the dispersion index for ordinal variables proposed by Giuseppe Leti. The procedure allows to avoid the problem of restriction of variance that…

Methodology · Statistics 2019-07-24 Giuseppe Bove , Pier Luigi Conti , Daniela Marella

Crowdsourcing offers an affordable and scalable means to collect relevance judgments for IR test collections. However, crowd assessors may show higher variance in judgment quality than trusted assessors. In this paper, we investigate how to…

Information Retrieval · Computer Science 2018-06-12 Mucahid Kutlu , Tyler McDonnell , Aashish Sheshadri , Tamer Elsayed , Matthew Lease

Multiple raters are often needed to be used interchangeably in practice for measurement or evaluation. Assessing agreement among these multiple raters via agreement indices are necessary before their participation. While the intuitively…

Methodology · Statistics 2020-06-09 Tongrong Wang , Huiman X. Barnhart

We study a problem where a group of agents has to decide how a joint reward should be shared among them. We focus on settings where the share that each agent receives depends on the subjective opinions of its peers concerning that agent's…

Computer Science and Game Theory · Computer Science 2013-05-23 Arthur Carvalho , Kate Larson

Today's society faces widening disagreement and conflicts among constituents with incompatible views. Escalated views and opinions are seen not only in radical ideology or extremism but also in many other scenes of our everyday life. Here…

Physics and Society · Physics 2020-07-08 Hiroki Sayama

Often exhibiting hierarchical and overlapping structures, communities or modular groups are fundamental and complex in network science. One of the most exploited tools to detect the mesoscopic structure is synchronization. Several phenomena…

Physics and Society · Physics 2018-07-05 Ren Ren , Jinliang Shao

Recommendation to groups of users is a challenging and currently only passingly studied task. Especially the evaluation aspect often appears ad-hoc and instead of truly evaluating on groups of users, synthesizes groups by merging individual…

Artificial Intelligence · Computer Science 2017-08-01 Zsolt Mezei , Carsten Eickhoff

One of the risks involved in multi agent community is in the identification of trustworthy agent partners for transaction. In this paper we aim to describe a trust model for measuring trust in the interacting agents. The trust metric model…

Computers and Society · Computer Science 2013-05-15 Sanat Kumar Bista , Keshav P. Dahal , Peter I. Cowling , Bhadra Man Tuladhar

When LLM-based multi-agent systems disagree, current practice treats this as noise to be resolved through consensus. We propose it can be signal. We focus on hate speech moderation, a domain where judgments depend on cultural context and…

Multiagent Systems · Computer Science 2026-04-07 Michał Wawer , Jarosław A. Chudziak