English
Related papers

Related papers: Inter-Rater: Software for analysis of inter-rater …

200 papers

Inter-rater reliability (IRR) is one of the commonly used tools for assessing the quality of ratings from multiple raters. However, applicant selection procedures based on ratings from multiple raters usually result in a binary outcome; the…

Methodology · Statistics 2025-06-17 František Bartoš , Patrícia Martinková

Cohen's and Fleiss' kappa are well-known measures of inter-rater agreement, but they restrict each rater to selecting only one category per subject. This limitation is consequential in contexts where subjects may belong to multiple…

Methodology · Statistics 2025-09-22 Filip Moons , Ellen Vandervieren

Qualitative analysis is typically limited to small datasets because it is time-intensive. Moreover, a second human rater is required to ensure reliable findings. Artificial intelligence tools may replace human raters if we demonstrate high…

Physics Education · Physics 2025-09-03 Nikhil Sanjay Borse , Ravishankar Chatta Subramaniam , N. Sanjay Rebello

We formulate three generalized Bayesian models for analyzing interrater and intrarater reliability in the presence of multilevel data. Stan implementations of these models provide new estimates of interrater and intrarater reliability. We…

Methodology · Statistics 2024-07-18 Nour Hawila , Arthur Berg

This paper investigates the inter-rater reliability of risk assessment instruments (RAIs). The main question is whether different, socially salient groups are affected differently by a lack of inter-rater reliability of RAIs, that is,…

Computers and Society · Computer Science 2023-08-30 Tim Räz

In recent years, the research on empirical software engineering that uses qualitative data analysis (e.g., cases studies, interview surveys, and grounded theory studies) is increasing. However, most of this research does not deep into the…

Software Engineering · Computer Science 2025-09-23 Ángel González-Prieto , Jorge Perez , Jessica Diaz , Daniel López-Fernández

Inter-coder agreement measures, like Cohen's kappa, correct the relative frequency of agreement between coders to account for agreement which simply occurs by chance. However, in some situations these measures exhibit behavior which make…

Applications · Statistics 2012-08-07 Dirk Schuster

In the Internet era the information overload and the challenge to detect quality content has raised the issue of how to rank both resources and users in online communities. In this paper we develop a general ranking method that can…

Physics and Society · Physics 2016-09-23 Hao Liao , Giulio Cimini , Matus Medo

We present a new approach to interpreting IRR that is empirical and contextualized. It is based upon benchmarking IRR against baseline measures in a replication, one of which is a novel cross-replication reliability (xRR) measure based on…

Applications · Statistics 2021-06-15 Ka Wong , Praveen Paritosh , Lora Aroyo

Inter-rater reliability (IRR), which is a prerequisite of high-quality ratings and assessments, may be affected by contextual variables such as the rater's or ratee's gender, major, or experience. Identification of such heterogeneity…

Methodology · Statistics 2023-02-17 Patrícia Martinková , František Bartoš , Marek Brabec

This work is motivated by the need to assess the degree of agreement between two independent groups of raters. It proposes two new methods.

Applications · Statistics 2018-06-18 Madhusmita Panda , Sharayu Paranjpe , Anil Gore

The purpose of this protocol is to be useful to identify, evaluate and synthesize reported knowledge about the measurement of interpersonal trust (IpT) in virtual software teams. To achieve this goal we applied a research technique known as…

Software Engineering · Computer Science 2020-02-13 Sergio Zapata , José Luis Barros-Justo , Gerardo Maturro , Samuel Sepúlveda

In this note, a connection between inter-rater reliability and individual fairness is established. It is shown that inter-rater reliability is a special case of individual fairness, a notion of fairness requiring that similar people are…

Computers and Society · Computer Science 2023-08-11 Tim Räz

Measurement of the interrater agreement (IRA) is critical in various disciplines. To correct for potential confounding chance agreement in IRA, Cohen's kappa and many other methods have been proposed. However, owing to the varied strategies…

Methodology · Statistics 2024-02-14 Zizhong Tian , Vernon M. Chinchilli , Chan Shen , Shouhao Zhou

This study was motivated by the problem of identifying fake documents on the Internet. To explore possible solutions to this problem we introduce a model of a network community in which members submit documents with verifiable content.…

Classical Analysis and ODEs · Mathematics 2018-12-20 Andrei Olifer

In this paper peer review reliability is investigated based on peer ratings of research teams at two Belgian universities. It is found that outcomes can be substantially influenced by the different ways in which experts attribute ratings.…

Digital Libraries · Computer Science 2013-07-29 Nadine Rons , Eric Spruyt

Reputation is crucial to enabling human or software agents to select among alternative providers. Although several effective reputation assessment methods exist, they typically distil reputation into a numerical representation, with no…

Artificial Intelligence · Computer Science 2020-06-17 Ingrid Nunes , Phillip Taylor , Lina Barakat , Nathan Griffiths , Simon Miles

Since the inception of crowdsourcing, aggregation has been a common strategy for dealing with unreliable data. Aggregate ratings are more reliable than individual ones. However, many natural language processing (NLP) applications that rely…

Artificial Intelligence · Computer Science 2022-03-25 Ka Wong , Praveen Paritosh

Classic Delphi and Fuzzy Delphi methods are used to test content validity of a data collection tools such as questionnaires. Fuzzy Delphi takes the opinion issued by judges from a linguistic perspective reducing ambiguity in opinions by…

Computers and Society · Computer Science 2024-03-12 Rosana Montes , Ana M. Sanchez , Pedro Villar , Francisco Herrera

Networked applications have software components that reside on different computers. Email, for example, has database, processing, and user interface components that can be distributed across a network and shared by users in different…

Methodology · Statistics 2007-08-03 John M. Chambers , David A. James , Diane Lambert , Scott Vander Wiel
‹ Prev 1 2 3 10 Next ›