English
Related papers

Related papers: Do peers share the same criteria for assessing gra…

200 papers

Iterative peer grading activities may keep students engaged during in-class project presentations. Effective methods for collecting and aggregating peer assessment data are essential. Students tend to grade projects favorably. So, while…

Computer Science and Game Theory · Computer Science 2025-03-25 Lihi Dery

The peer review system has been traditionally challenged due to its many limitations especially for allocating funding. Bibliometric indicators may well present themselves as a complement. Objective: We analyze the relationship between…

Digital Libraries · Computer Science 2013-07-02 Alvaro Cabezas-Clavijo , Nicolás Robinson-Garcia , Manuel Escabias , Evaristo Jiménez-Contreras

Crowdsourcing offers a practical method for ranking and scoring large amounts of items. To investigate the algorithms and incentives that can be used in crowdsourcing quality evaluations, we built CrowdGrader, a tool that lets students…

Social and Information Networks · Computer Science 2013-08-27 Luca de Alfaro , Michael Shavlovsky

Benchmarks underpin how progress in large language models (LLMs) is measured and trusted. Yet our analyses reveal that apparent convergence in benchmark accuracy can conceal deep epistemic divergence. Using two major reasoning benchmarks -…

Computation and Language · Computer Science 2026-02-13 Eddie Yang , Dashun Wang

With the rapid increase in paper submissions to academic conferences, the need for automated and accurate paper-reviewer matching is more critical than ever. Previous efforts in this area have considered various factors to assess the…

Information Retrieval · Computer Science 2025-02-18 Yu Zhang , Yanzhen Shen , SeongKu Kang , Xiusi Chen , Bowen Jin , Jiawei Han

Peer assessment systems are emerging in many social and multi-agent settings, such as peer grading in large (online) classes, peer review in conferences, peer art evaluation, etc. However, peer assessments might not be as accurate as expert…

Computers and Society · Computer Science 2021-11-09 Alireza A. Namanloo , Julie Thorpe , Amirali Salehi-Abari

Peer grading has emerged as a scalable solution for assessment in large and online classrooms, offering both logistical efficiency and pedagogical value. However, designing effective peer-grading systems remains challenging due to…

Computers and Society · Computer Science 2025-12-02 Uchswas Paul , Ananya Mantravadi , Jash Shah , Shail Shah , Sri Vaishnavi Mylavarapu , M Parvez Rashid , Edward Gehringer

Quality control is an ongoing concern in citizen science that is often managed by replication to consensus in online tasks such as image classification. Numerous factors can lead to disagreement, including image quality problems, interface…

Human-Computer Interaction · Computer Science 2021-10-18 Vinod Kumar Ahuja , Holly K. Rosser , Andrea Grover

The assumption that prediction-equivalent models produce equivalent explanations underlies many practices in explainable AI, including model selection, auditing, and regulatory evaluation. In this work, we show that this assumption does not…

Machine Learning · Computer Science 2026-03-18 Thackshanaramana B

Peer grading systems work well only if users have incentives to grade truthfully. An example of non-truthful grading, that we observed in classrooms, consists in students assigning the maximum grade to all submissions. With a naive grading…

Computer Science and Game Theory · Computer Science 2016-04-13 Luca de Alfaro , Michael Shavlovsky , Vassilis Polychronopoulos

Peer assessment has been widely applied across diverse academic fields over the last few decades and has demonstrated its effectiveness. However, the advantages of peer assessment can only be achieved with high-quality peer reviews.…

Computation and Language · Computer Science 2021-10-11 Qinjin Jia , Jialin Cui , Yunkai Xiao , Chengyuan Liu , Parvez Rashid , Edward F. Gehringer

Formality is one of the most important dimensions of writing style variation. In this study we conducted an inter-rater reliability experiment for assessing sentence formality on a five-point Likert scale, and obtained good agreement…

Computation and Language · Computer Science 2014-04-22 Shibamouli Lahiri , Xiaofei Lu

In many classification tasks, there is no definitive ground truth, only human judgments that may disagree. We address two challenges that arise in such settings: (1) how to use human raters to score classifiers, and (2) how to use them for…

Machine Learning · Computer Science 2026-04-24 Paul Resnick , Yuqing Kong , Grant Schoenebeck , Tim Weninger

This paper systematically investigates the performance of consensus-based distributed filtering under mismatched noise covariances. First, we introduce three performance evaluation indices for such filtering problems,namely the standard…

Systems and Control · Electrical Eng. & Systems 2024-08-14 Xiaoxu Lyu , Guanghui Wen , Ling Shi , Peihu Duan , Zhisheng Duan

Online social platforms increasingly rely on crowd-sourced systems to label misleading content at scale, but these systems must both aggregate users' evaluations and decide whose evaluations to trust. To address the latter, many platforms…

Social and Information Networks · Computer Science 2026-05-19 Yeganeh Alimohammadi , Karissa Huang , Christian Borgs , Jennifer Chayes

Crowdsourcing has evolved as an organizational approach to distributed problem solving and innovation. As contests are embedded in online communities and evaluation rights are assigned to the crowd, community members face a tension: they…

General Economics · Economics 2024-04-23 Christoph Riedl , Tom Grad , Christopher Lettl

Peer grading is an educational system in which students assess each other's work. It is commonly applied under Massive Open Online Course (MOOC) and offline classroom settings. With this system, instructors receive a reduced grading…

Applications · Statistics 2025-10-01 Giuseppe Mignemi , Yunxiao Chen , Irini Moustaki

The demand for global university league tables has been high over the past two decades. However, significant criticism of their methodologies is accumulating without being addressed. I revisit global university league tables by normalizing…

Physics and Society · Physics 2023-08-22 Saulo Mendes

It is common to see a handful of reviewers reject a highly novel paper, because they view, say, extensive experiments as far more important than novelty, whereas the community as a whole would have embraced the paper. More generally, the…

Artificial Intelligence · Computer Science 2020-03-03 Ritesh Noothigattu , Nihar B. Shah , Ariel D. Procaccia

The paper describes a potential platform to facilitate academic peer review with emphasis on early-stage research. This platform aims to make peer review more accurate and timely by rewarding reviewers on the basis of peer prediction…

Digital Libraries · Computer Science 2023-03-30 Alexander Ugarov