中文
相关论文

相关论文: Do peers share the same criteria for assessing gra…

200 篇论文

Iterative peer grading activities may keep students engaged during in-class project presentations. Effective methods for collecting and aggregating peer assessment data are essential. Students tend to grade projects favorably. So, while…

计算机科学与博弈论 · 计算机科学 2025-03-25 Lihi Dery

The peer review system has been traditionally challenged due to its many limitations especially for allocating funding. Bibliometric indicators may well present themselves as a complement. Objective: We analyze the relationship between…

Crowdsourcing offers a practical method for ranking and scoring large amounts of items. To investigate the algorithms and incentives that can be used in crowdsourcing quality evaluations, we built CrowdGrader, a tool that lets students…

社会与信息网络 · 计算机科学 2013-08-27 Luca de Alfaro , Michael Shavlovsky

Benchmarks underpin how progress in large language models (LLMs) is measured and trusted. Yet our analyses reveal that apparent convergence in benchmark accuracy can conceal deep epistemic divergence. Using two major reasoning benchmarks -…

计算与语言 · 计算机科学 2026-02-13 Eddie Yang , Dashun Wang

With the rapid increase in paper submissions to academic conferences, the need for automated and accurate paper-reviewer matching is more critical than ever. Previous efforts in this area have considered various factors to assess the…

信息检索 · 计算机科学 2025-02-18 Yu Zhang , Yanzhen Shen , SeongKu Kang , Xiusi Chen , Bowen Jin , Jiawei Han

Peer assessment systems are emerging in many social and multi-agent settings, such as peer grading in large (online) classes, peer review in conferences, peer art evaluation, etc. However, peer assessments might not be as accurate as expert…

计算机与社会 · 计算机科学 2021-11-09 Alireza A. Namanloo , Julie Thorpe , Amirali Salehi-Abari

Peer grading has emerged as a scalable solution for assessment in large and online classrooms, offering both logistical efficiency and pedagogical value. However, designing effective peer-grading systems remains challenging due to…

计算机与社会 · 计算机科学 2025-12-02 Uchswas Paul , Ananya Mantravadi , Jash Shah , Shail Shah , Sri Vaishnavi Mylavarapu , M Parvez Rashid , Edward Gehringer

Quality control is an ongoing concern in citizen science that is often managed by replication to consensus in online tasks such as image classification. Numerous factors can lead to disagreement, including image quality problems, interface…

人机交互 · 计算机科学 2021-10-18 Vinod Kumar Ahuja , Holly K. Rosser , Andrea Grover

The assumption that prediction-equivalent models produce equivalent explanations underlies many practices in explainable AI, including model selection, auditing, and regulatory evaluation. In this work, we show that this assumption does not…

机器学习 · 计算机科学 2026-03-18 Thackshanaramana B

Peer grading systems work well only if users have incentives to grade truthfully. An example of non-truthful grading, that we observed in classrooms, consists in students assigning the maximum grade to all submissions. With a naive grading…

计算机科学与博弈论 · 计算机科学 2016-04-13 Luca de Alfaro , Michael Shavlovsky , Vassilis Polychronopoulos

Peer assessment has been widely applied across diverse academic fields over the last few decades and has demonstrated its effectiveness. However, the advantages of peer assessment can only be achieved with high-quality peer reviews.…

计算与语言 · 计算机科学 2021-10-11 Qinjin Jia , Jialin Cui , Yunkai Xiao , Chengyuan Liu , Parvez Rashid , Edward F. Gehringer

Formality is one of the most important dimensions of writing style variation. In this study we conducted an inter-rater reliability experiment for assessing sentence formality on a five-point Likert scale, and obtained good agreement…

计算与语言 · 计算机科学 2014-04-22 Shibamouli Lahiri , Xiaofei Lu

In many classification tasks, there is no definitive ground truth, only human judgments that may disagree. We address two challenges that arise in such settings: (1) how to use human raters to score classifiers, and (2) how to use them for…

机器学习 · 计算机科学 2026-04-24 Paul Resnick , Yuqing Kong , Grant Schoenebeck , Tim Weninger

This paper systematically investigates the performance of consensus-based distributed filtering under mismatched noise covariances. First, we introduce three performance evaluation indices for such filtering problems,namely the standard…

系统与控制 · 电气工程与系统科学 2024-08-14 Xiaoxu Lyu , Guanghui Wen , Ling Shi , Peihu Duan , Zhisheng Duan

Online social platforms increasingly rely on crowd-sourced systems to label misleading content at scale, but these systems must both aggregate users' evaluations and decide whose evaluations to trust. To address the latter, many platforms…

社会与信息网络 · 计算机科学 2026-05-19 Yeganeh Alimohammadi , Karissa Huang , Christian Borgs , Jennifer Chayes

Crowdsourcing has evolved as an organizational approach to distributed problem solving and innovation. As contests are embedded in online communities and evaluation rights are assigned to the crowd, community members face a tension: they…

综合经济学 · 经济学 2024-04-23 Christoph Riedl , Tom Grad , Christopher Lettl

Peer grading is an educational system in which students assess each other's work. It is commonly applied under Massive Open Online Course (MOOC) and offline classroom settings. With this system, instructors receive a reduced grading…

应用统计 · 统计学 2025-10-01 Giuseppe Mignemi , Yunxiao Chen , Irini Moustaki

The demand for global university league tables has been high over the past two decades. However, significant criticism of their methodologies is accumulating without being addressed. I revisit global university league tables by normalizing…

物理与社会 · 物理学 2023-08-22 Saulo Mendes

It is common to see a handful of reviewers reject a highly novel paper, because they view, say, extensive experiments as far more important than novelty, whereas the community as a whole would have embraced the paper. More generally, the…

人工智能 · 计算机科学 2020-03-03 Ritesh Noothigattu , Nihar B. Shah , Ariel D. Procaccia

The paper describes a potential platform to facilitate academic peer review with emphasis on early-stage research. This platform aims to make peer review more accurate and timely by rewarding reviewers on the basis of peer prediction…

数字图书馆 · 计算机科学 2023-03-30 Alexander Ugarov