English

Quality Control in Open-Ended Crowdsourcing: A Survey

Human-Computer Interaction 2024-12-06 v1 Distributed, Parallel, and Cluster Computing

Abstract

Crowdsourcing provides a flexible approach for leveraging human intelligence to solve large-scale problems, gaining widespread acceptance in domains like intelligent information processing, social decision-making, and crowd ideation. However, the uncertainty of participants significantly compromises the answer quality, sparking substantial research interest. Existing surveys predominantly concentrate on quality control in Boolean tasks, which are generally formulated as simple label classification, ranking, or numerical prediction. Ubiquitous open-ended tasks like question-answering, translation, and semantic segmentation have not been sufficiently discussed. These tasks usually have large to infinite answer spaces and non-unique acceptable answers, posing significant challenges for quality assurance. This survey focuses on quality control methods applicable to open-ended tasks in crowdsourcing. We propose a two-tiered framework to categorize related works. The first tier introduces a holistic view of the quality model, encompassing key aspects like task, worker, answer, and system. The second tier refines the classification into more detailed categories, including quality dimensions, evaluation metrics, and design decisions, providing insights into the internal structures of the quality control framework in each aspect. We thoroughly investigate how these quality control methods are implemented in state-of-the-art works and discuss key challenges and potential future research directions.

Keywords

Cite

@article{arxiv.2412.03991,
  title  = {Quality Control in Open-Ended Crowdsourcing: A Survey},
  author = {Lei Chai and Hailong Sun and Jing Zhang},
  journal= {arXiv preprint arXiv:2412.03991},
  year   = {2024}
}

Comments

27 pages, 7 figures

R2 v1 2026-06-28T20:23:57.122Z