English
Related papers

Related papers: Crowdsourcing with Difficulty: A Bayesian Rating M…

200 papers

Machine learning from training data with a skewed distribution of examples per class can lead to models that favor performance on common classes at the expense of performance on rare ones. AudioSet has a very wide range of priors over its…

Machine Learning · Computer Science 2023-07-04 R. Channing Moore , Daniel P. W. Ellis , Eduardo Fonseca , Shawn Hershey , Aren Jansen , Manoj Plakal

Given a pre-trained classifier and multiple human experts, we investigate the task of online classification where model predictions are provided for free but querying humans incurs a cost. In this practical but under-explored setting,…

Machine Learning · Computer Science 2023-12-14 Sam Showalter , Alex Boyd , Padhraic Smyth , Mark Steyvers

We establish concentration rates for estimation of treatment effects in experiments that incorporate prior sources of information -- such as past pilots, related studies, or expert assessments -- whose external validity is uncertain. Each…

Econometrics · Economics 2026-03-24 Frederico Finan , Demian Pouzo

We consider the $M$-ary classification problem via crowdsourcing, where crowd workers respond to simple binary questions and the answers are aggregated via decision fusion. The workers have a reject option to skip answering a question when…

Human-Computer Interaction · Computer Science 2020-08-26 Baocheng Geng , Qunwei Li , Pramod K. Varshney

We consider the problem of sequential evaluation, in which an evaluator observes candidates in a sequence and assigns scores to these candidates in an online, irrevocable fashion. Motivated by the psychology literature that has studied…

Machine Learning · Statistics 2023-11-20 Jingyan Wang , Ashwin Pananjady

Fueled by the call for formative assessments, diagnostic classification models (DCMs) have recently gained popularity in psychometrics. Despite their potential for providing diagnostic information that aids in classroom instruction and…

Computation · Statistics 2022-08-26 Motonori Oka , Kensuke Okada

Modern machine learning models with a large number of parameters often generalize well despite perfectly interpolating noisy training data - a phenomenon known as benign overfitting. A foundational explanation for this in linear…

Machine Learning · Statistics 2025-11-18 Yuta Kondo

Regression-based predictive analytics used in modern kidney transplantation is known to inherit biases from training data. This leads to social discrimination and inefficient organ utilization, particularly in the context of a few social…

Human-Computer Interaction · Computer Science 2025-05-13 Mukund Telukunta , Venkata Sriram Siddhardh Nadendla , Morgan Stuart , Casey Canfield

We propose the Heterogeneous Thurstone Model (HTM) for aggregating ranked data, which can take the accuracy levels of different users into account. By allowing different noise distributions, the proposed HTM model maintains the generality…

Machine Learning · Computer Science 2019-12-04 Tao Jin , Pan Xu , Quanquan Gu , Farzad Farnoud

We consider learning from data of variable quality that may be obtained from different heterogeneous sources. Addressing learning from heterogeneous data in its full generality is a challenging problem. In this paper, we adopt instead a…

Machine Learning · Computer Science 2014-12-19 Shuang Song , Kamalika Chaudhuri , Anand D. Sarwate

The ground truth used for training image, video, or speech quality prediction models is based on the Mean Opinion Scores (MOS) obtained from subjective experiments. Usually, it is necessary to conduct multiple experiments, mostly with…

Audio and Speech Processing · Electrical Eng. & Systems 2021-12-15 Gabriel Mittag , Saman Zadtootaghaj , Thilo Michael , Babak Naderi , Sebastian Möller

To investigate the structure of individual differences in performance on behavioral tasks, Haaf and Rouder (2017) developed a class of hierarchical Bayesian mixed models with varying levels of constraint on the individual effects. The…

Applications · Statistics 2022-10-24 Thomas J. Faulkenberry

Online crowdsourcing provides a scalable and inexpensive means to collect knowledge (e.g. labels) about various types of data items (e.g. text, audio, video). However, it is also known to result in large variance in the quality of recorded…

Human-Computer Interaction · Computer Science 2018-12-10 Yuan Jin , Mark Carman , Ye Zhu , Yong Xiang

We explore the design of an effective crowdsourcing system for an $M$-ary classification task. Crowd workers complete simple binary microtasks whose results are aggregated to give the final classification decision. We consider the scenario…

Social and Information Networks · Computer Science 2017-04-05 Qunwei Li , Pramod K. Varshney

Crowdsourcing is a popular paradigm for effectively collecting labels at low cost. The Dawid-Skene estimator has been widely used for inferring the true labels from the noisy labels provided by non-expert crowdsourcing workers. However,…

Machine Learning · Statistics 2014-11-04 Yuchen Zhang , Xi Chen , Dengyong Zhou , Michael I. Jordan

We present three tiers of Bayesian consistency tests for the general case of $correlated$ datasets. Building on duplicates of the model parameters assigned to each dataset, these tests range from Bayesian evidence ratios as a global summary…

Cosmology and Nongalactic Astrophysics · Physics 2019-01-24 Fabian Köhlinger , Benjamin Joachimi , Marika Asgari , Massimo Viola , Shahab Joudaki , Tilman Tröster

Researchers have raised awareness about the harms of aggregating labels especially in subjective tasks that naturally contain disagreements among human annotators. In this work we show that models that are only provided aggregated labels…

Computation and Language · Computer Science 2024-03-08 Abhishek Anand , Negar Mokhberian , Prathyusha Naresh Kumar , Anweasha Saha , Zihao He , Ashwin Rao , Fred Morstatter , Kristina Lerman

A shortcoming of batch reinforcement learning is its requirement for rewards in data, thus not applicable to tasks without reward functions. Existing settings for lack of reward, such as behavioral cloning, rely on optimal demonstrations…

Machine Learning · Computer Science 2022-11-30 Guoxi Zhang , Hisashi Kashima

Oral cancer ranks among the most prevalent cancers globally, with a particularly high mortality rate in regions lacking adequate healthcare access. Early diagnosis is crucial for reducing mortality; however, challenges persist due to…

Computer Vision and Pattern Recognition · Computer Science 2025-10-03 Akshay Bhagwan Sonawane , Lena D. Swamikannan , Lakshman Tamil

Many operational AI systems depend on large-scale human annotation to detect rare but consequential events (e.g., fraud, defects, and medical abnormalities). When positives are rare, the prevalence effect induces systematic cognitive biases…

Human-Computer Interaction · Computer Science 2026-03-13 Gunnar P. Epping , Andrew Caplin , Erik Duhaime , William R. Holmes , Daniel Martin , Jennifer S. Trueblood
‹ Prev 1 8 9 10 Next ›