中文
相关论文

相关论文: Lurking Inferential Monsters? Quantifying bias in …

200 篇论文

Despite their popularity, machine learning predictions are sensitive to potential unobserved predictors. This paper proposes a general algorithm that assesses how the omission of an unobserved variable with high explanatory power could…

计量经济学 · 经济学 2021-02-09 Falco J. Bargagli Stoffi , Kenneth De Beckker , Joana E. Maldonado , Kristof De Witte

Despite growing interest in using Large Language Models (LLMs) for educational assessment, it remains unclear how closely they align with human scoring. We present a systematic evaluation of instruction-tuned LLMs across three open…

计算与语言 · 计算机科学 2026-04-02 Filip J. Kucia , Anirban Chakraborty , Anna Wróblewska

Recommender systems are seen as an effective tool to address information overload, but it is widely known that the presence of various biases makes direct training on large-scale observational data result in sub-optimal prediction…

信息检索 · 计算机科学 2023-04-19 Haoxuan Li , Yanghao Xiao , Chunyuan Zheng , Peng Wu

We clarify what fairness guarantees we can and cannot expect to follow from unconstrained machine learning. Specifically, we characterize when unconstrained learning on its own implies group calibration, that is, the outcome variable is…

机器学习 · 计算机科学 2019-01-28 Lydia T. Liu , Max Simchowitz , Moritz Hardt

Recovering and distinguishing between the strict-preference, indifference and/or indecisiveness parts of a decision maker's preferences is a challenging task but also important for testing theory and conducting welfare analysis. This paper…

理论经济学 · 经济学 2025-09-15 Georgios Gerasimou

Offline evaluation of language models from usage logs is biased when model choice is confounded: the same user-side factors that influence which model is used can also influence how its output is judged, so raw comparisons of logged scores…

机器学习 · 计算机科学 2026-05-05 Jikai Jin , Vasilis Syrgkanis

It is a truth universally acknowledged that an observed association without known mechanism must be in want of a causal estimate. However, causal estimation from observational data often relies on the (untestable) assumption of `no…

统计方法学 · 统计学 2020-12-10 Victor Veitch , Anisha Zaveri

For many interesting tasks, such as medical diagnosis and web page classification, a learner only has access to some positively labeled examples and many unlabeled examples. Learning from this type of data requires making assumptions about…

机器学习 · 计算机科学 2018-08-28 Jessa Bekker , Jesse Davis

Automatic Essay Scoring (AES) is a well-established educational pursuit that employs machine learning to evaluate student-authored essays. While much effort has been made in this area, current research primarily focuses on either (i)…

计算与语言 · 计算机科学 2024-01-12 Kaixun Yang , Mladen Raković , Yuyang Li , Quanlong Guan , Dragan Gašević , Guanliang Chen

Background: Although the missing covariate indicator method (MCIM) has been shown to be biased under extreme conditions, the degree and determinants of bias have not been formally assessed. We derived the formula for the relative bias in…

应用统计 · 统计学 2025-08-01 Gang Xu , Mingyang Song , Xin Zhou , Yilun Wu , Mathew Pazaris , Donna Spiegelman

In many applications of causal inference, the treatment received by one unit may influence the outcome of another, a phenomenon referred to as interference. Although there are several frameworks for conducting causal inference in the…

统计方法学 · 统计学 2025-11-27 Matvey Ortyashov , AmirEmad Ghassami

This paper deals with the issue of ecological bias in ecological inference. We provide an explicit formulation of the conditions required for the ordinary ecological regression to produce unbiased estimates and argue that, when these…

应用统计 · 统计学 2015-09-11 Michela Gnaldi , Venera Tomaselli , Antonio Forcina

Fairness for machine learning predictions is widely required in practice for legal, ethical, and societal reasons. Existing work typically focuses on settings without unobserved confounding, even though unobserved confounding can lead to…

机器学习 · 计算机科学 2024-10-23 Maresa Schröder , Dennis Frauen , Stefan Feuerriegel

Intercurrent events, common in clinical trials and observational studies, affect the existence or interpretation of final outcomes. Principal stratification addresses this challenge by defining local average treatment effect estimands…

统计方法学 · 统计学 2025-09-22 Jiaqi Tong , Brennan Kahan , Michael O. Harhay , Fan Li

Principal stratification is a popular framework for causal inference in the presence of an intermediate outcome. While the principal average treatment effects are the standard target of inference, they may be insufficient when interest lies…

统计方法学 · 统计学 2025-12-29 Xinyuan Chen , Fan Li

The possibility of unmeasured confounding is one of the main limitations for causal inference from observational studies. There are different methods for (partially) empirically assessing the plausibility of unconfoundedness. However, most…

统计方法学 · 统计学 2025-10-28 Fernando Pires Hartwig , Kate Tilling , George Davey Smith

Despite the impressive prediction ability, machine learning models show discrimination towards certain demographics and suffer from unfair prediction behaviors. To alleviate the discrimination, extensive studies focus on eliminating the…

机器学习 · 计算机科学 2023-07-11 Chia-Yuan Chang , Yu-Neng Chuang , Kwei-Herng Lai , Xiaotian Han , Xia Hu , Na Zou

In many application settings, the data have missing entries which make analysis challenging. An abundant literature addresses missing values in an inferential framework: estimating parameters and their variance from incomplete tables. Here,…

机器学习 · 统计学 2024-03-22 Julie Josse , Jacob M. Chen , Nicolas Prost , Erwan Scornet , Gaël Varoquaux

In the UK, US and elsewhere, school accountability systems increasingly compare schools using value-added measures of school performance derived from pupil scores in high-stakes standardised tests. Rather than naively comparing school…

应用统计 · 统计学 2018-11-26 George Leckie , Harvey Goldstein

As Large Language Models (LLMs) are increasingly embedded in real-world decision-making processes, it becomes crucial to examine the extent to which they exhibit cognitive biases. Extensively studied in the field of psychology, cognitive…

计算与语言 · 计算机科学 2025-09-30 R. Alexander Knipper , Charles S. Knipper , Kaiqi Zhang , Valerie Sims , Clint Bowers , Santu Karmaker