中文
相关论文

相关论文: DIF Analysis with Unknown Groups and Anchor Items

200 篇论文

Traditional ranking algorithms are designed to retrieve the most relevant items for a user's query, but they often inherit biases from data that can unfairly disadvantage vulnerable groups. Fairness in information access systems (IAS) is…

Uncertainty estimation is a key factor that makes deep learning reliable in practical applications. Recently proposed evidential neural networks explicitly account for different uncertainties by treating the network's outputs as evidence to…

机器学习 · 计算机科学 2023-07-03 Danruo Deng , Guangyong Chen , Yang Yu , Furui Liu , Pheng-Ann Heng

This study evaluated four multi-group differential item functioning (DIF) methods (the root mean square deviation approach, Wald-1, generalized logistic regression procedure, and generalized Mantel-Haenszel method) via Monte Carlo…

应用统计 · 统计学 2024-08-23 Dandan Chen Kaptur , Jinming Zhang

In many real-world regression tasks, the data distribution is heavily skewed, and models learn predominantly from abundant majority samples while failing to predict minority labels accurately. While imbalanced classification has been…

机器学习 · 计算机科学 2025-09-30 Shayan Alahyari

As machine learning (ML) algorithms are increasingly used in social domains to make predictions about humans, there is a growing concern that these algorithms may exhibit biases against certain social groups. Numerous notions of fairness…

机器学习 · 计算机科学 2025-09-30 Zhongteng Cai , Mohammad Mahdi Khalili , Xueru Zhang

Evaluating mathematical reasoning in LLMs is constrained by limited benchmark sizes and inherent model stochasticity, yielding high-variance accuracy estimates and unstable rankings across platforms. On difficult problems, an LLM may fail…

机器学习 · 计算机科学 2026-02-04 Zihan Dong , Zhixian Zhang , Yang Zhou , Can Jin , Ruijia Wu , Linjun Zhang

Most fair machine learning methods either highly rely on the sensitive information of the training samples or require a large modification on the target models, which hinders their practical application. To address this issue, we propose a…

机器学习 · 计算机科学 2023-12-27 Haonan Wang , Ziwei Wu , Jingrui He

This paper investigates efficient Difference-in-Differences (DiD) and Event Study (ES) estimation using short panel data sets within the heterogeneous treatment effect framework, free from parametric functional form assumptions and allowing…

计量经济学 · 经济学 2025-06-24 Xiaohong Chen , Pedro H. C. Sant'Anna , Haitian Xie

We present a semi-supervised learning algorithm for learning discrete factor analysis models with arbitrary structure on the latent variables. Our algorithm assumes that every latent variable has an "anchor", an observed variable with only…

机器学习 · 统计学 2015-11-12 Yoni Halpern , Steven Horng , David Sontag

Anomaly detection (AD) has been widely studied for decades in many real-world applications, including fraud detection in finance, and intrusion detection for cybersecurity, etc. Due to the imbalanced nature between protected and unprotected…

机器学习 · 计算机科学 2024-09-18 Ziwei Wu , Lecheng Zheng , Yuancheng Yu , Ruizhong Qiu , John Birge , Jingrui He

Within the educational context, students' assessment tests are routinely validated through Item Response Theory (IRT) models which assume unidimensionality and absence of Differential Item Functioning (DIF). In this paper, we investigate if…

应用统计 · 统计学 2012-12-04 Michela Gnaldi , Francesco Bartolucci , Silvia Bacci

Fairness-aware learning is a novel framework for classification tasks. Like regular empirical risk minimization (ERM), it aims to learn a classifier with a low error rate, and at the same time, for the predictions of the classifier to be…

机器学习 · 统计学 2015-06-26 Kazuto Fukuchi , Jun Sakuma

We propose a new method for estimating causal effects in longitudinal/panel data settings that we call generalized difference-in-differences. Our approach unifies two alternative approaches in these settings: ignorability estimators (e.g.,…

统计方法学 · 统计学 2023-12-12 Denis Agniel , Max Rubinstein , Jessie Coe , Maria DeYoreo

In the context of flexible manufacturing systems that are required to produce different types and quantities of products with minimal reconfiguration, this paper addresses the problem of unsupervised multi-class anomaly detection: develop a…

计算机视觉与模式识别 · 计算机科学 2023-07-18 Haonan Yin , Guanlong Jiao , Qianhui Wu , Borje F. Karlsson , Biqing Huang , Chin Yew Lin

The Difference-in-Differences (DiD) method is a fundamental tool for causal inference, yet its application is often complicated by missing data. Although recent work has developed robust DiD estimators for complex settings like staggered…

统计方法学 · 统计学 2026-01-27 Lorenzo Testa , Edward H. Kennedy , Matthew Reimherr

When using machine learning to aid decision-making, it is critical to ensure that an algorithmic decision is fair and does not discriminate against specific individuals/groups, particularly those from underprivileged populations. Existing…

机器学习 · 计算机科学 2024-11-20 Yifei Wang , Zhengyang Zhou , Liqin Wang , John Laurentiev , Peter Hou , Li Zhou , Pengyu Hong

Uncertainty-aware deep learning (DL) models recently gained attention in fault diagnosis as a way to promote the reliable detection of faults when out-of-distribution (OOD) data arise from unseen faults (epistemic uncertainty) or the…

机器学习 · 计算机科学 2024-12-30 Reza Jalayer , Masoud Jalayer , Andrea Mor , Carlotta Orsenigo , Carlo Vercellis

Multivariate time-series anomaly detection, which is critical for identifying unexpected events, has been explored in the field of machine learning for several decades. However, directly applying these methods to data from forceful tool use…

机器人学 · 计算机科学 2025-09-19 Yating Lin , Zixuan Huang , Fan Yang , Dmitry Berenson

Classic item response models assume that all items with the same difficulty have the same response probability among all respondents with the same ability. These assumptions, however, may very well be violated in practice, and it is not…

统计方法学 · 统计学 2021-08-23 Minjeong Jeon , Ick Hoon Jin , Michael Schweinberger , Samuel Baugh

Constructing an ensemble from a heterogeneous set of unsupervised anomaly detection methods is challenging because the class labels or the ground truth is unknown. Thus, traditional ensemble techniques that use the response variable or the…

机器学习 · 统计学 2021-06-14 Sevvandi Kandanaarachchi