中文
相关论文

相关论文: Representation-Aware Experimentation: Group Inequa…

200 篇论文

Ranking and scoring are ubiquitous. We consider the setting in which an institution, called a ranker, evaluates a set of individuals based on demographic, behavioral or other characteristics. The final output is a ranking that represents…

数据库 · 计算机科学 2016-10-28 Ke Yang , Julia Stoyanovich

Selection bias is a serious potential problem for inference about relationships of scientific interest based on samples without well-defined probability sampling mechanisms. Motivated by the potential for selection bias in (a) estimated…

Assessing the diversity of a dataset of information associated with people is crucial before using such data for downstream applications. For a given dataset, this often involves computing the imbalance or disparity in the empirical…

计算机与社会 · 计算机科学 2021-07-16 Vijay Keswani , L. Elisa Celis

Ranking functions that are used in decision systems often produce disparate results for different populations because of bias in the underlying data. Addressing, and compensating for, these disparate outcomes is a critical problem for fair…

机器学习 · 计算机科学 2024-04-23 Abraham Gale , Amélie Marian

In many applications, data can be heterogeneous in the sense of spanning latent groups with different underlying distributions. When predictive models are applied to such data the heterogeneity can affect both predictive performance and…

机器学习 · 统计学 2022-05-04 Thomas Lartigue , Sach Mukherjee

Data integration methods aim to extract low-dimensional embeddings from high-dimensional outcomes to remove unwanted variations, such as batch effects and unmeasured covariates, across heterogeneous datasets. However, multiple hypothesis…

统计方法学 · 统计学 2025-12-15 Jin-Hong Du , Kathryn Roeder , Larry Wasserman

In this review, we present econometric and statistical methods for analyzing randomized experiments. For basic experiments we stress randomization-based inference as opposed to sampling-based inference. In randomization-based inference,…

统计方法学 · 统计学 2017-10-26 Susan Athey , Guido Imbens

In the first stage of a two-stage study, the researcher uses a statistical model to impute the unobserved exposures. In the second stage, imputed exposures serve as covariates in epidemiological models. Imputation error in the first stage…

应用统计 · 统计学 2021-07-19 Ron Sarafian , Itai Kloog , Jonathan D. Rosenblatt

Although creativity is encouraged in the abstract it is often discouraged in educational and workplace settings. Using an agent-based model of cultural evolution, we investigated the idea that tempering the novelty-generating effects of…

多智能体系统 · 计算机科学 2019-03-15 Liane Gabora , Simon Tseng

Automatic Gender Recognition (AGR) systems are an increasingly widespread application in the Machine Learning (ML) landscape. While these systems are typically understood as detecting gender, they often classify datapoints based on…

计算机视觉与模式识别 · 计算机科学 2025-06-04 Camilla Quaresmini , Giacomo Zanotti

Motivated by scenarios where data is used for diverse prediction tasks, we study whether fair representation can be used to guarantee fairness for unknown tasks and for multiple fairness notions simultaneously. We consider seven group…

机器学习 · 计算机科学 2022-02-22 Xudong Shen , Yongkang Wong , Mohan Kankanhalli

As electrical generation becomes more distributed and volatile, and loads become more uncertain, controllability of distributed energy resources (DERs), regardless of their ownership status, will be necessary for grid reliability. Grid…

系统与控制 · 电气工程与系统科学 2024-10-22 Adam Lechowicz , Joshua Comden , Andrey Bernstein

Organizations cannot address demographic disparities that they cannot see. Recent research on machine learning and fairness has emphasized that awareness of sensitive attributes, such as race and sex, is critical to the development of…

计算机与社会 · 计算机科学 2019-12-16 Miranda Bogen , Aaron Rieke , Shazeda Ahmed

How should social scientists understand and communicate the uncertainty of statistically estimated causal effects? I propose we utilize the posterior distribution of a causal effect and present the probability of the effect being greater…

应用统计 · 统计学 2022-11-15 Akisato Suzuki

A/B testing refers to the statistical procedure of conducting an experiment to compare two treatments, A and B, applied to different testing subjects. It is widely used by technology companies such as Facebook, LinkedIn, and Netflix, to…

统计方法学 · 统计学 2026-05-12 Victoria Pokhiko , Qiong Zhang , Lulu Kang , D'arcy P. Mays

One fundamental statistical question for research areas such as precision medicine and health disparity is about discovering effect modification of treatment or exposure by observed covariates. We propose a semiparametric framework for…

统计方法学 · 统计学 2020-08-04 Muxuan Liang , Menggang Yu

A common problem in numerous research areas, particularly in clinical trials, is to test whether the effect of an explanatory variable on an outcome variable is equivalent across different groups. In practice, these tests are frequently…

统计方法学 · 统计学 2024-05-03 Niklas Hagemann , Kathrin Möllenhoff

Statistical models that include random effects are commonly used to analyze longitudinal and correlated data, often with strong and parametric assumptions about the random effects distribution. There is marked disagreement in the literature…

统计方法学 · 统计学 2012-01-11 Charles E. McCulloch , John M. Neuhaus

The foremost challenge to causal inference with real-world data is to handle the imbalance in the covariates with respect to different treatment options, caused by treatment selection bias. To address this issue, recent literature has…

机器学习 · 统计学 2022-02-23 Zhixuan Chu , Stephen Rathbun , Sheng Li

In sequential decision-making problems involving sensitive attributes like race and gender, reinforcement learning (RL) agents must carefully consider long-term fairness while maximizing returns. Recent works have proposed many different…

机器学习 · 计算机科学 2024-04-30 Zhihong Deng , Jing Jiang , Guodong Long , Chengqi Zhang