English
Related papers

Related papers: Improved Neymanian analysis for $2^K$ factorial de…

200 papers

Estimating the causal effect of a treatment or health policy with observational data can be challenging due to an imbalance of and a lack of overlap between treated and control covariate distributions. In the presence of limited overlap,…

Methodology · Statistics 2025-03-24 Martha Barnard , Jared D. Huling , Julian Wolfson

With reference to a baseline parametrization, we explore highly efficient fractional factorial designs for inference on the main effects and, perhaps, some interactions. Our tools include approximate theory together with certain carefully…

Statistics Theory · Mathematics 2014-05-14 Rahul Mukerjee , S. Huda

Factorial designs are widely used due to their ability to accommodate multiple factors simultaneously. The factor-based regression with main effects and some interactions is the dominant strategy for downstream data analysis, delivering…

Methodology · Statistics 2021-12-09 Anqi Zhao , Peng Ding

We investigate Bayesian predictive inference for finite population quantities when there are unequal probabilities of selection. Only limited information about the sample design is available; i.e., only the first-order selection…

Methodology · Statistics 2018-04-10 Junheng Ma , Joe Sedransk , Balgobin Nandram , Lu Chen

Estimation of genewise variance arises from two important applications in microarray data analysis: selecting significantly differentially expressed genes and validation tests for normalization of microarray data. We approach the problem by…

Statistics Theory · Mathematics 2010-11-11 Jianqing Fan , Yang Feng , Yue S. Niu

The multivariate coefficient of variation (MCV) is an attractive and easy-to-interpret effect size for the dispersion in multivariate data. Recently, the first inference methods for the MCV were proposed by Ditzhaus and Smaga (2022) for…

Methodology · Statistics 2023-01-31 Marc Ditzhaus , Łukasz Smaga

We propose a consistent estimator of sharp bounds on the variance of the difference-in-means estimator in completely randomized experiments. Generalizing Robins [Stat. Med. 7 (1988) 773-785], our results resolve a well-known identification…

Statistics Theory · Mathematics 2014-05-27 Peter M. Aronow , Donald P. Green , Donald K. K. Lee

Motivated by problems of anomaly detection, this paper implements the Neyman-Pearson paradigm to deal with asymmetric errors in binary classification with a convex loss. Given a finite collection of classifiers, we combine them and obtain a…

Machine Learning · Statistics 2011-03-01 Philippe Rigollet , Xin Tong

Micro-randomized trials are commonly conducted for optimizing mobile health interventions such as push notifications for behavior change. In analyzing such trials, causal excursion effects are often of primary interest, and their estimation…

Methodology · Statistics 2024-08-19 Yihan Bao , Lauren Bell , Elizabeth Williamson , Claire Garnett , Tianchen Qian

This paper proposes a flexible new framework for constructing Neyman-orthogonal scores in semiparametric models involving infinite-dimensional nuisance parameters. While locally estimation is vital for integrating machine learning into…

Methodology · Statistics 2026-04-30 Kun Ren , Wen Su , Li Liu , Ian W. McKeague , Xingqiu Zhao

Experimental design is crucial for inference where limitations in the data collection procedure are present due to cost or other restrictions. Optimal experimental designs determine parameters that in some appropriate sense make the data…

Machine Learning · Statistics 2016-03-11 Panagiotis Tsilifis , Roger G. Ghanem , Paris Hajali

We consider nonregular fractions of factorial experiments for a class of linear models. These models have a common general mean and main effects, however they may have different 2-factor interactions. Here we assume for simplicity that…

Computation · Statistics 2019-11-11 Shrabanti Chowdhury , Joshua Lukemire , Abhyuday Mandal

Regression adjustments are often made to experimental data. Since randomization does not justify the models, bias is likely; nor are the usual variance calculations to be trusted. Here, we evaluate regression adjustments using Neyman's…

Applications · Statistics 2008-12-18 David A. Freedman

Identifying causal effects is a key problem of interest across many disciplines. The two long-standing approaches to estimate causal effects are observational and experimental (randomized) studies. Observational studies can suffer from…

Machine Learning · Computer Science 2024-07-09 Sepehr Elahi , Sina Akbari , Jalal Etesami , Negar Kiyavash , Patrick Thiran

Since polynomial regression models are generally quite reliable for data with a linear trend, it is important to note that, in some cases, they may encounter overfitting issues during the training phase, which could result in negative…

Methodology · Statistics 2025-03-21 Anthony Torres-Hernandez

In 1990, Jakeman (see \cite{jakeman1990statistics}) defined the binomial process as a special case of the classical birth-death process, where the probability of birth is proportional to the difference between a fixed number and the number…

Statistics Theory · Mathematics 2024-05-15 Meena Sanjay Babulal , Sunil Kumar Gauttam , Aditya Maheshwari

In the partially-observed outcome setting, a recent set of proposals known as "prediction-powered inference" (PPI) involve (i) applying a pre-trained machine learning model to predict the response, and then (ii) using these predictions to…

Methodology · Statistics 2026-02-12 Runjia Zou , Daniela Witten , Brian Williamson

Consider a setting where there are $N$ heterogeneous units and $p$ interventions. Our goal is to learn unit-specific potential outcomes for any combination of these $p$ interventions, i.e., $N \times 2^p$ causal parameters. Choosing a…

Methodology · Statistics 2024-01-17 Abhineet Agarwal , Anish Agarwal , Suhas Vijaykumar

Under the Neyman causal model, it is well-known that OLS with treatment-by-covariate interactions cannot harm asymptotic precision of estimated treatment effects in completely randomized experiments. But do such guarantees extend to…

Statistics Theory · Mathematics 2018-03-19 Joel A. Middleton

Unbiased and consistent variance estimators generally do not exist for design-based treatment effect estimators because experimenters never observe more than one potential outcome for any unit. The problem is exacerbated by interference and…

Methodology · Statistics 2024-07-04 Christopher Harshaw , Joel A. Middleton , Fredrik Sävje