中文
相关论文

相关论文: On the role of the design phase in a linear regres…

200 篇论文

Consider a researcher estimating the parameters of a regression function based on data for all 50 states in the United States or on data for all visits to a website. What is the interpretation of the estimated parameters and the standard…

统计理论 · 数学 2019-06-25 Alberto Abadie , Susan Athey , Guido W. Imbens , Jeffrey M. Wooldridge

The Regression Discontinuity (RD) design is a widely used non-experimental method for causal inference and program evaluation. While its canonical formulation only requires a score and an outcome variable, it is common in empirical work to…

统计方法学 · 统计学 2022-08-25 Matias D. Cattaneo , Luke Keele , Rocio Titiunik

The design of experiments involves a compromise between covariate balance and robustness. This paper provides a formalization of this trade-off and describes an experimental design that allows experimenters to navigate it. The design is…

统计方法学 · 统计学 2023-11-16 Christopher Harshaw , Fredrik Sävje , Daniel Spielman , Peng Zhang

We consider in this paper the problem of optimal experiment design where a decision maker can choose which points to sample to obtain an estimate $\hat{\beta}$ of the hidden parameter $\beta^{\star}$ of an underlying linear model. The key…

机器学习 · 统计学 2021-01-01 Xavier Fontaine , Pierre Perrault , Michal Valko , Vianney Perchet

This paper discusses the problem of determining optimal designs for regression models, when the observations are dependent and taken on an interval. A complete solution of this challenging optimal design problem is given for a broad class…

统计方法学 · 统计学 2015-02-25 Holger Dette , Andrey Pepelyshev , Anatoly Zhigljavsky

We study regression discontinuity designs in which many predetermined covariates, possibly much more than the number of observations, can be used to increase the precision of treatment effect estimates. We consider a two-step estimator…

计量经济学 · 经济学 2022-05-06 Alexander Kreiß , Christoph Rothe

Controlled experiments are widely used in many applications to investigate the causal relationship between input factors and experimental outcomes. A completely randomized design is usually used to randomly assign treatment levels to…

统计方法学 · 统计学 2026-05-12 Yiou Li , Lulu Kang , Xiao Huang

Two-phase sampling designs are frequently employed in epidemiological studies and large-scale health surveys. In such designs, certain variables are exclusively collected within a second-phase random subsample of the initial first-phase…

统计方法学 · 统计学 2024-03-25 Lingxiao Wang

To increase statistical efficiency in a randomized experiment, researchers often use stratification (i.e., blocking) in the design stage. However, conventional practices of stratification fail to exploit valuable information about the…

统计方法学 · 统计学 2025-10-28 Zikai Li

The goal of subsampling is to select an informative subset of all observations, when using the full data for statistical analysis is not viable. We construct locally $ D $-optimal subsampling designs under a Poisson regression model with a…

统计理论 · 数学 2024-03-28 Torsten Reuter , Rainer Schwabe

Supervised learning under measurement constraints is a common challenge in statistical and machine learning. In many applications, despite extensive design points, acquiring responses for all points is often impractical due to resource…

统计方法学 · 统计学 2025-03-19 Lin Wang

A fundamental issue in causal inference for Big Observational Data is confounding due to covariate imbalances between treatment groups. This can be addressed by designing the data prior to analysis. Existing design methods, developed for…

统计方法学 · 统计学 2022-03-17 Yumin Zhang , Arman Sabbaghi

Randomization is a basis for the statistical inference of treatment effects without strong assumptions on the outcome-generating process. Appropriately using covariates further yields more precise estimators in randomized experiments. R. A.…

统计理论 · 数学 2020-01-03 Xinran Li , Peng Ding

We consider the problem of how to assign treatment in a randomized experiment, in which the correlation among the outcomes is informed by a network available pre-intervention. Working within the potential outcome causal framework, we…

统计方法学 · 统计学 2017-05-19 Guillaume W. Basse , Edoardo M. Airoldi

Observational studies often benefit from an abundance of observational units. This can lead to studies that -- while challenged by issues of internal validity -- have inferences derived from sample sizes substantially larger than randomized…

统计方法学 · 统计学 2020-08-24 Rachael C. Aikens , Dylan Greaves , Michael Baiocchi

The analysis of screening experiments is often done in two stages, starting with factor selection via an analysis under a main effects model. The success of this first stage is influenced by three components: (1) main effect estimators'…

统计方法学 · 统计学 2024-03-19 Jonathan W. Stallrich , Michael McKibben

In the first stage of a two-stage study, the researcher uses a statistical model to impute the unobserved exposures. In the second stage, imputed exposures serve as covariates in epidemiological models. Imputation error in the first stage…

应用统计 · 统计学 2021-07-19 Ron Sarafian , Itai Kloog , Jonathan D. Rosenblatt

Many applications of machine learning methods involve an iterative protocol in which data are collected, a model is trained, and then outputs of that model are used to choose what data to consider next. For example, one data-driven approach…

We study regression discontinuity designs when covariates are included in the estimation. We examine local polynomial estimators that include discrete or continuous covariates in an additive separable way, but without imposing any…

计量经济学 · 经济学 2019-07-02 Sebastian Calonico , Matias D. Cattaneo , Max H. Farrell , Rocio Titiunik

The survey experiment is widely used in economics and social sciences to evaluate the effects of treatments or programs. In a standard population-based survey experiment, the experimenter randomly draws experimental units from a target…

统计方法学 · 统计学 2026-05-11 Pengfei Tian , Jiyang Ren , Yingying Ma
‹ 上一页 1 2 3 10 下一页 ›