English
Related papers

Related papers: A prediction interval for the population-wise erro…

200 papers

We re-investigate the asymptotic properties of the traditional OLS (pooled) estimator, $\hat{\beta} _P$, in the context of cluster dependence. The present study considers various scenarios under various restrictions on the cluster sizes and…

Methodology · Statistics 2025-01-31 Subhodeep Dey , Gopal K. Basak , Samarjit Das

This study aims to evaluate the performance of power in the likelihood ratio test for changepoint detection by bootstrap sampling, and proposes a hypothesis test based on bootstrapped confidence interval lengths. Assuming i.i.d normally…

Methodology · Statistics 2020-11-10 Ryan Chen , Javier Cabrera

The Lasso is one of the most ubiquitous methods for variable selection in high-dimensional linear regression and has been studied extensively under different regimes. In a particular asymptotic setup entailing $n/p\to \text{constant}$, an…

Statistics Theory · Mathematics 2026-02-10 Lina Hidmi , Asaf Weinstein

Propensity scores are commonly used to estimate treatment effects from observational data. We argue that the probabilistic output of a learned propensity score model should be calibrated -- i.e., a predictive treatment probability of 90%…

Methodology · Statistics 2024-06-06 Shachi Deshpande , Volodymyr Kuleshov

Publication bias is a major concern in conducting systematic reviews and meta-analyses. Various sensitivity analysis or bias-correction methods have been developed based on selection models and they have some advantages over the widely used…

Methodology · Statistics 2021-09-28 Ao Huang , Kosuke Morikawa , Tim Friede , Satoshi Hattori

In this work, we revisit the problem of active sequential prediction-powered mean estimation, where at each round one must decide the query probability of the ground-truth label upon observing the covariates of a sample. Furthermore, if the…

Machine Learning · Statistics 2026-04-21 Maria-Eleni Sfyraki , Jun-Kun Wang

Doubly robust (DR) estimation is a crucial technique in causal inference and missing data problems. We propose a novel Propensity score Augmentved Doubly robust (PAD) estimator to enhance the commonly used DR estimator for average treatment…

Methodology · Statistics 2023-04-18 Liangbo Lyu , Molei Liu

Suppose one has a collection of parameters indexed by a (possibly infinite dimensional) set. Given data generated from some distribution, the objective is to estimate the maximal parameter in this collection evaluated at this distribution.…

Methodology · Statistics 2016-05-26 Alexander R. Luedtke , Mark J. van der Laan

Inverse probability weights are commonly used in epidemiology to estimate causal effects in observational studies. Researchers can typically focus on either the average treatment effect or the average treatment effect on the treated with…

Methodology · Statistics 2022-10-05 Eli Ben-Michael , Luke Keele

Data following an interval structure are increasingly prevalent in many scientific applications. In medicine, clinical events are often monitored between two clinical visits, making the exact time of the event unknown and generating…

Methodology · Statistics 2025-04-01 Carlos García Meixide , Michael R. Kosorok , Marcos Matabuena

The prediction interval has been increasingly used in meta-analyses as a useful measure for assessing the magnitude of treatment effect and between-studies heterogeneity. In calculations of the prediction interval, although the…

Methodology · Statistics 2021-07-14 Yuta Hamaguchi , Hisashi Noma , Kengo Nagashima , Tomohide Yamada , Toshi A. Furukawa

Most clinical trials conducted in drug development contain multiple endpoints in order to collectively assess the intended effects of the drug on various disease characteristics. Focusing on the estimation of the global win probability,…

Methodology · Statistics 2024-04-09 Di Shu , Guangyong Zou

Causal inference is crucial for understanding the true impact of interventions, policies, or actions, enabling informed decision-making and providing insights into the underlying mechanisms that shape our world. In this paper, we establish…

Methodology · Statistics 2024-03-26 Jingyue Huang , Changbao Wu , Leilei Zeng

Data-driven decision making frequently relies on predicting counterfactual outcomes. In practice, researchers commonly train counterfactual prediction models on a source dataset to inform decisions on a possibly separate target population.…

Machine Learning · Statistics 2026-04-07 Keith Barnatchez , Kevin P. Josey , Rachel C. Nethery , Giovanni Parmigiani

In an empirical Bayes analysis, we use data from repeated sampling to imitate inferences made by an oracle Bayesian with extensive knowledge of the data-generating distribution. Existing results provide a comprehensive characterization of…

Methodology · Statistics 2021-09-09 Nikolaos Ignatiadis , Stefan Wager

This paper extends my research applying statistical decision theory to treatment choice with sample data, using maximum regret to evaluate the performance of treatment rules. The specific new contribution is to study as-if optimization…

Econometrics · Economics 2021-10-05 Charles F. Manski

Inverse probability weighting (IPW) is widely used in many areas when data are subject to unrepresentativeness, missingness, or selection bias. An inevitable challenge with the use of IPW is that the IPW estimator can be remarkably unstable…

Methodology · Statistics 2021-11-29 Yukun Liu , Yan Fan

Many partial identification problems can be characterized by the optimal value of a function over a set where both the function and set need to be estimated by empirical data. Despite some progress for convex problems, statistical inference…

Methodology · Statistics 2022-08-31 Matthew Tudball , Rachael Hughes , Kate Tilling , Jack Bowden , Qingyuan Zhao

Panels with large time $(T)$ and cross-sectional $(N)$ dimensions are a key data structure in social sciences and other fields. A central question in panel data analysis is whether to pool data across individuals or to estimate separate…

Methodology · Statistics 2025-12-18 Tim Kutta , Martin Schumann , Holger Dette

Many trials are designed to collect outcomes at or around pre-specified times after randomization. If there is variability in the times when participants are actually assessed, this can pose a challenge to learning the effect of treatment,…