English
Related papers

Related papers: Bias in multivariable Mendelian randomization stud…

200 papers

The association between visit-to-visit systolic blood pressure variability and cardiovascular events has recently received a lot of attention in the cardiovascular literature. But blood pressure variability is usually estimated on a…

Applications · Statistics 2019-01-25 Jessica K. Barrett , Raphael Huille , Richard Parker , Yuichiro Yano , Michael Griswold

Repeating an imperfect biomarker test based on an initial result can introduce bias and influence misclassification risk. For example, in some blood donation settings, blood donors' hemoglobin is remeasured when the initial measurement…

Applications · Statistics 2026-02-17 Supun Manathunga , Mart P. Janssen , Yu Luo , W. Alton Russell , Mart Pothast

In epidemiology, obtaining accurate individual exposure measurements can be costly and challenging. Thus, these measurements are often subject to error. Regression calibration with a validation study is widely employed as a study design and…

Methodology · Statistics 2026-02-24 Zexiang Li , Donna Spiegelman , Molin Wang , Zuoheng Wang , Xin Zhou

Multi-collinearity is a wide-spread phenomenon in modern statistical applications and when ignored, can negatively impact model selection and statistical inference. Classic tools and measures that were developed for "$n>p$" data are not…

Methodology · Statistics 2022-03-22 Wei Q. Deng , Radu V. Craiu , Lei Sun

Estimating the prevalence of a category in a population using imperfect measurement devices (diagnostic tests, classifiers, or large language models) is fundamental to science, public health, and online trust and safety. Standard approaches…

Artificial Intelligence · Computer Science 2026-04-24 Fridolin Linder , Thomas Leeper , Daniel Haimovich , Niek Tax , Lorenzo Perini , Milan Vojnovic

Most conventional risk analysis methods rely on a single best estimate of exposure per person which does not allow for adjustment for exposure-related uncertainty. Here, we propose a Bayesian model averaging method to properly quantify the…

Applications · Statistics 2020-04-07 Deukwoo Kwon , F. Owen Hoffman , Brian E. Moroz , Steven L. Simon

Due to concerns about parametric model misspecification, there is interest in using machine learning to adjust for confounding when evaluating the causal effect of an exposure on an outcome. Unfortunately, exposure effect estimators that…

Methodology · Statistics 2025-01-08 Oliver Dukes , Stijn Vansteelandt , David Whitney

To date, we have seen the emergence of a large literature on multivariate disease mapping. That is, incidence of (or mortality from) multiple diseases is recorded at the scale of areal units where incidence (mortality) across the diseases…

Methodology · Statistics 2026-04-17 Garazi Retegui , María Dolores Ugarte , Jaione Etxeberria , Alan E. Gelfand

There is growing interest in Bayesian clinical trial designs with informative prior distributions, e.g. for extrapolation of adult data to pediatrics, or use of external controls. While the classical type I error is commonly used to…

Methodology · Statistics 2023-09-06 Nicky Best , Maxine Ajimi , Beat Neuenschwander , Gaelle Saint-Hilary , Simon Wandel

In epidemiological cohort studies, the relative risk (also known as risk ratio) is a major measure of association to summarize the results of two treatments or exposures. Generally, it measures the relative change in disease risk as a…

Methodology · Statistics 2022-07-05 Gopal Nath , Krishna K. Saha , Suojin Wang

There have been reports of correlation between estimates of prevalence and test accuracy across studies included in diagnostic meta-analyses. It has been hypothesized that this unexpected association arises because of certain biases…

Methodology · Statistics 2025-08-15 Yang Lu , Robert Platt , Nandini Dendukuri

Mendelian Randomization (MR) is a popular method in epidemiology and genetics that uses genetic variation as instrumental variables for causal inference. Existing MR methods usually assume most genetic variants are valid instrumental…

Applications · Statistics 2022-06-15 Daniel Iong , Qingyuan Zhao , Yang Chen

Typically, a randomized experiment is designed to test a hypothesis about the average treatment effect and sometimes hypotheses about treatment effect variation. The results of such a study may then be used to inform policy and practice for…

Methodology · Statistics 2026-05-01 Elizabeth Tipton , Michalis Mamakos

A common approach to statistical learning with big-data is to randomly split it among $m$ machines and learn the parameter of interest by averaging the $m$ individual estimates. In this paper, focusing on empirical risk minimization, or…

Machine Learning · Statistics 2016-06-14 Jonathan Rosenblatt , Boaz Nadler

Data from both a randomized trial and an observational study are sometimes simultaneously available for evaluating the effect of an intervention. The randomized data typically allows for reliable estimation of average treatment effects but…

Methodology · Statistics 2021-12-01 David Cheng , Tianxi Cai

Estimating the causal effect of a treatment or health policy with observational data can be challenging due to an imbalance of and a lack of overlap between treated and control covariate distributions. In the presence of limited overlap,…

Methodology · Statistics 2025-03-24 Martha Barnard , Jared D. Huling , Julian Wolfson

Longitudinal data tracking repeated measurements on individuals are highly valued for research because they offer controls for unmeasured individual heterogeneity that might otherwise bias results. Random effects or mixed models approaches,…

Applications · Statistics 2009-09-29 J. R. Lockwood , Daniel F. McCaffrey

Many lifestyle intervention trials depend on collecting self-reported outcomes, like dietary intake, to assess the intervention's effectiveness. Self-reported outcome measures are subject to measurement error, which could impact treatment…

Methodology · Statistics 2020-04-06 Benjamin Ackerman , Juned Siddique , Elizabeth A. Stuart

Nonparametric two-sample tests such as the Maximum Mean Discrepancy (MMD) are often used to detect differences between two distributions in machine learning applications. However, the majority of existing literature assumes that error-free…

Machine Learning · Statistics 2023-08-08 Ron Nafshi , Maggie Makar

You measure the value of a quantity x for a number of systems (cells, molecules, people, chunks of metal, DNA vectors, etc.). You repeat the whole set of measures in different occasions or assays, which you try to design as equal to one…

Quantitative Methods · Quantitative Biology 2013-11-04 Pablo Echenique-Robba , María Alejandra Nelo-Bazán , José A. Carrodeguas