English

Struggles with Survey Weighting and Regression Modeling

Methodology 2007-11-06 v1

Abstract

The general principles of Bayesian data analysis imply that models for survey responses should be constructed conditional on all variables that affect the probability of inclusion and nonresponse, which are also the variables used in survey weighting and clustering. However, such models can quickly become very complicated, with potentially thousands of poststratification cells. It is then a challenge to develop general families of multilevel probability models that yield reasonable Bayesian inferences. We discuss in the context of several ongoing public health and social surveys. This work is currently open-ended, and we conclude with thoughts on how research could proceed to solve these problems.

Keywords

Cite

@article{arxiv.0710.5005,
  title  = {Struggles with Survey Weighting and Regression Modeling},
  author = {Andrew Gelman},
  journal= {arXiv preprint arXiv:0710.5005},
  year   = {2007}
}

Comments

This paper commented in: [arXiv:0710.5009], [arXiv:0710.5012], [arXiv:0710.5013], [arXiv:0710.5015], [arXiv:0710.5016]. Rejoinder in [arXiv:0710.5019]. Published in at http://dx.doi.org/10.1214/088342306000000691 the Statistical Science (http://www.imstat.org/sts/) by the Institute of Mathematical Statistics (http://www.imstat.org)

R2 v1 2026-06-21T09:36:42.085Z