中文
相关论文

相关论文: Bootstrapping Clustered Data in R using lmeresampl…

200 篇论文

Randomized response (RR) designs are used to collect response data about sensitive behaviors (e.g., criminal behavior, sexual desires). The modeling of RR data is more complex, since it requires a description of the RR process. For the…

统计方法学 · 统计学 2021-06-21 Jean-Paul Fox , Konrad Klotzke , Duco Veen

We investigate the performance of model based bootstrap methods for constructing point-wise confidence intervals around the survival function with interval censored data. We show that bootstrapping from the nonparametric maximum likelihood…

统计方法学 · 统计学 2013-12-24 Bodhisattva Sen , Gongjun Xu

We consider penalized extremum estimation of a high-dimensional, possibly nonlinear model that is sparse in the sense that most of its parameters are zero but some are not. We use the SCAD penalty function, which provides model selection…

计量经济学 · 经济学 2024-02-23 Joel L. Horowitz , Ahnaf Rafi

Estimating sample size and statistical power is an essential part of a good study design. This R package allows users to conduct power analysis based on Monte Carlo simulations in settings in which consideration of the correlations between…

统计方法学 · 统计学 2024-04-16 Phuc H. Nguyen , Stephanie M. Engel , Amy H. Herring

Bootstrap methods have long been the cornerstone of ensemble learning in machine learning. This paper presents a theoretical analysis of bootstrap techniques applied to the Least Square Support Vector Machine (LSSVM) ensemble in the context…

Generalized additive models (GAMs) play an important role in modeling and understanding complex relationships in modern applied statistics. They allow for flexible, data-driven estimation of covariate effects. Yet researchers often have a…

统计方法学 · 统计学 2014-11-10 Benjamin Hofner , Thomas Kneib , Torsten Hothorn

Pre-training datasets are typically collected from web content and lack inherent domain divisions. For instance, widely used datasets like Common Crawl do not include explicit domain labels, while manually curating labeled datasets such as…

In this note we propose a vectorized implementation of the non-parametric bootstrap for statistics based on sample moments. Basically, we adopt the multinomial sampling formulation of the non-parametric bootstrap, and compute bootstrap…

统计计算 · 统计学 2014-12-12 E. Chaibub Neto

Measurement error and missing data in variables used in statistical models are common, and can at worst lead to serious biases in analyses if they are ignored. Yet, these problems are often not dealt with adequately, presumably in part…

统计方法学 · 统计学 2024-06-13 Emma Skarstein , Stefanie Muff

The bootstrap is a popular and powerful method for assessing precision of estimators and inferential methods. However, for massive datasets which are increasingly prevalent, the bootstrap becomes prohibitively costly in computation and its…

统计方法学 · 统计学 2015-08-06 Srijan Sengupta , Stanislav Volgushev , Xiaofeng Shao

Empirical best linear unbiased prediction (EBLUP) method uses a linear mixed model in combining information from different sources of information. This method is particularly useful in small area problems. The variability of an EBLUP is…

统计理论 · 数学 2008-12-18 Snigdhansu Chatterjee , Partha Lahiri , Huilin Li

Clustered sampling is prevalent in empirical regression discontinuity (RD) designs, but it has not received much attention in the theoretical literature. In this paper, we introduce a general model-based framework for such settings and…

计量经济学 · 经济学 2026-03-20 Claudia Noack , Tomasz Olma , Christoph Rothe

Meta-analysis combines pertinent information from existing studies to provide an overall estimate of population parameters/effect sizes, as well as to quantify and explain the differences between studies. However, testing the between-study…

统计方法学 · 统计学 2020-11-13 Han Du , Ge Jiang , Zijun Ke

This paper reports on application of bootstrap nonlinear regression method to a design of an experiment dataset with fewer experimental runs. Design with desired properties was augmented and verified using graphical techniques. The…

Nested-error regression models are widely used for analyzing clustered data. For example, they are often applied to two-stage sample surveys, and in biology and econometrics. Prediction is usually the main goal of such analyses, and…

统计理论 · 数学 2007-06-13 Peter Hall , Tapabrata Maiti

Reliable forward uncertainty quantification in engineering requires methods that account for aleatory and epistemic uncertainties. In many applications, epistemic effects arising from uncertain parameters and model form dominate prediction…

计算工程、金融与科学 · 计算机科学 2025-12-18 Akash Yadav , Ruda Zhang

Linear mixed models are widely used to analyze non-independent data, but inference for fixed effects can be unreliable under misspecification of the random-effects distribution, inaccurate Fisher information estimation, or convergence…

统计方法学 · 统计学 2026-05-01 Angela Andreella , Livio Finos

We investigate the problem of computing a nested expectation of the form $\mathbb{P}[\mathbb{E}[X|Y] \!\geq\!0]\!=\!\mathbb{E}[\textrm{H}(\mathbb{E}[X|Y])]$ where $\textrm{H}$ is the Heaviside function. This nested expectation appears, for…

计算金融 · 定量金融 2019-02-15 Michael B. Giles , Abdul-Lateef Haji-Ali

The bootstrap provides a simple and powerful means of assessing the quality of estimators. However, in settings involving large datasets---which are increasingly prevalent---the computation of bootstrap-based quantities can be prohibitively…

统计方法学 · 统计学 2012-06-29 Ariel Kleiner , Ameet Talwalkar , Purnamrita Sarkar , Michael I. Jordan

AI/ML methods are increasingly used in economics to generate binary variables (or labels) via classification algorithms. When these generated variables are included as covariates in regressions, even small misclassification errors can…

计量经济学 · 经济学 2026-04-28 Timothy Christensen , Silvia Goncalves , Benoit Perron