中文
相关论文

相关论文: Bootstrapping Clustered Data in R using lmeresampl…

200 篇论文

We describe a network clustering framework, based on finite mixture models, that can be applied to discrete-valued networks with hundreds of thousands of nodes and billions of edge variables. Relative to other recent model-based clustering…

统计计算 · 统计学 2020-03-13 Duy Q. Vu , David R. Hunter , Michael Schweinberger

Recently there has been much interest in data that, in statistical language, may be described as having a large crossed and severely unbalanced random effects structure. Such data sets arise for recommender engines and information retrieval…

应用统计 · 统计学 2007-12-18 Art B. Owen

The item response model in latent space (LSIRM; Jeon et al., 2021) uncovers unobserved interactions between respondents and items in the item response data by embedding both in a shared latent metric space. The R package lsirm12pl…

统计方法学 · 统计学 2025-03-11 Dongyoung Go , Gwanghee Kim , Jina Park , Junyong Park , Minjeong Jeon , Ick Hoon Jin

We implemented several multilabel classification algorithms in the machine learning package mlr. The implemented methods are binary relevance, classifier chains, nested stacking, dependent binary relevance and stacking, which can be used…

机器学习 · 统计学 2023-11-09 Philipp Probst , Quay Au , Giuseppe Casalicchio , Clemens Stachl , Bernd Bischl

We investigate the finite sample performance of sample splitting, cross-fitting and averaging for the estimation of the conditional average treatment effect. Recently proposed methods, so-called meta-learners, make use of machine learning…

统计方法学 · 统计学 2020-08-27 Daniel Jacob

A model based clustering procedure for data of mixed type, clustMD, is developed using a latent variable model. It is proposed that a latent variable, following a mixture of Gaussian distributions, generates the observed data of mixed type.…

统计方法学 · 统计学 2015-11-06 Damien McParland , Isobel Claire Gormley

In the context of paid research studies and clinical trials, budget considerations often require patient sampling from available populations which comes with inherent constraints. We introduce the R package CDsampling, which is the first to…

统计计算 · 统计学 2025-09-30 Yifei Huang , Liping Tong , Jie Yang

We present csSampling, an R package for estimation of Bayesian models for data collected from complex survey samples. csSampling combines functionality from the probabilistic programming language Stan (via the rstan and brms R packages) and…

统计计算 · 统计学 2023-08-15 Ryan Hornby , Matthew R. Williams , Terrance D. Savitsky , Mahmoud Elkasabi

Finite-sample bias is a pervasive challenge in the estimation of structural equation models (SEMs), especially when sample sizes are small or measurement reliability is low. A range of methods have been proposed to improve finite-sample…

统计方法学 · 统计学 2026-03-30 Haziq Jamil , Yves Rosseel , Oliver Kemp , Ioannis Kosmidis

In the fields of clinical trials, biomedical surveys, marketing, banking, with dichotomous response variable, the logistic regression is considered as an alternative convenient approach to linear regression. In this paper, we develop a…

统计理论 · 数学 2025-12-29 Debraj Das , Priyam Das

Bootstrap is commonly used as a tool for non-parametric statistical inference to estimate meaningful parameters in Variable Selection Models. However, for massive dataset that has exponential growth rate, the computation of Bootstrap…

统计计算 · 统计学 2016-12-26 Zhibing He , Yichen Qin , Ben-Chang Shia , Yang Li

In this article, we propose a penalized clustering method for large scale data with multiple covariates through a functional data approach. In the proposed method, responses and covariates are linked together through nonparametric…

统计方法学 · 统计学 2008-01-17 Ping Ma , Wenxuan Zhong

Meta-analyses require an effect-size estimate and its corresponding sampling variance from primary studies. In some cases, estimators for the sampling variance of a given effect size statistic may not exist, necessitating the derivation of…

The usage of psychological networks that conceptualize psychological behavior as a complex interplay of psychological and other components has gained increasing popularity in various fields of psychology. While prior publications have…

应用统计 · 统计学 2017-01-23 Sacha Epskamp , Denny Borsboom , Eiko I. Fried

The asymptotic validity of a resampling method for two sequential processes constructed from non-degenerate $U$-statistics is established under mixing conditions. The resampling schemes, referred to as {\em dependent multiplier bootstraps},…

统计理论 · 数学 2015-05-29 Axel Bücher , Ivan Kojadinovic

Binned scatter plots are a powerful statistical tool for empirical work in the social, behavioral, and biomedical sciences. Available methods rely on a quantile-based partitioning estimator of the conditional mean regression function to…

统计方法学 · 统计学 2024-07-23 Matias D. Cattaneo , Richard K. Crump , Max H. Farrell , Yingjie Feng

In high-dimensional time series, the component processes are often assembled into a matrix to display their interrelationship. We focus on detecting mean shifts with unknown change point locations in these matrix time series. Series that…

统计方法学 · 统计学 2024-07-16 Xinyu Zhang , Kung-Sik Chan

The increasing availability of time --and space-- resolved data describing human activities and interactions gives insights into both static and dynamic properties of human behavior. In practice, nevertheless, real-world datasets can often…

The overwhelming majority of empirical research that uses cluster-robust inference assumes that the clustering structure is known, even though there are often several possible ways in which a dataset could be clustered. We propose two tests…

计量经济学 · 经济学 2023-03-14 James G. MacKinnon , Morten Ørregaard Nielsen , Matthew D. Webb

This paper investigates the use of bootstrap-based bias correction of semi-parametric estimators of the long memory parameter in fractionally integrated processes. The re-sampling method involves the application of the sieve bootstrap to…

统计方法学 · 统计学 2014-02-28 D. S. Poskitt , Gael M. Martin , Simone D. Grose