中文
相关论文

相关论文: Prediction intervals for overdispersed Poisson dat…

200 篇论文

Survival analysis encompasses a broad range of methods for analyzing time-to-event data, with one key objective being the comparison of survival curves across groups. Traditional approaches for identifying clusters of survival curves often…

统计方法学 · 统计学 2025-12-19 Nora M. Villanueva , Marta Sestelo , Luis Meira-Machado

Mixture models are commonly used in applications with heterogeneity and overdispersion in the population, as they allow the identification of subpopulations. In the Bayesian framework, this entails the specification of suitable prior…

统计方法学 · 统计学 2023-06-21 Andrea Cremaschi , Timothy M. Wertz , Maria De Iorio

Violation of the assumptions underlying classical (Gaussian) limit theory often yields unreliable statistical inference. This paper shows that the bootstrap can detect such violations by delivering simple and powerful diagnostic tests that…

计量经济学 · 经济学 2025-10-09 Giuseppe Cavaliere , Luca Fanelli , Iliyan Georgiev

The tau statistic $\tau$ uses geolocation and, usually, symptom onset time to assess global spatiotemporal clustering from epidemiological data. We test different factors that could affect graphical hypothesis tests of clustering or bias…

The problem of multi-hypothesis testing with controlled sensing of observations is considered. The distribution of observations collected under each control is assumed to follow a single-parameter exponential family distribution. The goal…

统计理论 · 数学 2019-10-29 Aditya Deshmukh , Srikrishna Bhashyam , Venugopal V. Veeravalli

To make informative public policy decisions in battling the ongoing COVID-19 pandemic, it is important to know the disease prevalence in a population. There are two intertwined difficulties in estimating this prevalence based on testing…

统计方法学 · 统计学 2020-12-01 Bryan Cai , John P. A. Ioannidis , Eran Bendavid , Lu Tian

Cross-validation is a widely-used technique to estimate prediction error, but its behavior is complex and not fully understood. Ideally, one would like to think that cross-validation estimates the prediction error for the model at hand, fit…

统计方法学 · 统计学 2024-03-12 Stephen Bates , Trevor Hastie , Robert Tibshirani

For discrete-valued time series, predictive inference cannot be implemented through the construction of prediction intervals to some predetermined coverage level, as this is the case for real-valued time series. To address this problem, we…

统计方法学 · 统计学 2025-07-23 Maxime Faymonville , Carsten Jentsch , Efstathios Paparoditis

We investigate the performance of model based bootstrap methods for constructing point-wise confidence intervals around the survival function with interval censored data. We show that bootstrapping from the nonparametric maximum likelihood…

统计方法学 · 统计学 2013-12-24 Bodhisattva Sen , Gongjun Xu

Estimating causal effects from large experimental and observational data has become increasingly prevalent in both industry and research. The bootstrap is an intuitive and powerful technique used to construct standard errors and confidence…

统计方法学 · 统计学 2023-02-07 Matthew Kosko , Lin Wang , Michele Santacatterina

One of the most commonly used methods for forming confidence intervals for statistical inference is the empirical bootstrap, which is especially expedient when the limiting distribution of the estimator is unknown. However, despite its…

统计理论 · 数学 2020-11-24 Morgane Austern , Vasilis Syrgkanis

Linear mixed effects are considered excellent predictors of cluster-level parameters in various domains. However, previous work has shown that their performance can be seriously affected by departures from modelling assumptions. Since the…

统计方法学 · 统计学 2022-07-27 Katarzyna Reluga , Stefan Sperlich

Many of the data, particularly in medicine and disease mapping are count. Indeed, the under or overdispersion problem in count data distrusts the performance of the classical Poisson model. For taking into account this problem, in this…

统计方法学 · 统计学 2021-05-19 Mahsa Nadifar , Hossein Baghishani , Thomas Kneib , Afshin Fallah

In the practical industry, the most commonly used application of statistical analysis for monitoring the process mean is the control chart. Control charts are generated based on the presumption that we have a sample from a stable process.…

最优化与控制 · 数学 2025-10-07 Fahad Rafique , Saadia Masood , Shabbir Ahmad , Sadaf Amin

Maximum likelihood estimates (MLEs) are asymptotically normally distributed, and this property is used in meta-analyses to test the heterogeneity of estimates, either for a single cluster or for several sub-groups. More recently, MLEs for…

统计理论 · 数学 2022-02-28 Anthony J. Webster

Survival analysis models the distribution of time until an event of interest, such as discharge from the hospital or admission to the ICU. When a model's predicted number of events within any time interval is similar to the observed number,…

机器学习 · 计算机科学 2021-01-15 Mark Goldstein , Xintian Han , Aahlad Puli , Adler J. Perotte , Rajesh Ranganath

Cluster analysis of biological samples using gene expression measurements is a common task which aids the discovery of heterogeneous biological sub-populations having distinct mRNA profiles. Several model-based clustering algorithms have…

统计方法学 · 统计学 2012-01-30 Alberto Cozzini , Ajay Jasra , Giovanni Montana

Semi-supervised learning by self-training heavily relies on pseudo-label selection (PLS). The selection often depends on the initial model fit on labeled data. Early overfitting might thus be propagated to the final model by selecting…

机器学习 · 统计学 2023-06-27 Julian Rodemann , Jann Goschenhofer , Emilio Dorigatti , Thomas Nagler , Thomas Augustin

Clustering is a fundamental tool in statistical machine learning in the presence of heterogeneous data. Most recent results focus primarily on optimal mislabeling guarantees when data are distributed around centroids with sub-Gaussian…

统计理论 · 数学 2024-10-24 Soham Jana , Jianqing Fan , Sanjeev Kulkarni

Medical machine learning algorithms are typically evaluated based on accuracy vs. a clinician-defined ground truth, a reasonable initial choice since trained clinicians are usually better classifiers than ML models. However, this metric…

机器学习 · 计算机科学 2024-05-21 Charles B. Delahunt , Courosh Mehanian , Matthew P. Horning