English
Related papers

Related papers: Zero & $N$-inflated overdispersed binomial models …

200 papers

Many causal estimands are only partially identifiable since they depend on the unobservable joint distribution between potential outcomes. Stratification on pretreatment covariates can yield sharper bounds; however, unless the covariates…

Econometrics · Economics 2024-11-19 Wenlong Ji , Lihua Lei , Asher Spector

Binomial data with unknown sizes often appear in biological and medical sciences and are usually overdispersed. All previous methods used parametric models and only considered overdispersion due to the variation of sizes. The proposed…

Statistics Theory · Mathematics 2007-06-13 Wei Zhang

We consider a general, neutral, dynamical model of biodiversity. Individuals have i.i.d. lifetime durations, which are not necessarily exponentially distributed, and each individual gives birth independently at constant rate \lambda. We…

Populations and Evolution · Quantitative Biology 2010-09-02 Amaury Lambert

Multi-dimensional data frequently occur in many different fields, including risk management, insurance, biology, environmental sciences, and many more. In analyzing multivariate data, it is imperative that the underlying modelling…

Methodology · Statistics 2025-06-23 Orla A. Murphy , Juliana Schulz

Integrative analysis of datasets generated by multiple cohorts is a widely-used approach for increasing sample size, precision of population estimators, and generalizability of analysis results in epidemiological studies. However, often…

Determining spatial distributions of species and communities are key objectives of ecology and conservation. Joint species distribution models use multi-species detection-nondetection data to estimate species and community distributions.…

Applications · Statistics 2022-12-15 Jeffrey W. Doser , Andrew O. Finley , Sudipto Banerjee

Commonly in biomedical research, studies collect data in which an outcome measure contains informative excess zeros; for example when observing the burden of neuritic plaques in brain pathology studies, those who show none contribute to our…

Methodology · Statistics 2018-01-19 Matthew Goodman , Lori Chibnik , Tianxi Cai

Model selection aims to identify a sufficiently well performing model that is possibly simpler than the most complex model among a pool of candidates. However, the decision-making process itself can inadvertently introduce non-negligible…

Methodology · Statistics 2024-08-08 Yann McLatchie , Aki Vehtari

We introduce negative binomial matrix factorization (NBMF), a matrix factorization technique specially designed for analyzing over-dispersed count data. It can be viewed as an extension of Poisson matrix factorization (PF) perturbed by a…

Machine Learning · Computer Science 2018-01-08 Olivier Gouvert , Thomas Oberlin , Cédric Févotte

This paper develops forecasting methodology and application of new classes of dynamic models for time series of non-negative counts. Novel univariate models synthesise dynamic generalized linear models for binary and conditionally Poisson…

Methodology · Statistics 2022-06-07 Lindsay Berry , Mike West

Parallel recordings of neural spike counts have revealed the existence of context-dependent noise correlations in neural populations. Theories of population coding have also shown that such correlations can impact the information encoded by…

Machine Learning · Computer Science 2025-07-30 Sacha Sokoloski , Ruben Coen-Cagli

Zero-inflated explanatory variables are common in fields such as ecology and finance. In this paper we address the problem of having excess of zero values in some explanatory variables which are subject to multioutcome lasso-regularized…

Methodology · Statistics 2021-09-13 Jyrki Möttönen , Tero Lähderanta , Janne Salonen , Mikko J. Sillanpää

The rapid expansion of citizen science initiatives has led to a significant growth of biodiversity databases, and particularly presence-only (PO) observations. PO data are invaluable for understanding species distributions and their…

Machine Learning · Computer Science 2026-01-14 Maxime Ryckewaert , Diego Marcos , Christophe Botella , Maximilien Servajean , Pierre Bonnet , Alexis Joly

The inference of the causal relationship between a pair of observed variables is a fundamental problem in science, and most existing approaches are based on one single causal model. In practice, however, observations are often collected…

Machine Learning · Statistics 2018-11-13 Shoubo Hu , Zhitang Chen , Vahid Partovi Nia , Laiwan Chan , Yanhui Geng

This work is focussed on the inversion task of inferring the distribution over parameters of interest leading to multiple sets of observations. The potential to solve such distributional inversion problems is driven by increasing…

Machine Learning · Statistics 2026-05-06 Arnaud Vadeboncoeur , Mark Girolami , Andrew M. Stuart

We developed a statistical theory of zero-count-detector (ZCD), which is defined as a zero-class Poisson under conditions outlined in the paper. ZCD is often encountered in the studies of rare events in physics, health physics, and many…

Applications · Statistics 2024-12-30 Thomas M. Semkow

Causal mediation analysis is an important statistical tool to quantify effects transmitted by intermediate variables from a cause to an outcome. There is a gap in mediation analysis methods to handle mixture mediator data that are…

Methodology · Statistics 2025-07-22 Meilin Jiang , Seonjoo Lee , A. James O'Malley , Pengfei Li , Zhigang Li

By developing data augmentation methods unique to the negative binomial (NB) distribution, we unite seemingly disjoint count and mixture models under the NB process framework. We develop fundamental properties of the models and derive…

Machine Learning · Statistics 2013-02-18 Mingyuan Zhou , Lawrence Carin

Biological sequencing data consist of read counts, e.g. of specified taxa and often exhibit sparsity (zero-count inflation) and overdispersion (extra-Poisson variability). As most sequencing techniques provide an arbitrary total count,…

Applications · Statistics 2024-07-01 Noora Kartiosuo , Jaakko Nevalainen , Olli Raitakari , Katja Pahkala , Kari Auranen

Bayesian hierarchical models are commonly employed for inference in count datasets, as they account for multiple levels of variation by incorporating prior distributions for parameters at different levels. Examples include Beta-Binomial,…

Methodology · Statistics 2024-11-04 Yuexi Wang , Nicholas G. Polson
‹ Prev 1 8 9 10 Next ›