English
Related papers

Related papers: Overall marginalized models for longitudinal zero-…

200 papers

A frequent challenge encountered with compositional ecological data is how to interpret and model data with a high proportion of zeros and $N$'s. Such data frequently occur in ecological applications where counts of species are collected…

Methodology · Statistics 2025-08-04 James Sweeney , John Haslett , Dipankar Bandyopadhyay , Michael Fop , Andrew C. Parnell

Latent class models have been successfully used to handle complex datasets in different disciplines. For longitudinal outcomes, we often get a trajectory of the outcome for each individual, and on that basis, we cluster them for a powerful…

Methodology · Statistics 2025-09-08 Chitradipa Chakraborty , Kiranmoy Das

We consider three new classes of exponential dispersion models of discrete probability distributions which are defined by specifying their variance functions in their mean value parameterization. In a previous paper (Bar-Lev and Ridder,…

Methodology · Statistics 2020-04-01 Shaul K. Bar-Lev , Ad Ridder

We propose a new methodology to detect zero-inflation and overdispersion based on the comparison of the expected sample extremes among convexly ordered distributions. The method is very flexible and includes tests for the proportion of…

Methodology · Statistics 2008-09-25 A. Baillo , J. Carcamo , J. R. Berrendero

In this paper, we propose two important extensions to cluster-weighted models (CWMs). First, we extend CWMs to have generalized cluster-weighted models (GCWMs) by allowing modeling of non-Gaussian distribution of the continuous covariates,…

Applications · Statistics 2019-01-01 Nikola Pocuca , Petar Jevtic , Paul D. McNicholas , Tatjana Miljkovic

The scan statistic is widely used in spatial cluster detection applications of inhomogeneous Poisson processes. However, real data may present substantial departure from the underlying Poisson process. One of the possible departures has to…

Methodology · Statistics 2013-11-19 André L. F. Cançado , Cibele Q. da-Silva , Michel F. da Silva

In a variety of application areas, there is a growing interest in analyzing high dimensional sparse count data, with sparsity exhibited by an over-abundance of zeros and small non-zero counts. Existing approaches for analyzing multivariate…

Methodology · Statistics 2016-04-15 Jyotishka Datta , David B. Dunson

Longitudinal studies of a binary outcome are common in the health, social, and behavioral sciences. In general, a feature of random effects logistic regression models for longitudinal binary data is that the marginal functional form, when…

Zero-inflated continuous data ubiquitously appear in many fields, in which lots of exactly zero-valued data are observed while others distribute continuously. Due to the mixed structure of discreteness and continuity in its distribution,…

Methodology · Statistics 2024-10-28 Keita Hamamoto

This paper proposes a computationally efficient Bayesian factor model for multiple grouped count data. Adopting the link function approach, the proposed model can capture the association within and between the at-risk probabilities and…

Methodology · Statistics 2024-05-13 Genya Kobayashi , Yuta Yamauchi

Claim frequency data in insurance records the number of claims on insurance policies during a finite period of time. Given that insurance companies operate with multiple lines of insurance business where the claim frequencies on different…

Applications · Statistics 2022-12-05 Pengcheng Zhang , David Pitt , Xueyuan Wu

This paper proposes a method for analyzing count time series with inflation or deflation of zeros. In particular, zero-modified Poisson and zero-modified negative binomial series with intensities generated by non-negative Markov sequences…

Methodology · Statistics 2021-07-06 N. Balakrishna , Muhammed Anvar , Bovas Abraham

Modeling sparse data such as microbiome and transcriptomics (RNA-seq) data is very challenging due to the exceeded number of zeros and skewness of the distribution. Many probabilistic models have been used for modeling sparse data,…

Methodology · Statistics 2021-12-30 Hani Aldirawi , Jie Yang

Commonly used methods to analyze incomplete longitudinal clinical trial data include complete case analysis (CC) and last observation carried forward (LOCF). However, such methods rest on strong assumptions, including missing completely at…

Statistics Theory · Mathematics 2007-06-13 Ivy Jansen , Caroline Beunckens , Geert Molenberghs , Geert Verbeke , Craig Mallinckrodt

Generalized linear models, such as logistic regression, are widely used to model the association between a treatment and a binary outcome as a function of baseline covariates. However, the coefficients of a logistic regression model…

Methodology · Statistics 2022-01-04 Jiaqi Yin , Sonia Markes , Thomas S. Richardson , Linbo Wang

Detecting associations between microbial compositions and sample characteristics is one of the most important tasks in microbiome studies. Most of the existing methods apply univariate models to single microbial species separately, with…

Microorganisms play critical roles in human health and disease. It is well known that microbes live in diverse communities in which they interact synergistically or antagonistically. Thus for estimating microbial associations with clinical…

Count data with excessive zeros are often encountered when modelling infectious disease occurrence. The degree of zero inflation can vary over time due to non-epidemic periods as well as by age group or region. The existing endemic-epidemic…

Methodology · Statistics 2023-09-14 Junyi Lu , Sebastian Meyer

The COVID-19 pandemic has had a substantial impact on hospital services, as many institutions have observed a surge in healthcare-associated infections (HAIs) despite heightened adherence to isolation protocols and hand hygiene. According…

Applications · Statistics 2023-10-12 Paulo Dourado , Antonio C. Pedroso-de-Lima , Francisco M. M. Rocha

We present a new modelling approach for longitudinal count data that is motivated by the increasing availability of longitudinal RNA-sequencing experiments. The distribution of RNA-seq counts typically exhibits overdispersion,…

Methodology · Statistics 2020-08-27 Mirko Signorelli , Pietro Spitali , Roula Tsonaka