中文
相关论文

相关论文: Zero-inflated Poisson Factor Model with Applicatio…

200 篇论文

We study estimation and testing in the Poisson regression model with noisy high dimensional covariates, which has wide applications in analyzing noisy big data. Correcting for the estimation bias due to the covariate noise leads to a…

统计理论 · 数学 2023-01-03 Fei Jiang , Yeqing Zhou , Jianxuan Liu , Yanyuan Ma

Fractional equations have become the model of choice in several applications where heterogeneities at the microstructure result in anomalous diffusive behavior at the macroscale. In this work we introduce a new fractional operator…

数值分析 · 数学 2021-01-29 Marta D'Elia , Christian Glusa

Bayesian factor models are widely used for dimensionality reduction and pattern discovery in high-dimensional datasets across diverse fields. These models typically focus on imposing priors on factor loading to induce sparsity and improve…

统计方法学 · 统计学 2025-04-08 Yingjie Huang , Dafne Zorzetto , Roberta De Vito

Modeling sparse count data, which arise across numerous scientific fields, presents significant statistical challenges. This chapter addresses these challenges in the context of infectious disease prediction, with a focus on predicting…

机器学习 · 统计学 2026-02-05 Edwin Fong , Lancelot F. James , Juho Lee

We introduce negative binomial matrix factorization (NBMF), a matrix factorization technique specially designed for analyzing over-dispersed count data. It can be viewed as an extension of Poisson matrix factorization (PF) perturbed by a…

机器学习 · 计算机科学 2018-01-08 Olivier Gouvert , Thomas Oberlin , Cédric Févotte

Generalized linear models (GLMs) using a regression procedure to fit relationships between predictor and target variables are widely used in automobile insurance data. Here, in the process of ratemaking and in order to compute the premiums…

应用统计 · 统计学 2016-06-02 J. M. Pérez-Sánchez , E. Gómez-Déniz

Confirmatory factor analysis (CFA) is a statistical method for identifying and confirming the presence of latent factors among observed variables through the analysis of their covariance structure. Compared to alternative factor models, CFA…

统计方法学 · 统计学 2024-10-08 Yifan Yang , Tianzhou Ma , Chuan Bi , Shuo Chen

We introduce a novel approach to compositional data analysis based on $L^{\infty}$-normalization, addressing challenges posed by zero-rich high-throughput data. Traditional methods like Aitchison's transformations require excluding zeros,…

统计计算 · 统计学 2025-03-28 Pawel Gajer , Jacques Ravel

In recent years microbiome studies have become increasingly prevalent and large-scale. Through high-throughput sequencing technologies and well-established analytical pipelines, relative abundance data of operational taxonomic units and…

统计方法学 · 统计学 2022-05-12 Yan Li , Gen Li , Kun Chen

Claim frequency data in insurance records the number of claims on insurance policies during a finite period of time. Given that insurance companies operate with multiple lines of insurance business where the claim frequencies on different…

应用统计 · 统计学 2022-12-05 Pengcheng Zhang , David Pitt , Xueyuan Wu

Mediation analysis seeks to understand the mechanism by which a treatment affects an outcome. Count or zero-inflated count outcome are common in many studies in which mediation analysis is of interest. For example, in dental studies,…

统计方法学 · 统计学 2016-07-12 Zijian Guo , Dylan S. Small , Stuart A. Gansky , Jing Cheng

Functional principal component analysis (FPCA) is a fundamental tool and has attracted increasing attention in recent decades, while existing methods are restricted to data with a single or finite number of random functions (much smaller…

统计方法学 · 统计学 2021-01-22 Xiaoyu Hu , Fang Yao

It is more and more common to explore the genome at diverse levels and not only at a single omic level. Through integrative statistical methods, omics data have the power to reveal new biological processes, potential biomarkers, and…

统计方法学 · 统计学 2021-03-05 Morgane Pierre-Jean , Florence Mauger , Jean-François Deleuze , Edith Le Floch

In this paper, we introduce a new approach to generate flexible parametric families of distributions. These models arise on competitive and complementary risks scenario, in which the lifetime associated with a particular risk is not…

应用统计 · 统计学 2018-05-22 Pedro L. Ramos , Dipak K. Dey , Francisco Louzada , Victor H. Lachos

Charge and energy transfer in biological and synthetic organic materials are strongly influenced by the coupling of electronic states to high-frequency underdamped vibrations under dephasing noise. Non-perturbative simulations of these…

量子物理 · 物理学 2019-09-05 Alejandro D. Somoza , Oliver Marty , James Lim , Susana F. Huelga , Martin B. Plenio

Multi-omics data present significant challenges for statistical inference due to the complex interdependencies among biological layers. In this paper, we introduce a novel Multi-Omics Factor-Adjusted Cox (MOFA-Cox) model for analyzing…

统计方法学 · 统计学 2025-10-07 Heyuan Zhang , Meiling Hao , Lianqiang Qu , Liuquan Sun

We consider modeling and prediction of Capelin distribution in the Barents sea based on zero-inflated count observation data that vary continuously over a specified survey region. The model is a mixture of two components; a one-point…

统计方法学 · 统计学 2022-10-20 Shonosuke Sugasawa , Tomoyuki Nakagawa , Hiroko Kato Solvang , Sam Subbey , Salah Alrabeei

Understanding the spatial distribution of animals, during all their life phases, as well as how the distributions are influenced by environmental covariates, is a fundamental requirement for the effective management of animal populations.…

应用统计 · 统计学 2020-10-26 Soraia Pereira , Raquel Menezes , Maria Manuel Angélico , Tiago Marques

Pattern-mixture models provide a transparent approach for handling missing data, where the full-data distribution is factorized in a way that explicitly shows the parts that can be estimated from observed data alone, and the parts that…

统计方法学 · 统计学 2019-04-26 Yen-Chi Chen , Mauricio Sadinle

The growing use of high-throughput sequencing (HTS) has enabled the large-scale production of compositional count data, driving progress in microbiome research. However, such count data are often high-dimensional, over-dispersed, and…

其他统计学 · 统计学 2026-05-22 Wenqi Tang , Kamila Fačevicová , Klaus Nordhausen , Sara Taskinen