English
Related papers

Related papers: The Negative Binomial Chain-Ladder: A Full Likelih…

200 papers

Label noise in data has long been an important problem in supervised learning applications as it affects the effectiveness of many widely used classification methods. Recently, important real-world applications, such as medical diagnosis…

Machine Learning · Statistics 2021-12-02 Shunan Yao , Bradley Rava , Xin Tong , Gareth James

Negative probabilities arise primarily in physics, statistical quantum mechanics and quantum computing. Negative probabilities arise as mixing distributions of unobserved latent variables in Bayesian modeling. Our goal is to provide a link…

Quantum Physics · Physics 2024-09-06 Nick Polson , Vadim Sokolov

Continual learning (CL) is an important technique to allow artificial neural networks to work in open environments. CL enables a system to learn new tasks without severe interference to its performance on old tasks, i.e., overcome the…

Machine Learning · Computer Science 2024-07-08 Liangxuan Guo , Yang Chen , Shan Yu

Background: Analyses of elastic scattering with the optical model (OMP) are widely used in nuclear reactions. Purpose: Previous work compared a traditional frequentist approach and a Bayesian approach to quantify uncertainties in the OMP.…

Nuclear Theory · Physics 2024-03-04 C. D. Pruitt , A. E. Lovell , C. Hebborn , F. M. Nunes

The composite likelihood (CL) is amongst the computational methods used for estimation of the generalized linear mixed model (GLMM) in the context of bivariate meta-analysis of diagnostic test accuracy studies. Its advantage is that the…

Methodology · Statistics 2018-07-12 Aristidis K. Nikoloulopoulos

We consider a variant of the set covering problem with uncertain parameters, which we refer to as the chance-constrained set multicover problem (CC-SMCP). In this problem, we assume that there is uncertainty regarding whether a selected set…

Optimization and Control · Mathematics 2026-05-04 Shunyu Yao , Neng Fan , Pavlo Krokhmal

Beam search is the default decoding strategy for many sequence generation tasks in NLP. The set of approximate K-best items returned by the algorithm is a useful summary of the distribution for many applications; however, the candidates…

Computation and Language · Computer Science 2023-03-03 Clara Meister , Afra Amini , Tim Vieira , Ryan Cotterell

We develop a new class of dynamic multivariate Poisson count models that allow for fast online updating and we refer to these models as multivariate Poisson-scaled beta (MPSB). The MPSB model allows for serial dependence in the counts as…

Methodology · Statistics 2016-09-16 Tevfik Aktekin , Nicholas G. Polson , Refik Soyer

Leveraging the diversity and quantity of data provided by various graph-structured data augmentations while preserving intrinsic semantic information is challenging. Additionally, successive layers in graph neural network (GNN) tend to…

Machine Learning · Computer Science 2026-03-19 Jie Chen , Hua Mao , Chuanbin Liu , Zhu Wang , Xi Peng

A three-parameter discrete distribution is developed to describe the multiplicity distributions observed in total- and limited phase space volumes in different collision processes. The probability law is obtained by the Poisson transform of…

High Energy Physics - Phenomenology · Physics 2009-10-28 S. Hegyi

While several Gaussian mixture models-based biclustering approaches currently exist in the literature for continuous data, approaches to handle discrete data have not been well researched. A multivariate Poisson-lognormal (MPLN) model-based…

Methodology · Statistics 2025-03-13 Caitlin Kral , Evan Chance , Ryan Browne , Sanjeena Subedi

The paper presents a new statistical method that enables the use of systematic errors in the maximum-likelihood regression of integer-count Poisson data to a parametric model. The method is primarily aimed at the characterization of the…

Instrumentation and Methods for Astrophysics · Physics 2024-07-18 Max Bonamente , Yang Chen , Dale Zimmerman

Modern datasets are characterized by a large number of features that may conceal complex dependency structures. To deal with this type of data, dimensionality reduction techniques are essential. Numerous dimensionality reduction methods…

Methodology · Statistics 2021-06-02 Francesco Denti , Diego Doimo , Alessandro Laio , Antonietta Mira

Imputation of missing values is a strategy for handling non-responses in surveys or data loss in measurement processes, which may be more effective than ignoring them. When the variable represents a count, the literature dealing with this…

Applications · Statistics 2020-07-31 Gilma Hernández-Herrera , Albert Navarro , David Moriña

In subgroup analysis, testing the existence of a subgroup with a differential treatment effect serves as protection against spurious subgroup discovery. Despite its importance, this hypothesis testing possesses a complicated nature:…

Statistics Theory · Mathematics 2025-03-21 Shota Takeishi

In this work we show that the latest LHC data on multiplicity moments $C_2-C_5$ are well described by a two-step model in the form of a convolution of the Poisson distribution with energy-dependent source function. For the source function…

High Energy Physics - Phenomenology · Physics 2011-10-18 Michal Praszalowicz

Growth in both size and complexity of modern data challenges the applicability of traditional likelihood-based inference. Composite likelihood (CL) methods address the difficulties related to model selection and computational intractability…

Statistics Theory · Mathematics 2017-09-12 Zhendong Huang , Davide Ferrari

Poisson non-negative matrix factorization (NMF) is a widely used method to find interpretable "parts-based" decompositions of count data. While many variants of Poisson NMF exist, existing methods assume that the "parts" in the…

Machine Learning · Computer Science 2026-01-12 Eric Weine , Peter Carbonetto , Rafael A. Irizarry , Matthew Stephens

Multiple systems estimation using a Poisson loglinear model is a standard approach to quantifying hidden populations where data sources are based on lists of known cases. Information criteria are often used for selecting between the large…

Methodology · Statistics 2023-11-23 Bernard W. Silverman , Lax Chan , Kyle Vincent

Generalized linear models, such as logistic regression, are widely used to model the association between a treatment and a binary outcome as a function of baseline covariates. However, the coefficients of a logistic regression model…

Methodology · Statistics 2022-01-04 Jiaqi Yin , Sonia Markes , Thomas S. Richardson , Linbo Wang