English
Related papers

Related papers: Splitting models for multivariate count data

200 papers

A generalization of a distribution increases the flexibility particularly in studying of a phenomenon and its properties. Many generalizations of continuous univariate distributions are available in literature. In this study, an…

Applications · Statistics 2024-08-30 Brijesh P. Singh , Sandeep Singh , Utpal Dhar Das

Regression for count data is widely performed by models such as Poisson, negative binomial (NB) and zero-inflated regression. A challenge often faced by practitioners is the selection of the right model to take into account dispersion,…

Methodology · Statistics 2018-08-02 Hadeel S. Klakattawi , Veronica Vinciotti , Keming Yu

The paper considers multivariate discrete random sums with equal number of summands. Such distributions describe the total claim amount received by a company in a fixed time point. In Queuing theory they characterize cumulative waiting…

Probability · Mathematics 2016-12-12 Pavlina Jordanova

Neural networks and other machine learning models compute continuous representations, while humans communicate mostly through discrete symbols. Reconciling these two forms of communication is desirable for generating human-readable…

Machine Learning · Computer Science 2022-02-14 António Farinhas , Wilker Aziz , Vlad Niculae , André F. T. Martins

Interpreting a nonparametric regression model with many predictors is known to be a challenging problem. There has been renewed interest in this topic due to the extensive use of machine learning algorithms and the difficulty in…

Machine Learning · Statistics 2018-09-11 Xiaoyu Liu , Jie Chen , Joel Vaughan , Vijayan Nair , Agus Sudjianto

The recently introduced framework of universal inference provides a new approach to constructing hypothesis tests and confidence regions that are valid in finite samples and do not rely on any specific regularity assumptions on the…

Statistics Theory · Mathematics 2023-09-11 David Strieder , Mathias Drton

We propose a new framework for generating cross-sectional synthetic datasets via disjoint generative models. In this paradigm, a dataset is partitioned into disjoint subsets that are supplied to separate instances of generative models. The…

Machine Learning · Computer Science 2025-07-29 Anton Danholt Lautrup , Muhammad Rajabinasab , Tobias Hyrup , Arthur Zimek , Peter Schneider-Kamp

Many datasets are observed on a finite set of equally spaced directions instead of the exact angles, such as the wind direction data. However, in the statistical literature, bivariate models are only available for continuous circular random…

Methodology · Statistics 2026-02-16 Brajesh Kumar Dhakad , Jayant Jha , Debepsita Mukherjee

The spreading dynamics in social networks are often studied under the assumption that individuals' statuses, whether informed or infected, are fully observable. However, in many real-world situations, such statuses remain unobservable,…

Social and Information Networks · Computer Science 2026-02-23 Derrick Gilchrist Edward Manoharan , Anubha Goel , Alexandros Iosifidis , Henri Hansen , Juho Kanniainen

The discrete kernel method was developed to estimate count data distributions, distinguishing discrete associated kernels based on their asymptotic behaviour. This study investigates the class of discrete asymmetric kernels and their…

Methodology · Statistics 2017-02-07 Tristan Senga Kiessé

A discrete-time stochastic process derived from a model of basketball is used to generalize any discrete distribution. The generalized distributions can have one or two more parameters than the parent distribution. Those derived from…

Applications · Statistics 2020-06-25 Rose Baker

Circular variables arise in a multitude of data-modelling contexts ranging from robotics to the social sciences, but they have been largely overlooked by the machine learning community. This paper partially redresses this imbalance by…

Machine Learning · Statistics 2017-08-10 Alexandre K. W. Navarro , Jes Frellsen , Richard E. Turner

For exchangeable data, mixture models are an extremely useful tool for density estimation due to their attractive balance between smoothness and flexibility. When additional covariate information is present, mixture models can be extended…

Methodology · Statistics 2023-08-01 Sara Wade , Vanda Inacio , Sonia Petrone

Simulation has become the evaluation method of choice for many areas of distributing computing research. However, most existing simulation packages have several limitations on the size and complexity of the system being modeled. Fine…

Distributed, Parallel, and Cluster Computing · Computer Science 2011-07-01 Dobre Ciprian , Cristea Valentin , Iosif C. Legrand

We develop Bayesian models for density regression with emphasis on discrete outcomes. The problem of density regression is approached by considering methods for multivariate density estimation of mixed scale variables, and obtaining…

Methodology · Statistics 2019-08-14 Georgios Papageorgiou

Model averaging is a useful and robust method for dealing with model uncertainty in statistical analysis. Often, it is useful to consider data subset selection at the same time, in which model selection criteria are used to compare models…

Methodology · Statistics 2023-10-26 Ethan T. Neil , Jacob W. Sitison

Varying coefficient models are useful in applications where the effect of the covariate might depend on some other covariate such as time or location. Various applications of these models often give rise to case-specific prior distributions…

Methodology · Statistics 2019-12-05 Maria Franco-Villoria , Massimo Ventrucci , Håvard Rue

We extend the synthetic theories of discrete and Gaussian categorical probability by introducing a diagrammatic calculus for reasoning about hybrid probabilistic models in which continuous random variables, conditioned on discrete ones,…

Logic in Computer Science · Computer Science 2025-10-07 Mateo Torres-Ruiz , Robin Piedeleu , Alexandra Silva , Fabio Zanasi

A common assumption in causal modeling posits that the data is generated by a set of independent mechanisms, and algorithms should aim to recover this structure. Standard unsupervised learning, however, is often concerned with training a…

Machine Learning · Computer Science 2019-03-05 Francesco Locatello , Damien Vincent , Ilya Tolstikhin , Gunnar Rätsch , Sylvain Gelly , Bernhard Schölkopf

We propose a new modeling approach that is a generalization of generative and discriminative models. The core idea is to use an implicit parameterization of a joint probability distribution by specifying only the conditional distributions.…

Machine Learning · Computer Science 2016-12-06 Dmitrij Schlesinger , Carsten Rother