English
Related papers

Related papers: A subcopula characterization of dependence for the…

200 papers

This paper addresses the problem of quantification and propagation of uncertainties associated with dependence modeling when data for characterizing probability models are limited. Practically, the system inputs are often assumed to be…

Computation · Statistics 2020-04-14 Jiaxin Zhang , Michael D. Shields

Clustering multivariate binary data is of interest in many scientific fields, including ecology, biomedicine, and social policy. Beyond heuristic clustering algorithms, such data can be modelled using multivariate Bernoulli mixture models.…

Methodology · Statistics 2026-04-24 Luisa Ferrari , Maria Franco Villoria , Garritt L. Page , Alex Laini

We discuss the connection between information and copula theories by showing that a copula can be employed to decompose the information content of a multivariate distribution into marginal and dependence components, with the latter…

Statistical Finance · Quantitative Finance 2011-10-26 Rafael S. Calsaverini , Renato Vicente

In this paper, we deduce a new multivariate regression model designed to fit correlated binary data. The multivariate distribution is derived from a Bernoulli mixed model with a nonnormal random intercept on the marginal approach. The…

Methodology · Statistics 2024-06-10 Lizandra C. Fabio , Vanessa Barros , Cristian Villegas , Jalmar M. F. Carrasco

This paper proposes different methods to consistently detect multiple breaks in copula-based dependence measures, mainly focusing on Spearman's $\rho$. The leading model is a factor copula model due to its usefulness for analyzing data in…

Methodology · Statistics 2022-06-13 Marvin Borsch , Alexander Mayer , Dominik Wied

A method for estimating the Shannon differential entropy of multidimensional random variables using independent samples is described. The method is based on decomposing the distribution into a product of the marginal distributions and the…

Statistical Mechanics · Physics 2020-04-22 Gil Ariel , Yoram Louzoun

Clinical trials often evaluate multiple outcome variables to form a comprehensive picture of the effects of a new treatment. The resulting multidimensional insight contributes to clinically relevant and efficient decision-making about…

Methodology · Statistics 2023-08-14 X. M. Kavelaars , J. Mulder , M. C. Kaptein

Clustering methods with dimension reduction have been receiving considerable wide interest in statistics lately and a lot of methods to simultaneously perform clustering and dimension reduction have been proposed. This work presents a novel…

Methodology · Statistics 2014-06-17 Michio Yamamoto , Kenichi Hayashi

Copulas allow a flexible and simultaneous modeling of complicated dependence structures together with various marginal distributions. Especially if the density function can be represented as the product of the marginal density functions and…

Methodology · Statistics 2020-08-31 Jae Youn Ahn , Sebastian Fuchs , Rosy Oh

Building higher-dimensional copulas is generally recognized as a difficult problem. Regular-vines using bivariate copulas provide a flexible class of high-dimensional dependency models. In large dimensions, the drawback of the model is the…

Statistics Theory · Mathematics 2012-06-07 Edith Kovacs , Tamas Szantai

This paper introduces a novel class of models for binary data, which we call log-mean linear models. The characterizing feature of these models is that they are specified by linear constraints on the log-mean linear parameter, defined as a…

Methodology · Statistics 2013-01-14 Alberto Roverato , Monia Lupparelli , Luca La Rocca

We introduce a copula mixture model to perform dependency-seeking clustering when co-occurring samples from different data sources are available. The model takes advantage of the great flexibility offered by the copulas framework to extend…

Methodology · Statistics 2012-07-03 Melanie Rey , Volker Roth

High dimensional and heterogeneous count data are collected in various applied fields. In this paper, we look closely at high-resolution sequencing data on the microbiome, which have enabled researchers to study the genomes of entire…

Methodology · Statistics 2024-01-12 Veronica Vinciotti , Pariya Behrouzi , Reza Mohammadi

In fields such as hydrology and climatology, modelling the entire distribution of positive data is essential, as stakeholders require insights into the full range of values, from low to extreme. Traditional approaches often segment the…

Methodology · Statistics 2025-10-03 Carlo Gaetan , Philippe Naveau

Binary data are highly common in many applications, however it is usually modelled with the assumption that the data are independently and identically distributed. This is typically not the case in many real-world examples and such the…

Methodology · Statistics 2024-06-12 Louise Kimpton , Peter Challenor , Henry Wynn

We investigate the Conway--Maxwell multivariate Bernoulli distributions, a family of multivariate Bernoulli distributions derived from the Conway--Maxwell-binomial distribution. We show that it is possible to set the parametrization such…

Statistics Theory · Mathematics 2026-04-28 Hélène Cossette , Etienne Marceau , Alessandro Mutti , Patrizia Semeraro

Modern datasets commonly feature both substantial missingness and many variables of mixed data types, which present significant challenges for estimation and inference. Complete case analysis, which proceeds using only the observations with…

Methodology · Statistics 2023-04-10 Joseph Feldman , Daniel R. Kowal

Learning the joint dependence of discrete variables is a fundamental problem in machine learning, with many applications including prediction, clustering and dimensionality reduction. More recently, the framework of copula modeling has…

Machine Learning · Statistics 2013-11-15 Alfredo Kalaitzis , Ricardo Silva

We tackle the natural question of whether it is possible to estimate conditional distributions via Sklar's theorem by separately estimating the conditional distributions of the underlying copula and the marginals. Working with so-called…

Statistics Theory · Mathematics 2026-02-03 Kai Schärer , Wolfgang Trutschnig

To conduct Bayesian inference with large data sets, it is often convenient or necessary to distribute the data across multiple machines. We consider a likelihood function expressed as a product of terms, each associated with a subset of the…

Computation · Statistics 2020-04-09 Lewis J. Rendell , Adam M. Johansen , Anthony Lee , Nick Whiteley