English
Related papers

Related papers: Data driven partition-of-unity copulas with applic…

200 papers

Probability density estimation from observed data constitutes a central task in statistics. In this brief, we focus on the problem of estimating the copula density associated to any observed data, as it fully describes the dependence…

Machine Learning · Computer Science 2025-07-09 Nunzio A. Letizia , Nicola Novello , Andrea M. Tonello

Probability distributions produced by the cross-entropy loss for ordinal classification problems can possess undesired properties. We propose a straightforward technique to constrain discrete ordinal probability distributions to be unimodal…

Machine Learning · Statistics 2017-06-23 Christopher Beckham , Christopher Pal

We exploit Gaussian copulas to specify a class of multivariate circular distributions and obtain parametric models for the analysis of correlated circular data. This approach provides a straightforward extension of traditional multivariate…

Methodology · Statistics 2024-06-07 Francesco Lagona , Marco Mingione

A flexible semiparametric class of models is introduced that offers an alternative to classical regression models for count data as the Poisson and negative binomial model, as well as to more general models accounting for excess zeros that…

Methodology · Statistics 2020-03-30 Moritz Berger , Gerhard Tutz

The last decade witnessed an explosion in the availability of data for operations research applications. Motivated by this growing availability, we propose a novel schema for utilizing data to design uncertainty sets for robust optimization…

Optimization and Control · Mathematics 2014-11-25 Dimitris Bertsimas , Vishal Gupta , Nathan Kallus

Missing values with mixed data types is a common problem in a large number of machine learning applications such as processing of surveys and in different medical applications. Recently, Gaussian copula models have been suggested as a means…

Machine Learning · Statistics 2021-07-02 Benjamin Christoffersen , Mark Clements , Keith Humphreys , Hedvig Kjellström

We give derivations of some basic results for the Bernstein approximation in $n$ variables that are useful in investigating copulas. It is shown that Bernstein approximations of copulas are again copulas. We exhibit a stochastic…

Statistics Theory · Mathematics 2009-03-06 MD Taylor

We propose a semiparametric family of copulas based on a set of orthonormal functions and a matrix. This new copula permits to reach values of Spearman's Rho arbitrarily close to one without introducing a singular component. Moreover, it…

Statistics Theory · Mathematics 2013-10-22 Cécile Amblard , Stephane Girard , Ludovic Menneteau

The basic goal of computer engineering is the analysis of data. Such data are often large data sets distributed according to various distribution models. In this manuscript we focus on the analysis of non-Gaussian distributed data. In the…

Methodology · Statistics 2019-02-11 Krzysztof Domino

Copula-based models provide a great deal of flexibility in modelling multivariate distributions, allowing for the specifications of models for the marginal distributions separately from the dependence structure (copula) that links them to…

Methodology · Statistics 2021-09-09 Nicolás Kuschinski , Alejandro Jara

Leiner et al. [2023] introduce an important generalization of sample splitting, which they call data fission. They consider two cases of data fission: P1 fission and P2 fission. While P1 fission is extremely useful and easy to use, Leiner…

Methodology · Statistics 2024-09-06 Anna Neufeld , Ameer Dharamshi , Lucy L. Gao , Daniela Witten , Jacob Bien

Modern datasets commonly feature both substantial missingness and many variables of mixed data types, which present significant challenges for estimation and inference. Complete case analysis, which proceeds using only the observations with…

Methodology · Statistics 2023-04-10 Joseph Feldman , Daniel R. Kowal

In this work, tests of symmetry for bivariate copulas are introduced and studied using empirical Bernstein copula process. Three statistics are proposed and their asymptotic properties are established. Besides, a multiplier bootstrap…

Methodology · Statistics 2024-05-14 Guanjie Lyu , Mohamed Belalia

Modern quantitative risk management relies on an adequate modeling of the tail dependence and a possibly accurate quantification of risk measures, like Value at Risk (VaR), at high confidence levels like 1 in 100 or even 1 in 2000. Quantum…

Methodology · Statistics 2020-03-10 Janusz Milek

We present a joint copula-based model for insurance claims and sizes. It uses bivariate copulae to accommodate for the dependence between these quantities. We derive the general distribution of the policy loss without the restrictive…

Statistics Theory · Mathematics 2012-09-25 Nicole Kraemer , Eike C. Brechmann , Daniel Silvestrini , Claudia Czado

Learning the joint dependence of discrete variables is a fundamental problem in machine learning, with many applications including prediction, clustering and dimensionality reduction. More recently, the framework of copula modeling has…

Machine Learning · Statistics 2013-11-15 Alfredo Kalaitzis , Ricardo Silva

Regression for count data is widely performed by models such as Poisson, negative binomial (NB) and zero-inflated regression. A challenge often faced by practitioners is the selection of the right model to take into account dispersion,…

Methodology · Statistics 2018-08-02 Hadeel S. Klakattawi , Veronica Vinciotti , Keming Yu

Copulas are mathematical tools for modeling joint probability distributions. Since copulas enable one to conveniently treat the marginal distribution of each variable and the interdependencies among variables separately, in the past 60…

Quantum Physics · Physics 2022-06-28 Daiwei Zhu , Weiwei Shen , Annarita Giani , Saikat Ray Majumder , Bogdan Neculaes , Sonika Johri

Many types of bounded data defined on the unit interval arise naturally as ratios of the form $X/(X + Y)$. In the existing literature, the main statistical models proposed for this type of bounded data typically based on the assumption that…

Methodology · Statistics 2026-03-04 Roberto Vila , Felipe Quintino , Marcelo Bourguignon

Grouped data are commonly encountered in applications. The Bernstein polynomial model is proposed as an approximate model in this paper for estimating a univariate density function based on grouped data. The coefficients of the Bernstein…

Methodology · Statistics 2015-07-21 Zhong Guan