English
Related papers

Related papers: Learning Vine Copula Models For Synthetic Data Gen…

200 papers

This paper proposes a new method to generate synthetic data sets based on copula models. Our goal is to produce surrogate data resembling real data in terms of marginal and joint distributions. We present a complete and reliable algorithm…

Machine Learning · Computer Science 2022-04-01 Regis Houssou , Mihai-Cezar Augustin , Efstratios Rappos , Vivien Bonvin , Stephan Robert-Nicoud

Future electricity consumption is fundamentally uncertain and dependent on many variables such as economic activity, weather, electricity rates and demand side management. The stochasticity of system load as well as power generation from…

Applications · Statistics 2017-08-21 Swasti R. Khuntia , Jose L. Rueda , Mart A. M. M. van der Meijden

Dependence strucuture estimation is one of the important problems in machine learning domain and has many applications in different scientific areas. In this paper, a theoretical framework for such estimation based on copula and copula…

Machine Learning · Computer Science 2019-09-11 Jian Ma , Zengqi Sun

We introduce a new goodness-of-fit test for regular vine (R-vine) copula models. R-vine copulas are a very flexible class of multivariate copulas based on a pair-copula construction (PCC). The test arises from the information matrix…

Computation · Statistics 2013-06-05 Ulf Schepsmeier

Time-varying dependence is often modeled with dynamic correlations or Gaussian graphical models, but multivariate systems can change through tail behavior, asymmetry, or conditional structure even when correlations are nearly stable. We…

Machine Learning · Statistics 2026-05-08 Houman Safaai , Alessandro Marin Vargas

Estimating dependence relationships between variables is a crucial issue in many applied domains, such as medicine, social sciences and psychology. When several variables are entertained, these can be organized into a network which encodes…

Methodology · Statistics 2025-01-01 Federico Castelletti

We propose to use nonparametric Bernstein copulas as bivariate pair-copulas in high-dimensional vine models. The resulting smooth and nonparametric vine copulas completely obviate the error-prone need for choosing the pair-copulas from…

Risk Management · Quantitative Finance 2012-10-09 Gregor Weiß , Marcus Scheffer

We propose stepwise variational inference (VI) with vine copulas: a universal VI procedure that combines vine copulas with a novel stepwise estimation procedure of the variational parameters. Vine copulas consist of a nested sequence of…

Machine Learning · Statistics 2026-03-25 Elisabeth Griesbauer , Leiv Rønneberg , Arnoldo Frigessi , Claudia Czado , Ingrid Hobæk Haff

Vine copulas constitute a flexible way for modeling of dependences using only pair copulas as building blocks. The pair-copula constructions introduced by Joe (1997) are able to encode more types of dependences in the same time since they…

Methodology · Statistics 2016-04-12 Edith Kovács , Tamás Szántai

Deep generative models offer powerful tools for multivariate data analysis, but their black-box architectures are often unidentified and difficult to interpret. We introduce the Deep Discrete Encoder (DDE) Copula, an identifiable and…

Machine Learning · Statistics 2026-05-28 Joseph Feldman , Yuqi Gu

Population synthesis involves generating synthetic yet realistic representations of a target population of micro-agents for behavioral modeling and simulation. Traditional methods, often reliant on target population samples, such as census…

Can we improve machine-learning (ML) emulators with synthetic data? If data are scarce or expensive to source and a physical model is available, statistically generated data may be useful for augmenting training sets cheaply. Here we…

Machine Learning · Computer Science 2021-09-28 David Meyer , Thomas Nagler , Robin J. Hogan

We introduce a new goodness-of-fit test for regular vine (R-vine) copula models, a flexible class of multivariate copulas based on a pair-copula construction (PCC). The test arises from the information matrix ratio. The corresponding test…

Computation · Statistics 2013-09-24 Ulf Schepsmeier

By sampling from the latent space of an autoencoder and decoding the latent space samples to the original data space, any autoencoder can simply be turned into a generative model. For this to work, it is necessary to model the autoencoder's…

Machine Learning · Statistics 2023-09-19 Maximilian Coblenz , Oliver Grothe , Fabian Kächele

We propose TVineSynth, a vine copula based synthetic tabular data generator, which is designed to balance privacy and utility, using the vine tree structure and its truncation to do the trade-off. Contrary to synthetic data generators that…

Machine Learning · Computer Science 2025-03-21 Elisabeth Griesbauer , Claudia Czado , Arnoldo Frigessi , Ingrid Hobæk Haff

A synthetic dataset is a data object that is generated programmatically, and it may be valuable to creating a single dataset from multiple sources when direct collection is difficult or costly. Although it is a fundamental step for many…

Applications · Statistics 2020-09-22 Zheng Li , Yue Zhao , Jialin Fu

Verification and validation of fully automated vehicles is linked to an almost intractable challenge of reflecting the real world with all its interactions in a virtual environment. Influential stochastic parameters need to be extracted…

Applications · Statistics 2022-11-22 Katrin Lotto , Thomas Nagler , Mladjan Radic

Copulas are a fundamental tool for modelling multivariate dependencies in data, forming the method of choice in diverse fields and applications. However, the adoption of existing models for multimodal and high-dimensional dependencies is…

Machine Learning · Statistics 2026-05-20 David Huk , Theodoros Damoulas

With the advancements of computer architectures, the use of computational models proliferates to solve complex problems in many scientific applications such as nuclear physics and climate research. However, the potential of such models is…

Computation · Statistics 2021-07-05 Vojtech Kejzlar , Tapabrata Maiti

Discrete diffusion models have recently shown significant progress in modeling complex data, such as natural languages and DNA sequences. However, unlike diffusion models for continuous data, which can generate high-quality samples in just…

Machine Learning · Computer Science 2025-03-20 Anji Liu , Oliver Broadrick , Mathias Niepert , Guy Van den Broeck