English
Related papers

Related papers: Model selection in sparse high-dimensional vine co…

200 papers

The varying-coefficient model is an important nonparametric statistical model that allows us to examine how the effects of covariates vary with exposure variables. When the number of covariates is big, the issue of variable selection…

Statistics Theory · Mathematics 2013-03-05 Jianqing Fan , Yunbei Ma , Wei Dai

Vine copulas (or pair-copula constructions) have become an important tool for high-dimensional dependence modeling. Typically, so called simplified vine copula models are estimated where bivariate conditional copulas are approximated by…

Methodology · Statistics 2017-05-19 Christian Schellhase , Fabian Spanhel

Copula models have become one of the most widely used tools in the applied modelling of multivariate data. Similarly, Bayesian methods are increasingly used to obtain efficient likelihood-based inference. However, to date, there has been…

Methodology · Statistics 2015-10-13 Michael Stanley Smith

Model selection is an indispensable part of data analysis dealing very frequently with fitting and prediction purposes. In this paper, we tackle the problem of model selection in a general linear regression where the parameter matrix…

Signal Processing · Electrical Eng. & Systems 2022-09-19 Prakash B. Gohain , Magnus Jansson

Copulas allow to learn marginal distributions separately from the multivariate dependence structure (copula) that links them together into a density function. Vine factorizations ease the learning of high-dimensional copulas by constructing…

Methodology · Statistics 2013-02-19 David Lopez-Paz , José Miguel Hernández-Lobato , Zoubin Ghahramani

Mixture model-based clustering has become an increasingly popular data analysis technique since its introduction over fifty years ago, and is now commonly utilized within a family setting. Families of mixture models arise when the component…

Methodology · Statistics 2019-11-11 Sanjeena Subedi , Paul D. McNicholas

Model selection is indispensable to high-dimensional sparse modeling in selecting the best set of covariates among a sequence of candidate models. Most existing work assumes implicitly that the model is correctly specified or of fixed…

Statistics Theory · Mathematics 2014-12-24 Pallavi Basu , Yang Feng , Jinchi Lv

While the Bayesian Information Criterion (BIC) and Akaike Information Criterion (AIC) are powerful tools for model selection in linear regression, they are built on different prior assumptions and thereby apply to different data generation…

Methodology · Statistics 2017-12-15 MB de Kock , HC Eggers

In many conventional scientific investigations with high or ultra-high dimensional feature spaces, the relevant features, though sparse, are large in number compared with classical statistical problems, and the magnitude of their effects…

Statistics Theory · Mathematics 2011-07-14 Shan Luo , Zehua Chen

Vine copulas, constructed using bivariate copulas as building blocks, provide a flexible framework for modeling multi-dimensional dependencies. However, this flexibility is accompanied by rapidly increasing complexity as dimensionality…

Methodology · Statistics 2025-04-25 Ichiro Nishi , Yoshinori Kawasaki

We consider a sparse linear regression model, when the number of available predictors, $p$, is much larger than the sample size, $n$, and the number of non-zero coefficients, $p_0$, is small. To choose the regression model in this…

Statistics Theory · Mathematics 2018-05-31 Piotr Szulc

We consider the problem of modeling the dependence among many time series. We build high dimensional time-varying copula models by combining pair-copula constructions (PCC) with stochastic autoregressive copula (SCAR) models to capture…

Methodology · Statistics 2012-02-10 Carlos Almeida , Claudia Czado , Hans Manner

Vine copulas are a useful statistical tool to describe the dependence structure between several random variables, especially when the number of variables is very large. When modeling data with vine copulas, one often is confronted with a…

Methodology · Statistics 2017-05-10 Matthias Killiches , Daniel Kraus , Claudia Czado

The original development of Shapley values for prediction explanation relied on the assumption that the features being described were independent. If the features in reality are dependent this may lead to incorrect explanations. Hence,…

Methodology · Statistics 2021-02-15 Kjersti Aas , Thomas Nagler , Martin Jullum , Anders Løland

We demonstrate how the uncertainty of parameter point estimates can be assessed in a maximum likelihood framework in order to prevent overfitting and erroneous detection of time-inhomogeneity. The class of models we consider are regular…

Computation · Statistics 2012-05-23 Jakob Stöber , Ulf Schepsmeier

We consider approximate Bayesian model choice for model selection problems that involve models whose Fisher-information matrices may fail to be invertible along other competing submodels. Such singular models do not obey the regularity…

Methodology · Statistics 2016-03-24 Mathias Drton , Martyn Plummer

In many studies multivariate event time data are generated from clusters having a possibly complex association pattern. Flexible models are needed to capture this dependence. Vine copulas serve this purpose. Inference methods for vine…

Applications · Statistics 2017-07-25 Nicole Barthel , Candida Geerdens , Matthias Killiches , Paul Janssen , Claudia Czado

Penalized regression models are popularly used in high-dimensional data analysis to conduct variable selection and model fitting simultaneously. Whereas success has been widely reported in literature, their performances largely depend on…

Machine Learning · Statistics 2013-12-16 Wei Sun , Junhui Wang , Yixin Fang

Modeling high-dimensional dependencies while keeping likelihoods tractable remains challenging. Classical vine-copula pipelines are interpretable but can be expensive, while many neural estimators are flexible but less structured. In this…

Machine Learning · Computer Science 2026-05-08 Houman Safaai

Standard penalized methods of variable selection and parameter estimation rely on the magnitude of coefficient estimates to decide which variables to include in the final model. However, coefficient estimates are unreliable when the design…

Methodology · Statistics 2018-02-13 Jonathan P Williams , Jan Hannig