English
Related papers

Related papers: Exploring and measuring non-linear correlations: C…

200 papers

This paper presents a new methodology for clustering multivariate time series leveraging optimal transport between copulas. Copulas are used to encode both (i) intra-dependence of a multivariate time series, and (ii) inter-dependence…

Machine Learning · Computer Science 2016-01-12 Gautier Marti , Frank Nielsen , Philippe Donnat

A frequent task in exploratory data analysis consists in examining pairwise dependencies between data variables. Popular approaches include visualizing correlation or scatter plot matrices. However, both methods can be misleading. The…

Applications · Statistics 2022-04-04 Arturo Erdely , Manuel Rubio-Sanchez

Simultaneous recordings from many neurons hide important information and the connections characterizing the network remain generally undiscovered despite the progresses of statistical and machine learning techniques. Discerning the presence…

Applications · Statistics 2019-03-21 Pietro Verzelli , Laura Sacerdote

We introduce a copula mixture model to perform dependency-seeking clustering when co-occurring samples from different data sources are available. The model takes advantage of the great flexibility offered by the copulas framework to extend…

Methodology · Statistics 2012-07-03 Melanie Rey , Volker Roth

We present a methodology for clustering N objects which are described by multivariate time series, i.e. several sequences of real-valued random variables. This clustering methodology leverages copulas which are distributions encoding the…

Machine Learning · Statistics 2016-11-15 Gautier Marti , Sébastien Andler , Frank Nielsen , Philippe Donnat

This study outlines a comprehensive methodology utilizing copulas to discern inconsistencies in the behavior exhibited by pairs of financial assets. It introduces a robust approach to establishing the interrelationship between the returns…

Computational Finance · Quantitative Finance 2023-12-05 Alexander Shulzhenko

This paper proposes a new paradigm and computational framework for identification of correspondences between sub-structures of distinct composite systems. For this, we define and investigate a variant of traditional data clustering, termed…

Machine Learning · Computer Science 2007-05-23 Zvika Marx , Ido Dagan , Joachim Buhmann

In this work, the possibility of clustering correlated random variables was examined, both because of their mutual similarity and because of their similarity to the principal components. The k-means algorithm and spectral algorithms were…

Machine Learning · Computer Science 2019-09-10 Zenon Gniazdowski , Dawid Kaliszewski

This paper proposes multivariate copula models for hierarchical data. They account for two types of correlation: one is between variables measured on the same unit and the other is a correlation between units in the same cluster. This model…

Methodology · Statistics 2023-04-24 Talagbe Gabin Akpo , Louis-Paul Rivest

Handling highly dependent data is crucial in clinical trials, particularly in fields related to ophthalmology. Incorrectly specifying the dependency structure can lead to biased inferences. Traditionally, models rely on three fixed…

Methodology · Statistics 2025-09-30 Shuyi Liang , Takeshi Emura , Chang-Xing Ma , Yijing Xin , Xin-Wei Huang

Testing for pairwise independence for the case where the number of variables may be of the same size or even larger than the sample size has received increasing attention in the recent years. We contribute to this branch of the literature…

Statistics Theory · Mathematics 2024-09-18 Axel Bücher , Cambyse Pakzad

We introduce the coverage correlation coefficient, a novel nonparametric measure of statistical association designed to quantifies the extent to which two random variables have a joint distribution concentrated on a singular subset with…

Methodology · Statistics 2025-08-18 Xuzhi Yang , Mona Azadkia , Tengyao Wang

When scholars study joint distributions of multiple variables, copulas are useful. However, if the variables are not linearly correlated with each other yet are still not independent, most of conventional copulas are not up to the task.…

Methodology · Statistics 2023-08-08 Kentaro Fukumoto

Distance correlation coefficient (DCC) can be used to identify new associations and correlations between multiple variables. The distance correlation coefficient applies to variables of any dimension, can be used to determine smaller sets…

Statistical Finance · Quantitative Finance 2023-01-13 J. E. Salgado-Hernández , Manan Vyas

In a standard cluster analysis, such as k-means, in addition to clusters locations and distances between them, it's important to know if they are connected or well separated from each other. The main focus of this paper is discovering the…

Machine Learning · Statistics 2017-05-22 Evgeny Bauman , Konstantin Bauman

Regression analysis is one of the most popularly used statistical technique which only measures the direct effect of independent variables on dependent variable. Path analysis looks for both direct and indirect effects of independent…

Methodology · Statistics 2024-06-26 Alam Ali , Ashok Kumar Pathak , Mohd Arshad , Ayyub Sheikhi

Clustering under pairwise constraints is an important knowledge discovery tool that enables the learning of appropriate kernels or distance metrics to improve clustering performance. These pairwise constraints, which come in the form of…

Machine Learning · Computer Science 2022-03-24 Benedikt Boecking , Vincent Jeanselme , Artur Dubrawski

We consider estimation in a high-dimensional linear model with strongly correlated variables. We propose to cluster the variables first and do subsequent sparse estimation such as the Lasso for cluster-representatives or the group Lasso…

Methodology · Statistics 2015-01-14 Peter Bühlmann , Philipp Rütimann , Sara van de Geer , Cun-Hui Zhang

Probability density estimation from observed data constitutes a central task in statistics. In this brief, we focus on the problem of estimating the copula density associated to any observed data, as it fully describes the dependence…

Machine Learning · Computer Science 2025-07-09 Nunzio A. Letizia , Nicola Novello , Andrea M. Tonello

The majority of model-based clustering techniques is based on multivariate Normal models and their variants. In this paper copulas are used for the construction of flexible families of models for clustering applications. The use of copulas…

Methodology · Statistics 2018-02-16 Ioannis Kosmidis , Dimitris Karlis
‹ Prev 1 2 3 10 Next ›