English
Related papers

Related papers: A Circular Chatterjee's Correlation Coefficient

200 papers

Randomness or mutual independence is an important underlying assumption for most widely used statistical methods for circular data. Verifying this assumption is essential to ensure the validity and reliability of the resulting inferences.…

Methodology · Statistics 2025-07-01 Shriya Gehlot , Arnab Kumar Laha

In causal inference, a fundamental task is to estimate the effect resulting from a specific treatment, which is often handled with inverse probability weighting. Despite an abundance of attention to the advancement of this task, most…

Methodology · Statistics 2025-07-30 Kuan-Hsun Wu

Distance covariance is a popular measure of dependence between random variables. It has some robustness properties, but not all. We prove that the influence function of the usual distance covariance is bounded, but that its breakdown value…

Methodology · Statistics 2025-08-26 Sarah Leyder , Jakob Raymaekers , Peter J. Rousseeuw

Estimating the dependences between random variables, and ranking them accordingly, is a prevalent problem in machine learning. Pursuing frequentist and information-theoretic approaches, we first show that the p-value and the mutual…

Machine Learning · Computer Science 2012-07-02 Harald Steck

Distance correlation is a novel class of multivariate dependence measure, taking positive values between 0 and 1, and applicable to random vectors of arbitrary dimensions, not necessarily equal. It offers several advantages over the…

Computation · Statistics 2024-05-06 Blanca E. Monroy-Castillo , M. A , Jácome , Ricardo Cao

In the process of building (structural learning) a probabilistic graphical model from a set of observed data, the directional, cyclic dependencies between the random variables of the model are often found. Existing graphical models such as…

Machine Learning · Computer Science 2023-10-26 Oleksii Sirotkin

We introduce the Randomized Dependence Coefficient (RDC), a measure of non-linear dependence between random variables of arbitrary dimension based on the Hirschfeld-Gebelein-R\'enyi Maximum Correlation Coefficient. RDC is defined in terms…

Machine Learning · Statistics 2013-06-04 David Lopez-Paz , Philipp Hennig , Bernhard Schölkopf

Circular permutation connects the N and C termini of a protein and concurrently cleaves elsewhere in the chain, providing an important mechanism for generating novel protein fold and functions. However, their in genomes is unknown because…

Biomolecules · Quantitative Biology 2016-11-17 T. Andrew Binkowski , Bhaskar DasGupta , Jie Liang

In this article, we consider the complete independence test of high-dimensional data. Based on Chatterjee coefficient, we pioneer the development of quadratic test and extreme value test which possess good testing performance for…

Statistics Theory · Mathematics 2024-09-17 Liqi Xia , Ruiyuan Cao , Jiang Du , Jun Dai

Fr\'echet regression extends the principles of linear regression to accommodate responses valued in generic metric spaces. While this approach has primarily focused on exploring relationships between Euclidean predictors and non-Euclidean…

Statistics Theory · Mathematics 2026-02-25 Chang Jun Im , Jeong Min Jeon

We propose the extension of Fr\'{e}chet-Hoeffding copula bounds for circular data. The copula is a powerful tool for describing the dependency of random variables. In two dimensions, the Fr\'{e}chet-Hoeffding upper (lower) bound indicates…

Statistics Theory · Mathematics 2023-11-17 Hiroaki Ogata

We study route choice in a repeated routing game where an uncertain state of nature determines link latency functions, and agents receive private route recommendation. The state is sampled in an i.i.d. manner in every round from a publicly…

Computer Science and Game Theory · Computer Science 2022-08-02 Yixian Zhu , Ketan Savla

Clustered data are common in practice. Clustering arises when subjects are measured repeatedly, or subjects are nested in groups (e.g., households, schools). It is often of interest to evaluate the correlation between two variables with…

Methodology · Statistics 2025-01-16 Shengxin Tu , Chun Li , Bryan E. Shepherd

Circular data are data measured in angles and occur in a variety of scientific disciplines. Bayesian methods promise to allow for flexible analysis of circular data. Three existing MCMC methods (Gibbs, Metropolis-Hastings, and Rejection)…

Computation · Statistics 2015-05-12 Kees Tim Mulder , Irene Klugkist

The quotient correlation is defined here as an alternative to Pearson's correlation that is more intuitive and flexible in cases where the tail behavior of data is important. It measures nonlinear dependence where the regular correlation…

Statistics Theory · Mathematics 2008-12-18 Zhengjun Zhang

Modern regression analysis often involves responses and predictors taking values in the same or distinct metric spaces. To rank non-Euclidean heterogeneous predictors in regression by explanatory strength, analogous to the classical $R^2$,…

Methodology · Statistics 2026-04-28 Shuaida He , Yangzhou Chen , Xin Chen

Scatter plots carry an implicit if subtle message about causality. Whether we look at functions of one variable in pure mathematics, plots of experimental measurements as a function of the experimental conditions, or scatter plots of…

Human-Computer Interaction · Computer Science 2018-09-26 Carl T. Bergstrom , Jevin D. West

We introduce a correlation coefficient that is designed to deal with a variety of ranking formats including those containing non-strict (i.e., with-ties) and incomplete (i.e., unknown) preferences. The correlation coefficient is designed to…

Applications · Statistics 2019-02-19 Yeawon Yoo , Adolfo R. Escobedo , J. Kyle Skolfield

We extend the scope of Azadkia-Chatterjee's dependence coefficient between a scalar response $Y$ and a multivariate covariate $X$ to the case where $X$ takes values in a general metric space. Particular attention is paid to the case where…

Statistics Theory · Mathematics 2025-01-16 Siegfried Hörmann , Daniel Strenger

This paper introduces a new property of estimators of the strength of statistical association, which helps characterize how well an estimator will perform in scenarios where dependencies between continuous and discrete random variables need…

Machine Learning · Statistics 2021-01-12 Kiran Karra , Lamine Mili