English
Related papers

Related papers: The Randomized Dependence Coefficient

200 papers

(To appear in The American Statistician.) Distance covariance (Sz\'ekely, Rizzo, and Bakirov, 2007) is a fascinating recent notion, which is popular as a test for dependence of any type between random variables $X$ and $Y$. This approach…

Methodology · Statistics 2024-07-08 Jakob Raymaekers , Peter J. Rousseeuw

Langevin Monte Carlo (LMC) is a popular Bayesian sampling method. For the log-concave distribution function, the method converges exponentially fast, up to a controllable discretization error. However, the method requires the evaluation of…

Machine Learning · Statistics 2025-03-07 Zhiyan Ding , Qin Li

Following our previous work on copula-based nonsymmetric dependence measures, we introduce similar measures for discrete random variables. The measures cover the range between two extremes: independence and complete dependence, which take…

Methodology · Statistics 2015-12-29 Hui Li

We propose a new measure related with tail dependence in terms of correlation: quantile correlation coefficient of random variables X, Y. The quantile correlation is defined by the geometric mean of two quantile regression slopes of X on Y…

Methodology · Statistics 2018-03-19 Ji-Eun Choi , Dong Wan Shin

Distance covariance and distance correlation have long been regarded as natural measures of dependence between two random vectors, and have been used in a variety of situations for testing independence. Despite their popularity, the…

Methodology · Statistics 2025-03-31 Hallin Marc , Davide La Vecchia , Hang Liu , Xinyi Xu

Ridge regression with random coefficients provides an important alternative to fixed coefficients regression in high dimensional setting when the effects are expected to be small but not zeros. This paper considers estimation and prediction…

Machine Learning · Statistics 2023-06-29 Hongzhe Zhang , Hongzhe Li

Classical dependence measures such as Pearson correlation, Spearman's $\rho$, and Kendall's $\tau$ can detect only monotonic or linear dependence. To overcome these limitations, Szekely et al.(2007) proposed distance covariance as a…

Computation · Statistics 2019-02-07 Arin Chaudhuri , Wenhao Hu

The Regression Discontinuity (RD) design is one of the most widely used non-experimental methods for causal inference and program evaluation. Over the last two decades, statistical and econometric methods for RD analysis have expanded and…

Econometrics · Economics 2022-02-25 Matias D. Cattaneo , Rocio Titiunik

We introduce kernel integrated $R^2$, a new measure of statistical dependence that combines the local normalization principle of the recently introduced integrated $R^2$ with the flexibility of reproducing kernel Hilbert spaces (RKHSs). The…

Machine Learning · Statistics 2026-02-27 Pouya Roudaki , Shakeel Gavioli-Akilagun , Florian Kalinke , Mona Azadkia , Zoltán Szabó

We propose a new \textit{randomized Bregman (block) coordinate descent} (RBCD) method for minimizing a composite problem, where the objective function could be either convex or nonconvex, and the smooth part are freed from the global…

Optimization and Control · Mathematics 2020-01-16 Tianxiang Gao , Songtao Lu , Jia Liu , Chris Chu

We consider the distribution of the sum and the maximum of a collection of independent exponentially distributed random variables. The focus is laid on the explicit form of the density functions (pdf) of non-i.i.d. sequences. Those are…

Probability · Mathematics 2013-07-16 Markus Bibinger

Many tools exist to detect dependence between random variables, a core question across a wide range of machine learning, statistical, and scientific endeavors. Although several statistical tests guarantee eventual detection of any…

Machine Learning · Statistics 2026-03-23 Nathaniel Xu , Feng Liu , Danica J. Sutherland

In this short technical report, we define on the sample space R^D a distance between data points which depends on their correlation. We also derive an expression for the center of mass of a set of points with respect to this distance.

Information Retrieval · Computer Science 2007-05-23 Jean-Luc Falcone , Paul Albuquerque

Random forests is a common non-parametric regression technique which performs well for mixed-type data and irrelevant covariates, while being robust to monotonic variable transformations. Existing random forest implementations target…

Machine Learning · Statistics 2018-05-04 Taylor Pospisil , Ann B. Lee

Conventionally, regression discontinuity analysis contrasts a univariate regression's limits as its independent variable, $R$, approaches a cut-point, $c$, from either side. Alternative methods target the average treatment effect in a small…

Applications · Statistics 2021-06-21 Adam C. Sales , Ben B. Hansen

We propose a coefficient that measures the dependence among large values for spatial processes of maxima. Its main properties are: a) $k$ locations can be taken into account; b) it takes values in $[0,1]$ and higher values indicate stronger…

Statistics Theory · Mathematics 2015-06-22 Helena Ferreira , Luisa Pereira

We consider a residuals-based distributionally robust optimization (DRO) model, where the underlying uncertainty depends on both covariate information and our decisions. We adopt both parametric and nonparametric regression models to learn…

Optimization and Control · Mathematics 2026-05-21 Qing Zhu , Xian Yu , Guzin Bayraksan

The spectra of random feature matrices provide essential information on the conditioning of the linear system used in random feature regression problems and are thus connected to the consistency and generalization of random feature models.…

Machine Learning · Statistics 2022-12-13 Zhijun Chen , Hayden Schaeffer , Rachel Ward

Additive models play an essential role in studying non-linear relationships. Despite many recent advances in estimation, there is a lack of methods and theories for inference in high-dimensional additive models, including confidence…

Statistics Theory · Mathematics 2022-02-18 Zijian Guo , Wei Yuan , Cun-Hui Zhang

Identifying statistical dependence between the features and the label is a fundamental problem in supervised learning. This paper presents a framework for estimating dependence between numerical features and a categorical label using…

Machine Learning · Computer Science 2021-10-01 Silu Zhang , Xin Dang , Dao Nguyen , Dawn Wilkins , Yixin Chen
‹ Prev 1 8 9 10 Next ›