English
Related papers

Related papers: Minimum HGR Correlation Principle: From Marginals …

200 papers

Consider the binary classification problem of predicting a target variable $Y$ from a discrete feature vector $X = (X_1,...,X_d)$. When the probability distribution $\mathbb{P}(X,Y)$ is known, the optimal classifier, leading to the minimum…

Machine Learning · Computer Science 2015-11-06 Meisam Razaviyayn , Farzan Farnia , David Tse

The Hirschfeld-Gebelein-R\'enyi (HGR) correlation coefficient is an extension of Pearson's correlation that is not limited to linear correlations, with potential applications in algorithmic fairness, scientific analysis, and causal…

Machine Learning · Computer Science 2025-09-12 Luca Giuliani , Michele Lombardi

For independent random variables $(X_i)_{1\leq i\leq n}$, we consider the maximal correlation coefficient $R=R(\min_{i:1\leq i\leq m}X_i,\min_{j:\ell+1\leq j\leq n}X_j)$. If $X_1,X_2,\ldots,X_n$ are identically distributed with the same…

Probability · Mathematics 2026-03-27 Yinshan Chang , Qinwei Chen

Given two discrete random variables $X$ and $Y$, with probability distributions ${\bf p} =(p_1, \ldots , p_n)$ and ${\bf q}=(q_1, \ldots , q_m)$, respectively, denote by ${\cal C}({\bf p}, {\bf q})$ the set of all couplings of ${\bf p}$ and…

Information Theory · Computer Science 2017-03-29 Ferdinando Cicalese , Luisa Gargano , Ugo Vaccaro

Consider a bivariate Geometric random variable where the first component has parameter $p_1$ and the second parameter $p_2$. It is not possible to make the correlation between the marginals equal to -1. Here the properties of this minimum…

Probability · Mathematics 2014-08-29 Mark Huber , Nevena Maric

The maximal (or Hilbertian) correlation coefficient between two random variables X and Y, denoted by \{X:Y\}, is the supremum of the |Corr(f(X),g(Y))| for real measurable functions f, g, where "Corr" denotes Pearson's correlation…

Probability · Mathematics 2011-01-04 Remi Peyre

The joint distribution $P(X,Y)$ cannot be determined from its marginals $P(X)$ and $P(Y)$ alone; one also needs one of the conditionals $P(X|Y)$ or $P(Y|X)$. But is there a best guess, given only the marginals? Here we answer this question…

Statistics Theory · Mathematics 2020-05-26 Richard Rohwer

Consider the problem of drawing random variates $(X_1,\ldots,X_n)$ from a distribution where the marginal of each $X_i$ is specified, as well as the correlation between every pair $X_i$ and $X_j$. For given marginals, the…

Probability · Mathematics 2016-12-30 Mark Huber , Nevena Maric

Given two discrete random variables $X$ and $Y,$ with probability distributions ${\bf p}=(p_1, \ldots , p_n)$ and ${\bf q}=(q_1, \ldots , q_m)$, respectively, denote by ${\cal C}({\bf p}, {\bf q})$ the set of all couplings of ${\bf p}$ and…

Information Theory · Computer Science 2019-01-24 Ferdinando Cicalese , Luisa Gargano , Ugo Vaccaro

We consider a distributed logistic regression problem where labeled data pairs $(X_i,Y_i)\in \mathbb{R}^d\times\{-1,1\}$ for $i=1,\ldots,n$ are distributed across multiple machines in a network and must be communicated to a centralized…

Information Theory · Computer Science 2019-10-04 Leighton Pate Barnes , Ayfer Ozgur

We propose a lower bound on the log marginal likelihood of Gaussian process regression models that can be computed without matrix factorisation of the full kernel matrix. We show that approximate maximum likelihood learning of model…

Machine Learning · Statistics 2021-02-17 Artem Artemev , David R. Burt , Mark van der Wilk

We correct claims about lower bounds on mutual information (MI) between real-valued random variables made in A. Kraskov {\it et al.}, Phys. Rev. E {\bf 69}, 066138 (2004). We show that non-trivial lower bounds on MI in terms of linear…

Data Analysis, Statistics and Probability · Physics 2013-05-29 David V. Foster , Peter Grassberger

Hybrid Bayesian Networks (HBNs), which contain both discrete and continuous variables, arise naturally in many application areas (e.g., image understanding, data fusion, medical diagnosis, fraud detection). This paper concerns inference in…

Artificial Intelligence · Computer Science 2019-05-21 Cheol Young Park , Kathryn Blackmond Laskey , Paulo C. G. Costa , Shou Matsumoto

We construct optimal low-rank approximations for the Gaussian posterior distribution in linear Gaussian inverse problems with possibly infinite-dimensional separable Hilbert parameter spaces and finite-dimensional data spaces. We first…

Statistics Theory · Mathematics 2026-04-09 Giuseppe Carere , Han Cheng Lie

The Hirschfeld-Gebelein-R\'{e}nyi (HGR) maximal correlation and the corresponding functions have been shown useful in many machine learning scenarios. In this paper, we study the sample complexity of estimating the HGR maximal correlation…

Information Theory · Computer Science 2021-09-15 Shao-Lun Huang , Xiangxiang Xu

Many types of bounded data defined on the unit interval arise naturally as ratios of the form $X/(X + Y)$. In the existing literature, the main statistical models proposed for this type of bounded data typically based on the assumption that…

Methodology · Statistics 2026-03-04 Roberto Vila , Felipe Quintino , Marcelo Bourguignon

Graph cuts are among the most prominent tools for clustering and classification analysis. While intensively studied from geometric and algorithmic perspectives, graph cut-based statistical inference still remains elusive to a certain…

Statistics Theory · Mathematics 2025-12-11 Leo Suchan , Housen Li , Axel Munk

Many inference problems involving questions of optimality ask for the maximum or the minimum of a finite set of unknown quantities. This technical report derives the first two posterior moments of the maximum of two correlated Gaussian…

Machine Learning · Statistics 2009-10-02 Philipp Hennig

George R. Terrell (1983, {Ann. Probab., vol. 11(3), pp. 823--826) showed that the Pearson coefficient of correlation of an ordered pair from a random sample of size two is at most one-half, and the equality is attained only for rectangular…

Probability · Mathematics 2022-05-31 Nickos Papadatos

Let $X=(x_{ij})\in\mathbb{R}^{N\times n}$ be a rectangular random matrix with i.i.d. entries (we assume $N/n\to\mathbf{a}>1$), and denote by $\sigma_{min}(X)$ its smallest singular value. When entries have mean zero and unit second moment,…

Probability · Mathematics 2025-07-30 Yi Han
‹ Prev 1 2 3 10 Next ›