English

An Information Retrieval Approach to Finding Dependent Subspaces of Multiple Views

Machine Learning 2016-01-11 v2 Machine Learning

Abstract

Finding relationships between multiple views of data is essential both for exploratory analysis and as pre-processing for predictive tasks. A prominent approach is to apply variants of Canonical Correlation Analysis (CCA), a classical method seeking correlated components between views. The basic CCA is restricted to maximizing a simple dependency criterion, correlation, measured directly between data coordinates. We introduce a new method that finds dependent subspaces of views directly optimized for the data analysis task of \textit{neighbor retrieval between multiple views}. We optimize mappings for each view such as linear transformations to maximize cross-view similarity between neighborhoods of data samples. The criterion arises directly from the well-defined retrieval task, detects nonlinear and local similarities, is able to measure dependency of data relationships rather than only individual data coordinates, and is related to well understood measures of information retrieval quality. In experiments we show the proposed method outperforms alternatives in preserving cross-view neighborhood similarities, and yields insights into local dependencies between multiple views.

Keywords

Cite

@article{arxiv.1511.06423,
  title  = {An Information Retrieval Approach to Finding Dependent Subspaces of Multiple Views},
  author = {Ziyuan Lin and Jaakko Peltonen},
  journal= {arXiv preprint arXiv:1511.06423},
  year   = {2016}
}

Comments

9 pages, 15 figures. Submitted for ICLR 2016; the authors contributed equally

R2 v1 2026-06-22T11:49:59.690Z