中文

Hierarchical Clustering Using Mutual Information

定量方法 2007-05-23 v1 计算复杂性 数据分析、统计与概率

摘要

We present a method for hierarchical clustering of data called {\it mutual information clustering} (MIC) algorithm. It uses mutual information (MI) as a similarity measure and exploits its grouping property: The MI between three objects X,Y,X, Y, and ZZ is equal to the sum of the MI between XX and YY, plus the MI between ZZ and the combined object (XY)(XY). We use this both in the Shannon (probabilistic) version of information theory and in the Kolmogorov (algorithmic) version. We apply our method to the construction of phylogenetic trees from mitochondrial DNA sequences and to the output of independent components analysis (ICA) as illustrated with the ECG of a pregnant woman.

关键词

引用

@article{arxiv.q-bio/0311037,
  title  = {Hierarchical Clustering Using Mutual Information},
  author = {Alexander Kraskov and Harald Stoegbauer and Ralph G. Andrzejak and Peter Grassberger},
  journal= {arXiv preprint arXiv:q-bio/0311037},
  year   = {2007}
}

备注

4 pages, 4 figures