English

Joint universal lossy coding and identification of stationary mixing sources with general alphabets

Information Theory 2016-11-15 v1 Machine Learning math.IT

Abstract

We consider the problem of joint universal variable-rate lossy coding and identification for parametric classes of stationary β\beta-mixing sources with general (Polish) alphabets. Compression performance is measured in terms of Lagrangians, while identification performance is measured by the variational distance between the true source and the estimated source. Provided that the sources are mixing at a sufficiently fast rate and satisfy certain smoothness and Vapnik-Chervonenkis learnability conditions, it is shown that, for bounded metric distortions, there exist universal schemes for joint lossy compression and identification whose Lagrangian redundancies converge to zero as Vnlogn/n\sqrt{V_n \log n /n} as the block length nn tends to infinity, where VnV_n is the Vapnik-Chervonenkis dimension of a certain class of decision regions defined by the nn-dimensional marginal distributions of the sources; furthermore, for each nn, the decoder can identify nn-dimensional marginal of the active source up to a ball of radius O(Vnlogn/n)O(\sqrt{V_n\log n/n}) in variational distance, eventually with probability one. The results are supplemented by several examples of parametric sources satisfying the regularity conditions.

Keywords

Cite

@article{arxiv.0901.1904,
  title  = {Joint universal lossy coding and identification of stationary mixing sources with general alphabets},
  author = {Maxim Raginsky},
  journal= {arXiv preprint arXiv:0901.1904},
  year   = {2016}
}

Comments

16 pages, 1 figure; accepted to IEEE Transactions on Information Theory