English

Local identifiability of $l_1$-minimization dictionary learning: a sufficient and almost necessary condition

Machine Learning 2016-07-13 v2

Abstract

We study the theoretical properties of learning a dictionary from NN signals xiRK\mathbf x_i\in \mathbb R^K for i=1,...,Ni=1,...,N via l1l_1-minimization. We assume that xi\mathbf x_i's are i.i.d.i.i.d. random linear combinations of the KK columns from a complete (i.e., square and invertible) reference dictionary D0RK×K\mathbf D_0 \in \mathbb R^{K\times K}. Here, the random linear coefficients are generated from either the ss-sparse Gaussian model or the Bernoulli-Gaussian model. First, for the population case, we establish a sufficient and almost necessary condition for the reference dictionary D0\mathbf D_0 to be locally identifiable, i.e., a local minimum of the expected l1l_1-norm objective function. Our condition covers both sparse and dense cases of the random linear coefficients and significantly improves the sufficient condition by Gribonval and Schnass (2010). In addition, we show that for a complete μ\mu-coherent reference dictionary, i.e., a dictionary with absolute pairwise column inner-product at most μ[0,1)\mu\in[0,1), local identifiability holds even when the random linear coefficient vector has up to O(μ2)O(\mu^{-2}) nonzeros on average. Moreover, our local identifiability results also translate to the finite sample case with high probability provided that the number of signals NN scales as O(KlogK)O(K\log K).

Keywords

Cite

@article{arxiv.1505.04363,
  title  = {Local identifiability of $l_1$-minimization dictionary learning: a sufficient and almost necessary condition},
  author = {Siqi Wu and Bin Yu},
  journal= {arXiv preprint arXiv:1505.04363},
  year   = {2016}
}