English

A statistical model for word discovery in child directed speech

Computation and Language 2007-05-23 v1 Machine Learning

Abstract

A statistical model for segmentation and word discovery in child directed speech is presented. An incremental unsupervised learning algorithm to infer word boundaries based on this model is described and results of empirical tests showing that the algorithm is competitive with other models that have been used for similar tasks are also presented.

Keywords

Cite

@article{arxiv.cs/9910011,
  title  = {A statistical model for word discovery in child directed speech},
  author = {Anand Venkataraman},
  journal= {arXiv preprint arXiv:cs/9910011},
  year   = {2007}
}

Comments

48 pgs, 10 figs

R2 v1 2026-07-22T12:29:22.635Z