English

Generalization error bounds for learning to rank: Does the length of document lists matter?

Machine Learning 2016-08-24 v1

Abstract

We consider the generalization ability of algorithms for learning to rank at a query level, a problem also called subset ranking. Existing generalization error bounds necessarily degrade as the size of the document list associated with a query increases. We show that such a degradation is not intrinsic to the problem. For several loss functions, including the cross-entropy loss used in the well known ListNet method, there is \emph{no} degradation in generalization ability as document lists become longer. We also provide novel generalization error bounds under 1\ell_1 regularization and faster convergence rates if the loss function is smooth.

Keywords

Cite

@article{arxiv.1603.01860,
  title  = {Generalization error bounds for learning to rank: Does the length of document lists matter?},
  author = {Ambuj Tewari and Sougata Chaudhuri},
  journal= {arXiv preprint arXiv:1603.01860},
  year   = {2016}
}

Comments

Appeared in ICML 2015. arXiv admin note: substantial text overlap with arXiv:1405.0586

R2 v1 2026-06-22T13:04:45.724Z