English

Using Web Page Titles to Rediscover Lost Web Pages

Information Retrieval 2010-02-15 v1

Abstract

Titles are denoted by the TITLE element within a web page. We queried the title against the the Yahoo search engine to determine the page's status (found, not found). We conducted several tests based on elements of the title. These tests were used to discern whether we could predict a pages status based on the title. Our results increase our ability to determine bad titles but not our ability to determine good titles.

Cite

@article{arxiv.1002.2439,
  title  = {Using Web Page Titles to Rediscover Lost Web Pages},
  author = {Jeffery L. Shipman and Martin Klein and Michael L. Nelson},
  journal= {arXiv preprint arXiv:1002.2439},
  year   = {2010}
}

Comments

49 pages, 18 figures, CS project report

R2 v1 2026-06-21T14:46:13.784Z