English

Product/Brand extraction from WikiPedia

Information Retrieval 2012-12-14 v1 Artificial Intelligence

Abstract

In this paper we describe the task of extracting product and brand pages from wikipedia. We present an experimental environment and setup built on top of a dataset of wikipedia pages we collected. We introduce a method for recognition of product pages modelled as a boolean probabilistic classification task. We show that this approach can lead to promising results and we discuss alternative approaches we considered.

Cite

@article{arxiv.1212.3013,
  title  = {Product/Brand extraction from WikiPedia},
  author = {K. Massoudi and G. Modena},
  journal= {arXiv preprint arXiv:1212.3013},
  year   = {2012}
}

Comments

17 pages. Manuscript first creation date: November 27, 2009. At the time of first creation both authors were affiliated with the University of Amsterdam (The Netherlands)

R2 v1 2026-06-21T22:53:39.663Z