Product/Brand extraction from WikiPedia
Information Retrieval
2012-12-14 v1 Artificial Intelligence
Abstract
In this paper we describe the task of extracting product and brand pages from wikipedia. We present an experimental environment and setup built on top of a dataset of wikipedia pages we collected. We introduce a method for recognition of product pages modelled as a boolean probabilistic classification task. We show that this approach can lead to promising results and we discuss alternative approaches we considered.
Cite
@article{arxiv.1212.3013,
title = {Product/Brand extraction from WikiPedia},
author = {K. Massoudi and G. Modena},
journal= {arXiv preprint arXiv:1212.3013},
year = {2012}
}
Comments
17 pages. Manuscript first creation date: November 27, 2009. At the time of first creation both authors were affiliated with the University of Amsterdam (The Netherlands)