English

Webpage Segmentation for Extracting Images and Their Surrounding Contextual Information

Multimedia 2020-05-21 v1 Computer Vision and Pattern Recognition Information Retrieval

Abstract

Web images come in hand with valuable contextual information. Although this information has long been mined for various uses such as image annotation, clustering of images, inference of image semantic content, etc., insufficient attention has been given to address issues in mining this contextual information. In this paper, we propose a webpage segmentation algorithm targeting the extraction of web images and their contextual information based on their characteristics as they appear on webpages. We conducted a user study to obtain a human-labeled dataset to validate the effectiveness of our method and experiments demonstrated that our method can achieve better results compared to an existing segmentation algorithm.

Keywords

Cite

@article{arxiv.2005.09639,
  title  = {Webpage Segmentation for Extracting Images and Their Surrounding Contextual Information},
  author = {F. Fauzi and H. J. Long and M. Belkhatir},
  journal= {arXiv preprint arXiv:2005.09639},
  year   = {2020}
}

Comments

arXiv admin note: substantial text overlap with arXiv:2005.02156

R2 v1 2026-06-23T15:40:07.425Z