English

Extraction of Salient Sentences from Labelled Documents

Computation and Language 2015-03-03 v2 Information Retrieval Machine Learning

Abstract

We present a hierarchical convolutional document model with an architecture designed to support introspection of the document structure. Using this model, we show how to use visualisation techniques from the computer vision literature to identify and extract topic-relevant sentences. We also introduce a new scalable evaluation technique for automatic sentence extraction systems that avoids the need for time consuming human annotation of validation data.

Keywords

Cite

@article{arxiv.1412.6815,
  title  = {Extraction of Salient Sentences from Labelled Documents},
  author = {Misha Denil and Alban Demiraj and Nando de Freitas},
  journal= {arXiv preprint arXiv:1412.6815},
  year   = {2015}
}

Comments

arXiv admin note: substantial text overlap with arXiv:1406.3830

R2 v1 2026-06-22T07:39:57.229Z