English

An Annotation Scheme of A Large-scale Multi-party Dialogues Dataset for Discourse Parsing and Machine Comprehension

Computation and Language 2019-11-12 v1

Abstract

In this paper, we propose the scheme for annotating large-scale multi-party chat dialogues for discourse parsing and machine comprehension. The main goal of this project is to help understand multi-party dialogues. Our dataset is based on the Ubuntu Chat Corpus. For each multi-party dialogue, we annotate the discourse structure and question-answer pairs for dialogues. As we know, this is the first large scale corpus for multi-party dialogues discourse parsing, and we firstly propose the task for multi-party dialogues machine reading comprehension.

Keywords

Cite

@article{arxiv.1911.03514,
  title  = {An Annotation Scheme of A Large-scale Multi-party Dialogues Dataset for Discourse Parsing and Machine Comprehension},
  author = {Jiaqi Li and Ming Liu and Bing Qin and Zihao Zheng and Ting Liu},
  journal= {arXiv preprint arXiv:1911.03514},
  year   = {2019}
}