English

Adaptation of Hierarchical Structured Models for Speech Act Recognition in Asynchronous Conversation

Computation and Language 2019-04-09 v1 Machine Learning Machine Learning

Abstract

We address the problem of speech act recognition (SAR) in asynchronous conversations (forums, emails). Unlike synchronous conversations (e.g., meetings, phone), asynchronous domains lack large labeled datasets to train an effective SAR model. In this paper, we propose methods to effectively leverage abundant unlabeled conversational data and the available labeled data from synchronous domains. We carry out our research in three main steps. First, we introduce a neural architecture based on hierarchical LSTMs and conditional random fields (CRF) for SAR, and show that our method outperforms existing methods when trained on in-domain data only. Second, we improve our initial SAR models by semi-supervised learning in the form of pretrained word embeddings learned from a large unlabeled conversational corpus. Finally, we employ adversarial training to improve the results further by leveraging the labeled data from synchronous domains and by explicitly modeling the distributional shift in two domains.

Keywords

Cite

@article{arxiv.1904.04021,
  title  = {Adaptation of Hierarchical Structured Models for Speech Act Recognition in Asynchronous Conversation},
  author = {Tasnim Mohiuddin and Thanh-Tung Nguyen and Shafiq Joty},
  journal= {arXiv preprint arXiv:1904.04021},
  year   = {2019}
}

Comments

To appear in NAACL 2019

R2 v1 2026-06-23T08:32:48.813Z