English

Predicting User Engagement Status for Online Evaluation of Intelligent Assistants

Computation and Language 2021-06-02 v2 Human-Computer Interaction

Abstract

Evaluation of intelligent assistants in large-scale and online settings remains an open challenge. User behavior-based online evaluation metrics have demonstrated great effectiveness for monitoring large-scale web search and recommender systems. Therefore, we consider predicting user engagement status as the very first and critical step to online evaluation for intelligent assistants. In this work, we first proposed a novel framework for classifying user engagement status into four categories -- fulfillment, continuation, reformulation and abandonment. We then demonstrated how to design simple but indicative metrics based on the framework to quantify user engagement levels. We also aim for automating user engagement prediction with machine learning methods. We compare various models and features for predicting engagement status using four real-world datasets. We conducted detailed analyses on features and failure cases to discuss the performance of current models as well as challenges.

Keywords

Cite

@article{arxiv.2010.00656,
  title  = {Predicting User Engagement Status for Online Evaluation of Intelligent Assistants},
  author = {Rui Meng and Zhen Yue and Alyssa Glass},
  journal= {arXiv preprint arXiv:2010.00656},
  year   = {2021}
}

Comments

Paper has been accepted by ECIR 2021 (43rd edition of the annual European Conference on Information Retrieval)