English

Building Proactive Voice Assistants: When and How (not) to Interact

Human-Computer Interaction 2020-05-05 v1

Abstract

Voice assistants have recently achieved remarkable commercial success. However, the current generation of these devices is typically capable of only reactive interactions. In other words, interactions have to be initiated by the user, which somewhat limits their usability and user experience. We propose, that the next generation of such devices should be able to proactively provide the right information in the right way at the right time, without being prompted by the user. However, achieving this is not straightforward, since there is the danger it could interrupt what the user is doing too much, resulting in it being distracting or even annoying. Furthermore, it could unwittingly, reveal sensitive/private information to third parties. In this report, we discuss the challenges of developing proactively initiated interactions, and suggest a framework for when it is appropriate for the device to intervene. To validate our design assumptions, we describe firstly, how we built a functioning prototype and secondly, a user study that was conducted to assess users' reactions and reflections when in the presence of a proactive voice assistant. This pre-print summarises the state, ideas and progress towards a proactive device as of autumn 2018.

Keywords

Cite

@article{arxiv.2005.01322,
  title  = {Building Proactive Voice Assistants: When and How (not) to Interact},
  author = {O. Miksik and I. Munasinghe and J. Asensio-Cubero and S. Reddy Bethi and S-T. Huang and S. Zylfo and X. Liu and T. Nica and A. Mitrocsak and S. Mezza and R. Beard and R. Shi and R. Ng and P. Mediano and Z. Fountas and S-H. Lee and J. Medvesek and H. Zhuang and Y. Rogers and P. Swietojanski},
  journal= {arXiv preprint arXiv:2005.01322},
  year   = {2020}
}

Comments

17 pages, technical report

R2 v1 2026-06-23T15:17:04.517Z