English
Related papers

Related papers: Super Kawaii Vocalics: Amplifying the "Cute" Facto…

200 papers

Humans vary their expressivity when speaking for extended periods to maintain engagement with their listener. Although social robots tend to be deployed with ``expressive'' joyful voices, they lack this long-term variation found in human…

With the rise of human-machine communication, machines are increasingly designed with humanlike characteristics, such as gender, which can inadvertently trigger cognitive biases. Many conversational agents (CAs), such as voice assistants…

Human-Computer Interaction · Computer Science 2024-01-09 Weizi Liu

As machine learning approaches are increasingly used to augment human decision-making, eXplainable Artificial Intelligence (XAI) research has explored methods for communicating system behavior to humans. However, these approaches often fail…

Computer Vision and Pattern Recognition · Computer Science 2021-10-18 Luke Guerdan , Alex Raymond , Hatice Gunes

Do speakers of different languages talk differently about what they see? Behavioural and cognitive studies report cultural effects on perception; however, these are mostly limited in scope and hard to replicate. In this work, we conduct the…

Computation and Language · Computer Science 2024-10-15 Uri Berger , Edoardo M. Ponti

Voice Cloning has rapidly advanced in today's digital world, with many researchers and corporations working to improve these algorithms for various applications. This article aims to establish a standardized terminology for voice cloning…

Sound · Computer Science 2025-05-02 Hussam Azzuni , Abdulmotaleb El Saddik

We introduce PuppetAI, a modular soft robot interaction platform. This platform offers a scalable cable-driven actuation system and a customizable, puppet-inspired robot gesture framework, supporting a multitude of interaction gesture robot…

Human-Computer Interaction · Computer Science 2026-05-05 Jiaye Li , Tongshun Chen , Siyi Ma , Elizabeth Churchill , Ke Wu

The sweet spot can be interpreted as the region where acoustic sources create a spatial auditory illusion. We study the problem of maximizing this sweet spot when reproducing a desired sound wave using an array of loudspeakers. To achieve…

Audio and Speech Processing · Electrical Eng. & Systems 2022-05-05 Pedro Izquierdo Lehmann , Rodrigo F. Cadiz , Carlos A. Sing Long

We present pathways of investigation regarding conversational user interfaces (CUIs) for children in the classroom. We highlight anticipated challenges to be addressed in order to advance knowledge on CUIs for children. Further, we discuss…

Human-Computer Interaction · Computer Science 2021-12-02 Garrett Allen , Jie Yang , Maria Soledad Pera , Ujwal Gadiraju

Recent work in computer vision has yielded impressive results in automatically describing images with natural language. Most of these systems generate captions in a sin- gle language, requiring multiple language-specific models to build a…

Computer Vision and Pattern Recognition · Computer Science 2017-06-21 Satoshi Tsutsui , David Crandall

Text-based communication, such as text chat, is commonly employed in various contexts, both professional and personal. However, it lacks the rich emotional cues present in verbal and visual forms of communication, such as facial expressions…

Human-Computer Interaction · Computer Science 2023-09-14 Rintaro Chujo , Atsunobu Suzuki , Ari Hautasaari

This study extracted and analyzed the linguistic speech patterns that characterize Japanese anime or game characters. Conventional morphological analyzers, such as MeCab, segment words with high performance, but they are unable to segment…

Computation and Language · Computer Science 2022-03-08 Mika Kishino , Kanako Komiya

Wake-up words (WUW) is a short sentence used to activate a speech recognition system to receive the user's speech input. WUW utterances include not only the lexical information for waking up the system but also non-lexical information such…

Sound · Computer Science 2022-11-08 Taesu Kim , SeungHeon Doh , Gyunpyo Lee , Hyungseok Jeon , Juhan Nam , Hyeon-Jeong Suk

Audio captioning aims at describing the content of audio clips with human language. Due to the ambiguity of audio, different people may perceive the same audio differently, resulting in caption disparities (i.e., one audio may correlate to…

Sound · Computer Science 2022-04-19 Yiming Zhang , Hong Yu , Ruoyi Du , Zhanyu Ma , Yuan Dong

Audio deepfakes have reached a level of realism that makes it increasingly difficult to distinguish between human and artificial voices, which poses risks such as identity theft or spread of disinformation. Despite these concerns, research…

Audio and Speech Processing · Electrical Eng. & Systems 2025-12-11 Eugenia San Segundo , Aurora López-Jareño , Xin Wang , Junichi Yamagishi

ChatGPT is a conversational agent built on a large language model. Trained on a significant portion of human output, ChatGPT can mimic people to a degree. As such, we need to consider what social identities ChatGPT simulates (or can be…

Human-Computer Interaction · Computer Science 2024-05-15 Takao Fujii , Katie Seaborn , Madeleine Steeds

It can be difficult to critically reflect on technology that has become part of everyday rituals and routines. To combat this, speculative and fictional approaches have previously been used by HCI to decontextualise the familiar and imagine…

Human-Computer Interaction · Computer Science 2020-03-09 William Seymour , Max Van Kleek

Humans have clear cross-modal preferences when matching certain novel words to visual shapes. Evidence suggests that these preferences play a prominent role in our linguistic processing, language learning, and the origins of signal-meaning…

Computation and Language · Computer Science 2024-07-26 Tessa Verhoef , Kiana Shahrasbi , Tom Kouwenhoven

This paper explores the growing presence of emotionally responsive artificial intelligence through a critical and interdisciplinary lens. Bringing together the voices of early-career researchers from multiple fields, it explores how AI…

Recent years have seen an explosion in the availability of Voice User Interfaces. However, user surveys suggest that there are issues with respect to usability, and it has been hypothesised that contemporary voice-enabled systems are…

Human-Computer Interaction · Computer Science 2019-07-29 Roger K. Moore

Current computational-emotion research has focused on applying acoustic properties to analyze how emotions are perceived mathematically or used in natural language processing machine learning models. While recent interest has focused on…

Sound · Computer Science 2021-07-06 Daniel Szelogowski