中文
相关论文

相关论文: Super Kawaii Vocalics: Amplifying the "Cute" Facto…

200 篇论文

Humans vary their expressivity when speaking for extended periods to maintain engagement with their listener. Although social robots tend to be deployed with ``expressive'' joyful voices, they lack this long-term variation found in human…

机器人学 · 计算机科学 2025-07-31 Paige Tuttösí , Shivam Mehta , Zachary Syvenky , Bermet Burkanova , Gustav Eje Henter , Angelica Lim

With the rise of human-machine communication, machines are increasingly designed with humanlike characteristics, such as gender, which can inadvertently trigger cognitive biases. Many conversational agents (CAs), such as voice assistants…

人机交互 · 计算机科学 2024-01-09 Weizi Liu

As machine learning approaches are increasingly used to augment human decision-making, eXplainable Artificial Intelligence (XAI) research has explored methods for communicating system behavior to humans. However, these approaches often fail…

计算机视觉与模式识别 · 计算机科学 2021-10-18 Luke Guerdan , Alex Raymond , Hatice Gunes

Do speakers of different languages talk differently about what they see? Behavioural and cognitive studies report cultural effects on perception; however, these are mostly limited in scope and hard to replicate. In this work, we conduct the…

计算与语言 · 计算机科学 2024-10-15 Uri Berger , Edoardo M. Ponti

Voice Cloning has rapidly advanced in today's digital world, with many researchers and corporations working to improve these algorithms for various applications. This article aims to establish a standardized terminology for voice cloning…

声音 · 计算机科学 2025-05-02 Hussam Azzuni , Abdulmotaleb El Saddik

We introduce PuppetAI, a modular soft robot interaction platform. This platform offers a scalable cable-driven actuation system and a customizable, puppet-inspired robot gesture framework, supporting a multitude of interaction gesture robot…

人机交互 · 计算机科学 2026-05-05 Jiaye Li , Tongshun Chen , Siyi Ma , Elizabeth Churchill , Ke Wu

The sweet spot can be interpreted as the region where acoustic sources create a spatial auditory illusion. We study the problem of maximizing this sweet spot when reproducing a desired sound wave using an array of loudspeakers. To achieve…

音频与语音处理 · 电气工程与系统科学 2022-05-05 Pedro Izquierdo Lehmann , Rodrigo F. Cadiz , Carlos A. Sing Long

We present pathways of investigation regarding conversational user interfaces (CUIs) for children in the classroom. We highlight anticipated challenges to be addressed in order to advance knowledge on CUIs for children. Further, we discuss…

人机交互 · 计算机科学 2021-12-02 Garrett Allen , Jie Yang , Maria Soledad Pera , Ujwal Gadiraju

Recent work in computer vision has yielded impressive results in automatically describing images with natural language. Most of these systems generate captions in a sin- gle language, requiring multiple language-specific models to build a…

计算机视觉与模式识别 · 计算机科学 2017-06-21 Satoshi Tsutsui , David Crandall

Text-based communication, such as text chat, is commonly employed in various contexts, both professional and personal. However, it lacks the rich emotional cues present in verbal and visual forms of communication, such as facial expressions…

人机交互 · 计算机科学 2023-09-14 Rintaro Chujo , Atsunobu Suzuki , Ari Hautasaari

This study extracted and analyzed the linguistic speech patterns that characterize Japanese anime or game characters. Conventional morphological analyzers, such as MeCab, segment words with high performance, but they are unable to segment…

计算与语言 · 计算机科学 2022-03-08 Mika Kishino , Kanako Komiya

Wake-up words (WUW) is a short sentence used to activate a speech recognition system to receive the user's speech input. WUW utterances include not only the lexical information for waking up the system but also non-lexical information such…

声音 · 计算机科学 2022-11-08 Taesu Kim , SeungHeon Doh , Gyunpyo Lee , Hyungseok Jeon , Juhan Nam , Hyeon-Jeong Suk

Audio captioning aims at describing the content of audio clips with human language. Due to the ambiguity of audio, different people may perceive the same audio differently, resulting in caption disparities (i.e., one audio may correlate to…

声音 · 计算机科学 2022-04-19 Yiming Zhang , Hong Yu , Ruoyi Du , Zhanyu Ma , Yuan Dong

Audio deepfakes have reached a level of realism that makes it increasingly difficult to distinguish between human and artificial voices, which poses risks such as identity theft or spread of disinformation. Despite these concerns, research…

音频与语音处理 · 电气工程与系统科学 2025-12-11 Eugenia San Segundo , Aurora López-Jareño , Xin Wang , Junichi Yamagishi

ChatGPT is a conversational agent built on a large language model. Trained on a significant portion of human output, ChatGPT can mimic people to a degree. As such, we need to consider what social identities ChatGPT simulates (or can be…

人机交互 · 计算机科学 2024-05-15 Takao Fujii , Katie Seaborn , Madeleine Steeds

It can be difficult to critically reflect on technology that has become part of everyday rituals and routines. To combat this, speculative and fictional approaches have previously been used by HCI to decontextualise the familiar and imagine…

人机交互 · 计算机科学 2020-03-09 William Seymour , Max Van Kleek

Humans have clear cross-modal preferences when matching certain novel words to visual shapes. Evidence suggests that these preferences play a prominent role in our linguistic processing, language learning, and the origins of signal-meaning…

计算与语言 · 计算机科学 2024-07-26 Tessa Verhoef , Kiana Shahrasbi , Tom Kouwenhoven

This paper explores the growing presence of emotionally responsive artificial intelligence through a critical and interdisciplinary lens. Bringing together the voices of early-career researchers from multiple fields, it explores how AI…

Recent years have seen an explosion in the availability of Voice User Interfaces. However, user surveys suggest that there are issues with respect to usability, and it has been hypothesised that contemporary voice-enabled systems are…

人机交互 · 计算机科学 2019-07-29 Roger K. Moore

Current computational-emotion research has focused on applying acoustic properties to analyze how emotions are perceived mathematically or used in natural language processing machine learning models. While recent interest has focused on…

声音 · 计算机科学 2021-07-06 Daniel Szelogowski