English
Related papers

Related papers: Super Kawaii Vocalics: Amplifying the "Cute" Facto…

200 papers

Persuasion is a fundamental aspect of communication, influencing decision-making across diverse contexts, from everyday conversations to high-stakes scenarios such as politics, marketing, and law. The rise of conversational AI systems has…

Computation and Language · Computer Science 2026-03-24 Nimet Beyza Bozdag , Shuhaib Mehri , Xiaocheng Yang , Hyeonjeong Ha , Zirui Cheng , Esin Durmus , Jiaxuan You , Heng Ji , Gokhan Tur , Dilek Hakkani-Tür

Computational developments--particularly artificial intelligence--are reshaping social scientific research and raise new questions for in-depth methods such as ethnography and qualitative interviewing. Building on classic debates about…

Computers and Society · Computer Science 2025-11-13 Corey M. Abramson , Tara Prendergast , Zhuofan Li , Daniel Dohan

Voice-enabled interactions provide more human-like experiences in many popular IoT systems. Cloud-based speech analysis services extract useful information from voice input using speech recognition techniques. The voice signal is a rich…

Cryptography and Security · Computer Science 2019-08-13 Ranya Aloufi , Hamed Haddadi , David Boyle

This paper focuses on the often-overlooked aspect of perceived voice femininity in singing voices. While existing research has examined perceived voice femininity in speech, the same concept has not yet been studied in singing voice. The…

Sound · Computer Science 2025-11-05 Yuexuan Kong , Viet-Anh Tran , Romain Hennequin

As voice assistants (VAs) become increasingly integrated into daily life, the need for emotion-aware systems that can recognize and respond appropriately to user emotions has grown. While significant progress has been made in speech emotion…

Human-Computer Interaction · Computer Science 2025-02-24 Yong Ma , Yuchong Zhang , Di Fu , Stephanie Zubicueta Portales , Danica Kragic , Morten Fjeld

Voice assistants (VAs) are typically evaluated through task performance metrics and self-report questionnaires, but people's voices themselves carry rich paralinguistic cues that reveal affect, effort, and interaction breakdowns. We present…

Human-Computer Interaction · Computer Science 2026-03-23 Yong Ma , Xuesong Zhang , Xuedong Zhang , Natalia Bartłomiejczyk , Seungwoo Je , Adrian Holzer , Morten Fjeld , Andreas Butz

Cognitive Behavioral Therapy (CBT) is a goal-oriented psychotherapy for mental health concerns implemented in a conversational setting with broad empirical support for its effectiveness across a range of presenting problems and client…

Audio and Speech Processing · Electrical Eng. & Systems 2020-10-16 Zhuohao Chen , Nikolaos Flemotomos , Victor Ardulov , Torrey A. Creed , Zac E. Imel , David C. Atkins , Shrikanth Narayanan

We present the first English corpus study on abusive language towards three conversational AI systems gathered "in the wild": an open-domain social bot, a rule-based chatbot, and a task-based system. To account for the complexity of the…

Computation and Language · Computer Science 2021-09-21 Amanda Cercas Curry , Gavin Abercrombie , Verena Rieser

The advancements of AI-synthesized human voices have introduced a growing threat of impersonation and disinformation. It is therefore of practical importance to developdetection methods for synthetic human voices. This work proposes a new…

Sound · Computer Science 2023-04-28 Chengzhe Sun , Shan Jia , Shuwei Hou , Ehab AlBadawy , Siwei Lyu

Voice cloning is the task of learning to synthesize the voice of an unseen speaker from a few samples. While current voice cloning methods achieve promising results in Text-to-Speech (TTS) synthesis for a new voice, these approaches lack…

Sound · Computer Science 2021-02-02 Paarth Neekhara , Shehzeen Hussain , Shlomo Dubnov , Farinaz Koushanfar , Julian McAuley

As social robots become more common, many have adopted cute aesthetics aiming to enhance user comfort and acceptance. However, the effect of this aesthetic choice on human feedback in reinforcement learning scenarios remains unclear.…

Robotics · Computer Science 2025-02-10 Jessica Eggers , Angela Dai , Matthew C. Gombolay

Existing Voice Cloning (VC) tasks aim to convert a paragraph text to a speech with desired voice specified by a reference audio. This has significantly boosted the development of artificial speech applications. However, there also exist…

Computer Vision and Pattern Recognition · Computer Science 2021-11-29 Qi Chen , Yuanqing Li , Yuankai Qi , Jiaqiu Zhou , Mingkui Tan , Qi Wu

Role-playing chatbots built on large language models have drawn interest, but better techniques are needed to enable mimicking specific fictional characters. We propose an algorithm that controls language models via an improved prompt and…

Computation and Language · Computer Science 2023-08-21 Cheng Li , Ziang Leng , Chenxi Yan , Junyi Shen , Hao Wang , Weishi MI , Yaying Fei , Xiaoyang Feng , Song Yan , HaoSheng Wang , Linkang Zhan , Yaokai Jia , Pingyu Wu , Haozhen Sun

We construct Japanese Idol Speech Corpus (JIS) to advance research in speech generation AI, including text-to-speech synthesis (TTS) and voice conversion (VC). JIS will facilitate more rigorous evaluations of speaker similarity in TTS and…

Sound · Computer Science 2025-07-17 Yuto Kondo , Hirokazu Kameoka , Kou Tanaka , Takuhiro Kaneko

Cartoonization is a task that renders natural photos into cartoon styles. Previous deep cartoonization methods only have focused on end-to-end translation, which may hinder editability. Instead, we propose a novel solution with editing…

Computer Vision and Pattern Recognition · Computer Science 2023-03-09 Namhyuk Ahn , Patrick Kwon , Jihye Back , Kibeom Hong , Seungkwon Kim

We present JNV (Japanese Nonverbal Vocalizations) corpus, a corpus of Japanese nonverbal vocalizations (NVs) with diverse phrases and emotions. Existing Japanese NV corpora lack phrase or emotion diversity, which makes it difficult to…

Sound · Computer Science 2023-05-23 Detai Xin , Shinnosuke Takamichi , Hiroshi Saruwatari

Speech technology systems struggle with many downstream tasks for child speech due to small training corpora and the difficulties that child speech pose. We apply a novel dataset, SpeechMaturity, to state-of-the-art transformer models to…

Computation and Language · Computer Science 2025-06-11 Theo Zhang , Madurya Suresh , Anne S. Warlaumont , Kasia Hitczenko , Alejandrina Cristia , Margaret Cychosz

This paper explores public perceptions around the role of affective computing in the workplace. It uses a series of design fictions with 46 UK based participants, unpacking their perspectives on the advantages and disadvantages of tracking…

Human-Computer Interaction · Computer Science 2022-05-18 Lachlan Urquhart , Alex Laffer , Diana Miranda

Pervasive voice interaction enables deceptive patterns through subtle voice characteristics, yet empirical investigation into this manipulation lags behind, especially within major non-English language contexts. Addressing this gap, our…

Human-Computer Interaction · Computer Science 2025-08-04 Shuning Zhang , Han Chen , Yabo Wang , Yiqun Xu , Jiaqi Bai , Yuanyuan Wu , Shixuan Li , Xin Yi , Chunhui Wang , Hewu Li

We present a method for automatically producing human-like vocal imitations of sounds: the equivalent of "sketching," but for auditory rather than visual representation. Starting with a simulated model of the human vocal tract, we first try…

Graphics · Computer Science 2024-09-23 Matthew Caren , Kartik Chandra , Joshua B. Tenenbaum , Jonathan Ragan-Kelley , Karima Ma