中文
相关论文

相关论文: Kawaii Game Vocalics: A Preliminary Model

200 篇论文

Enhancing AI systems with efficient communication skills for effective human assistance necessitates proactive initiatives from the system side to discern specific circumstances and interact aptly. This research focuses on a collective…

计算与语言 · 计算机科学 2024-02-09 Jack Zhang

This paper presents the first large-scale analysis of public-facing chatbots on Character$.$AI, a rapidly growing social media platform where users create and interact with chatbots. Character$.$AI is distinctive in that it merges…

社会与信息网络 · 计算机科学 2026-03-16 Owen Lee , Kenneth Joseph

In text-to-speech, controlling voice characteristics is important in achieving various-purpose speech synthesis. Considering the success of text-conditioned generation, such as text-to-image, free-form text instruction should be useful for…

In recent years, several high-performance conversational systems have been proposed based on the Transformer encoder-decoder model. Although previous studies analyzed the effects of the model parameters and the decoding method on subjective…

Gender bias in machine translation (MT) systems has been extensively documented, but bias in automatic quality estimation (QE) metrics remains comparatively underexplored. Existing studies suggest that QE metrics can also exhibit gender…

A voice AI agent that blends seamlessly into daily life would interact with humans in an autonomous, real-time, and emotionally expressive manner. Rather than merely reacting to commands, it would continuously listen, reason, and respond…

人工智能 · 计算机科学 2025-05-06 Yemin Shi , Yu Shu , Siwei Dong , Guangyi Liu , Jaward Sesay , Jingwen Li , Zhiting Hu

Rises in the number of animal abuse cases are reported around the world. While chatbots have been effective in influencing their users' perceptions and behaviors, little if any research has hitherto explored the design of chatbots that…

人机交互 · 计算机科学 2025-12-08 Jingshu Li , Aaditya Patwari , Yi-Chieh Lee

Zero-shot talking avatar generation aims at synthesizing natural talking videos from speech and a single portrait image. Previous methods have relied on domain-specific heuristics such as warping-based motion representation and 3D Morphable…

计算机视觉与模式识别 · 计算机科学 2024-03-15 Tianyu He , Junliang Guo , Runyi Yu , Yuchi Wang , Jialiang Zhu , Kaikai An , Leyi Li , Xu Tan , Chunyu Wang , Han Hu , HsiangTao Wu , Sheng Zhao , Jiang Bian

ChatGPT is a conversational agent built on a large language model. Trained on a significant portion of human output, ChatGPT can mimic people to a degree. As such, we need to consider what social identities ChatGPT simulates (or can be…

人机交互 · 计算机科学 2024-05-15 Takao Fujii , Katie Seaborn , Madeleine Steeds

AI agents -- systems that execute multi-step reasoning workflows with persistent state, tool access, and specialist skills -- represent a qualitative shift from prior automation technologies in social science. Unlike chatbots that respond…

人工智能 · 计算机科学 2026-03-10 Yongjun Zhang

In this demo paper, we introduce SAPIEN, a platform for high-fidelity virtual agents driven by large language models that can hold open domain conversations with users in 13 different languages, and display emotions through facial…

人机交互 · 计算机科学 2024-01-23 Masum Hasan , Cengiz Ozel , Sammy Potter , Ehsan Hoque

In this work we investigate how children ages 5-12 perceive, understand, and use generative AI models such as a text-based LLMs ChatGPT and a visual-based model DALL-E. Generative AI is newly being used widely since chatGPT. Children are…

人机交互 · 计算机科学 2024-05-24 Eliza Kosoy , Soojin Jeong , Anoop Sinha , Alison Gopnik , Tanya Kraljic

Japan faces many challenges related to its aging society, including increasing rates of cognitive decline in the population and a shortage of caregivers. Efforts have begun to explore solutions using artificial intelligence (AI), especially…

Korean aegyo is a socially recognized childlike speaking style used predominantly in romantic interactions among adults. This study examined vowel space modification in aegyo by analyzing formant frequencies from twelve Seoul Korean…

计算与语言 · 计算机科学 2026-04-29 Ji-eun Kim , Volker Dellwo

Conversational user interfaces (CUIs) have become an everyday technology for people the world over, as well as a booming area of research. Advances in voice synthesis and the emergence of chatbots powered by large language models (LLMs),…

Audio deepfakes have reached a level of realism that makes it increasingly difficult to distinguish between human and artificial voices, which poses risks such as identity theft or spread of disinformation. Despite these concerns, research…

音频与语音处理 · 电气工程与系统科学 2025-12-11 Eugenia San Segundo , Aurora López-Jareño , Xin Wang , Junichi Yamagishi

Voice Assistants (VAs) can assist users in various everyday tasks, but many users are reluctant to rely on VAs for intricate tasks like online shopping. This study aims to examine whether the vocal characteristics of VAs can serve as an…

人机交互 · 计算机科学 2024-06-14 Sabid Bin Habib Pias , Ran Huang , Donald Williamson , Minjeong Kim , Apu Kapadia

Video game localisation, a field highly impacted by the lack of visual environment and text linearity, forces translators to create inclusive solutions in terms of gender to overcome the hurdles created by variables. This paper will…

计算机与社会 · 计算机科学 2024-09-25 Maria Isabel Rivas Ginel , Sarah Theroine

Machine learning is frequently used in affective computing, but presents challenges due the opacity of state-of-the-art machine learning methods. Because of the impact affective machine learning systems may have on an individual's life, it…

机器学习 · 计算机科学 2025-10-07 David S. Johnson , Olya Hakobyan , Hanna Drimalla

The dialogue experience with conversational agents can be greatly enhanced with multimodal and immersive interactions in virtual reality. In this work, we present an open-source architecture with the goal of simplifying the development of…

人工智能 · 计算机科学 2023-08-08 Michele Yin , Gabriel Roccabruna , Abhinav Azad , Giuseppe Riccardi