中文
相关论文

相关论文: Why Are Conversational Assistants Still Black Boxe…

200 篇论文

While the advances in artificial intelligence and machine learning empower a new generation of autonomous systems for assisting human performance, one major concern arises from the human factors perspective: Humans have difficulty…

人机交互 · 计算机科学 2020-08-04 Ruikun Luo , Na Du , X. Jessie Yang

We introduce CONCORD, a privacy-aware asynchronous assistant-to-assistant (A2A) framework that leverages collaboration between proactive speech-based AI. As agents evolve from reactive to always-listening assistants, they face a core…

人工智能 · 计算机科学 2026-04-16 Tanmay Srivastava , Amartya Basu , Shubham Jain , Vaishnavi Ranganathan

Equipped with artificial intelligence (AI) and advanced sensing capabilities, social robots are gaining interest among consumers in the United States. These robots seem like a natural evolution of traditional smart home devices. However,…

计算机与社会 · 计算机科学 2025-07-17 Henry Bell , Jabari Kwesi , Hiba Laabadli , Pardis Emami-Naeini

The growing use of voice user interfaces has led to a surge in the collection and storage of speech data. While data collection allows for the development of efficient tools powering most speech services, it also poses serious privacy…

密码学与安全 · 计算机科学 2024-03-04 Pierre Champion

The interactive use of large language models (LLMs) in AI assistants (at work, home, etc.) introduces a new set of inference-time privacy risks: LLMs are fed different types of information from multiple sources in their inputs and are…

人工智能 · 计算机科学 2024-07-02 Niloofar Mireshghallah , Hyunwoo Kim , Xuhui Zhou , Yulia Tsvetkov , Maarten Sap , Reza Shokri , Yejin Choi

Calls for transparency in AI systems are growing in number and urgency from diverse stakeholders ranging from regulators to researchers to users (with a comparative absence of companies developing AI). Notions of transparency for AI abound,…

密码学与安全 · 计算机科学 2025-02-03 Peter Hall , Olivia Mundahl , Sunoo Park

Displaying a written transcript of what a human said (i.e. producing an "automatic speech recognition transcript") is a common feature for smartphone vocal assistants: the utterance produced by a human speaker (e.g. a question) is displayed…

人机交互 · 计算机科学 2025-04-08 Damien Rudaz , Christian Licoppe

Background: Public speaking is a vital professional skill, yet it remains a source of significant anxiety for many individuals. Traditional training relies heavily on expert coaching, but recent advances in AI has led to novel types of…

人机交互 · 计算机科学 2025-07-14 Nesrine Fourati , Alisa Barkar , Marion Dragée , Liv Danthon-Lefebvre , Mathieu Chollet

As users increasingly rely on cloud-based computing services, it is important to ensure that uploaded speech data remains private. Existing solutions rely either on server-side methods or focus on hiding speaker identity. While these…

音频与语音处理 · 电气工程与系统科学 2021-10-26 Peter Wu , Paul Pu Liang , Jiatong Shi , Ruslan Salakhutdinov , Shinji Watanabe , Louis-Philippe Morency

Recently emerged intelligent assistants on smartphones and home electronics (e.g., Siri and Alexa) can be seen as novel hybrids of domain-specific task-oriented spoken dialogue systems and open-domain non-task-oriented ones. To realize such…

计算与语言 · 计算机科学 2018-07-25 Satoshi Akasaki , Nobuhiro Kaji

Smart speakers are gaining popularity. However, such devices can put the user's privacy at risk whenever hot-words are misinterpreted and voice data is recorded without the user's consent. To mitigate such risks, smart speakers provide…

人机交互 · 计算机科学 2019-11-19 Christian Tiefenau , Maximilian Häring , Eva Gerlitz , Emanuel von Zezschwitz

In contemporary society, voice-controlled devices, such as smartphones and home assistants, have become pervasive due to their advanced capabilities and functionality. The always-on nature of their microphones offers users the convenience…

密码学与安全 · 计算机科学 2023-09-27 Yuchen Liu , Apu Kapadia , Donald Williamson

The current "notice and consent" paradigm is broken: consent dialogues are often manipulative, and users cannot realistically read or understand every privacy policy. While recent LLM-based tools empower users seeking active control, many…

人机交互 · 计算机科学 2026-04-24 Vincent Freiberger

The performance of a voice anonymization system is typically measured according to its ability to hide the speaker's identity and keep the data's utility for downstream tasks. This means that the requirements the anonymization should…

音频与语音处理 · 电气工程与系统科学 2025-08-11 Sarina Meyer , Ngoc Thang Vu

The growing use of AI applications among freelance workers is reshaping trust and relationships with clients. This paper investigates how both workers and clients perceive AI use and disclosure in the freelance economy through a three-stage…

人机交互 · 计算机科学 2026-03-10 Angel Hsing-Chi Hwang , Senya Wong , Baixiao Chen , Jessica He , Hyo Jin Do

Machine learning systems are increasingly used to support public sector decision-making across a variety of sectors. Given concerns around accountability in these domains, and amidst accusations of intentional or unintentional bias, there…

计算机与社会 · 计算机科学 2018-11-06 Michael Veale

The popularity of voice-controlled smart speakers with intelligent personal assistants (IPAs) like the Amazon Echo and their increasing use as an interface for other Internet of Things (IoT) technologies in the home provides opportunities…

人机交互 · 计算机科学 2019-10-07 Mirzel Avdic , Jo Vermeulen

Advanced AI assistants combine frontier LLMs and tool access to autonomously perform complex tasks on behalf of users. While the helpfulness of such assistants can increase dramatically with access to user information including emails and…

Recent advancements in multi-turn voice interaction models have improved user-model communication. However, while closed-source models effectively retain and recall past utterances, whether open-source models share this ability remains…

声音 · 计算机科学 2025-05-26 Heeseung Kim , Che Hyun Lee , Sangkwon Park , Jiheum Yeom , Nohil Park , Sangwon Yu , Sungroh Yoon

Non-Display Smart Glasses hold the potential to support everyday activities by combining continuous environmental sensing with voice-only interaction powered by large language models (LLMs). Understanding how conversational successes and…

人机交互 · 计算机科学 2026-04-03 Xiuqi Tommy Zhu , Xiaoan Liu , Casper Harteveld , Smit Desai , Eileen McGivney