中文
相关论文

相关论文: Towards Machines that Trust: AI Agents Learn to Tr…

200 篇论文

Scientists and philosophers have debated whether humans can trust advanced artificial intelligence (AI) agents to respect humanity's best interests. Yet what about the reverse? Will advanced AI agents trust humans? Gauging an AI agent's…

人工智能 · 计算机科学 2022-12-29 Tim Johnson , Nick Obradovich

Handling trust is one of the core requirements for facilitating effective interaction between the human and the AI agent. Thus, any decision-making framework designed to work with humans must possess the ability to estimate and leverage…

人工智能 · 计算机科学 2023-01-31 Zahra Zahedi , Sarath Sreedharan , Subbarao Kambhampati

This chapter explores the symbiotic relationship between Artificial Intelligence (AI) and trust in networked systems, focusing on how these two elements reinforce each other in strategic cybersecurity contexts. AI's capabilities in data…

人工智能 · 计算机科学 2024-11-21 Yunfei Ge , Quanyan Zhu

Practical uses of Artificial Intelligence (AI) in the real world have demonstrated the importance of embedding moral choices into intelligent agents. They have also highlighted that defining top-down ethical constraints on AI according to…

多智能体系统 · 计算机科学 2023-08-31 Elizaveta Tennant , Stephen Hailes , Mirco Musolesi

Trust is a central component of the interaction between people and AI, in that 'incorrect' levels of trust may cause misuse, abuse or disuse of the technology. But what, precisely, is the nature of trust in AI? What are the prerequisites…

人工智能 · 计算机科学 2021-01-21 Alon Jacovi , Ana Marasović , Tim Miller , Yoav Goldberg

As artificial intelligence (AI) assistants become more widely adopted in safety-critical domains, it becomes important to develop safeguards against potential failures or adversarial attacks. A key prerequisite to developing these…

人机交互 · 计算机科学 2025-04-04 Abed Kareem Musaffar , Anand Gokhale , Sirui Zeng , Rasta Tadayon , Xifeng Yan , Ambuj Singh , Francesco Bullo

Behavioral experiments on the trust game have shown that trust and trustworthiness are universal among human beings, contradicting the prediction by assuming \emph{Homo economicus} in orthodox Economics. This means some mechanism must be at…

种群与进化 · 定量生物学 2024-12-20 Guozhong Zheng , Jiqiang Zhang , Jing Zhang , Weiran Cai , Li Chen

In human society, trust is an essential component of social attitude that helps build and maintain long-term, healthy relationships which creates a strong foundation for cooperation, enabling individuals to work together effectively and…

人机交互 · 计算机科学 2025-07-30 Anushka Debnath , Stephen Cranefield , Emiliano Lorini , Bastin Tony Roy Savarimuthu

How do people build up trust with artificial agents? Here, we study a key component of interpersonal trust: people's ability to evaluate the competence of another agent across repeated interactions. Prior work has largely focused on…

人机交互 · 计算机科学 2022-05-25 Erik Brockbank , Haoliang Wang , Justin Yang , Suvir Mirchandani , Erdem Bıyık , Dorsa Sadigh , Judith E. Fan

An ambitious goal for machine learning is to create agents that behave ethically: The capacity to abide by human moral norms would greatly expand the context in which autonomous agents could be practically and safely deployed, e.g. fully…

人工智能 · 计算机科学 2021-07-21 Adrien Ecoffet , Joel Lehman

To achieve desirable performance, current AI systems often require huge amounts of training data. This is especially problematic in domains where collecting data is both expensive and time-consuming, e.g., where AI systems require having…

人工智能 · 计算机科学 2022-10-11 Ardavan S. Nobandegani , Thomas R. Shultz , Irina Rish

Artificial reinforcement learning (RL) is a widely used technique in artificial intelligence that provides a general method for training agents to perform a wide variety of behaviours. RL as used in computer science has striking parallels…

人工智能 · 计算机科学 2014-10-31 Brian Tomasik

How can we build AI systems that can learn any set of individual human values both quickly and safely, avoiding causing harm or violating societal standards for acceptable behavior during the learning process? We explore the effects of…

人工智能 · 计算机科学 2024-11-11 Andrea Wynn , Ilia Sucholutsky , Thomas L. Griffiths

We present an overview of the literature on trust in AI and AI trustworthiness and argue for the need to distinguish these concepts more clearly and to gather more empirically evidence on what contributes to people s trusting behaviours. We…

人工智能 · 计算机科学 2023-09-20 Andreas Duenser , David M. Douglas

One critical aspect of building human-centered, trustworthy artificial intelligence (AI) systems is maintaining calibrated trust: appropriate reliance on AI systems outperforms both overtrust (e.g., automation bias) and undertrust (e.g.,…

计算与语言 · 计算机科学 2026-02-03 Siyu Yan , Lusha Zhu , Jian-Qiao Zhu

Artificial intelligence systems increasingly involve continual learning to enable flexibility in general situations that are not encountered during system training. Human interaction with autonomous systems is broadly studied, but research…

Human trust research uncovered important catalysts for trust building between interaction partners such as appearance or cognitive factors. The introduction of robots into social interactions calls for a reevaluation of these findings and…

机器人学 · 计算机科学 2023-11-15 Anna L. Lange , Murat Kirtay , Verena V. Hafner

This study investigated whether human trust in a social robot with anthropomorphic physicality is similar to that in an AI agent or in a human in order to clarify how anthropomorphic physicality influences human trust in an agent. We…

人机交互 · 计算机科学 2023-04-05 Akihiro Maehigashi , Takahiro Tsumura , Seiji Yamada

There is general agreement that fostering trust and cooperation within the AI development ecosystem is essential to promote the adoption of trustworthy AI systems. By embedding Large Language Model (LLM) agents within an evolutionary…

This study explores the dynamics of trust in artificial intelligence (AI) agents, particularly large language models (LLMs), by introducing the concept of "deferred trust", a cognitive mechanism where distrust in human agents redirects…

人机交互 · 计算机科学 2025-11-24 Johan Sebastián Galindez-Acosta , Juan José Giraldo-Huertas
‹ 上一页 1 2 3 10 下一页 ›