中文
相关论文

相关论文: Mitigating Negative Side Effects via Environment S…

200 篇论文

This paper contributes a preliminary report on the advantages and disadvantages of incorporating simultaneous human control and feedback signals in the training of a reinforcement learning robotic agent. While robotic human-machine…

人机交互 · 计算机科学 2016-06-23 Kory W. Mathewson , Patrick M. Pilarski

We consider the problem of creating assistants that can help agents solve new sequential decision problems, assuming the agent is not able to specify the reward function explicitly to the assistant. Instead of acting in place of the agent…

机器学习 · 计算机科学 2022-12-01 Sebastiaan De Peuter , Samuel Kaski

Neuroevolution (NE) has recently proven a competitive alternative to learning by gradient descent in reinforcement learning tasks. However, the majority of NE methods and associated simulation environments differ crucially from biological…

神经与进化计算 · 计算机科学 2023-08-07 Gautier Hamon , Eleni Nisioti , Clément Moulin-Frier

We present a general framework for training safe agents whose naive incentives are unsafe. As an example, manipulative or deceptive behaviour can improve rewards but should be avoided. Most approaches fail here: agents maximize expected…

人工智能 · 计算机科学 2022-04-22 Sebastian Farquhar , Ryan Carey , Tom Everitt

Agents often exert influence when interacting with humans and non-human agents. However, the ethical status of such influence is often unclear. In this paper, we present the SHAPE framework, which lists reasons why influence may be…

多智能体系统 · 计算机科学 2023-11-08 Elfia Bezou-Vrakatseli , Benedikt Brückner , Luke Thorburn

Human behaviors are regularized by a variety of norms or regulations, either to maintain orders or to enhance social welfare. If artificially intelligent (AI) agents make decisions on behalf of human beings, we would hope they can also…

计算机科学与博弈论 · 计算机科学 2019-10-28 Fan-Yun Sun , Yen-Yu Chang , Yueh-Hua Wu , Shou-De Lin

Both entropy-minimizing and entropy-maximizing (curiosity) objectives for unsupervised reinforcement learning (RL) have been shown to be effective in different environments, depending on the environment's level of natural entropy. However,…

机器学习 · 计算机科学 2024-08-19 Adriana Hugessen , Roger Creus Castanyer , Faisal Mohamed , Glen Berseth

In many situations, communication between agents is a critical component of cooperative multi-agent systems, however, it can be difficult to learn or evolve. In this paper, we investigate a simple way in which the emergence of communication…

多智能体系统 · 计算机科学 2024-05-28 Dylan Cope , Peter McBurney

Balancing user agency and system automation is essential for effective human-AI interactions. Fully automated systems can deliver efficiency but risk undermining usability and user autonomy, while purely manual tools are often inefficient…

人机交互 · 计算机科学 2025-02-20 Thomas Langerak

In cooperative multiagent planning, it can often be beneficial for an agent to make commitments about aspects of its behavior to others, allowing them in turn to plan their own behaviors without taking the agent's detailed behavior into…

人工智能 · 计算机科学 2017-03-16 Qi Zhang , Satinder Singh , Edmund Durfee

Imitation Learning techniques enable programming the behavior of agents through demonstrations rather than manual engineering. However, they are limited by the quality of available demonstration data. Interactive Imitation Learning…

机器人学 · 计算机科学 2022-03-09 Snehal Jauhri , Carlos Celemin , Jens Kober

Movement is how people interact with and affect their environment. For realistic character animation, it is necessary to synthesize such interactions between virtual characters and their surroundings. Despite recent progress in character…

图形学 · 计算机科学 2023-02-03 Mohamed Hassan , Yunrong Guo , Tingwu Wang , Michael Black , Sanja Fidler , Xue Bin Peng

Obtaining reliable feedback from the environment is a fundamental capability for intelligent agents to evaluate the correctness of their actions and to accumulate reusable knowledge. However, most existing approaches rely on predefined…

人工智能 · 计算机科学 2026-01-09 Hong Su

The potential for negative impacts of AI has rapidly become more pervasive around the world, and this has intensified a need for responsible AI governance. While many regulatory bodies endorse risk-based approaches and a multitude of risk…

计算机与社会 · 计算机科学 2025-02-24 Julia Barnett , Kimon Kieslich , Natali Helberger , Nicholas Diakopoulos

To widen their accessibility and increase their utility, intelligent agents must be able to learn complex behaviors as specified by (non-expert) human users. Moreover, they will need to learn these behaviors within a reasonable amount of…

机器学习 · 计算机科学 2019-02-13 Dilip Arumugam , Jun Ki Lee , Sophie Saskin , Michael L. Littman

To promote cooperation and strengthen the individual impact on the collective outcome in social dilemmas, we propose the Environmental-impact Multi-Agent Reinforcement Learning (EMuReL) method where each agent estimates the "environmental…

人工智能 · 计算机科学 2023-11-09 Farinaz Alamiyan-Harandi , Pouria Ramazi

Constructing personalized and anthropomorphic agents holds significant importance in the simulation of social networks. However, there are still two key problems in existing works: the agent possesses world knowledge that does not belong to…

计算与语言 · 计算机科学 2024-04-03 Junkai Zhou , Liang Pang , Ya Jing , Jia Gu , Huawei Shen , Xueqi Cheng

Seamlessly interacting with humans or robots is hard because these agents are non-stationary. They update their policy in response to the ego agent's behavior, and the ego agent must anticipate these changes to co-adapt. Inspired by humans,…

机器人学 · 计算机科学 2020-11-16 Annie Xie , Dylan P. Losey , Ryan Tolsma , Chelsea Finn , Dorsa Sadigh

This work explores learning agent-agnostic synthetic environments (SEs) for Reinforcement Learning. SEs act as a proxy for target environments and allow agents to be trained more efficiently than when directly trained on the target…

机器学习 · 计算机科学 2021-02-09 Fabio Ferreira , Thomas Nierhoff , Frank Hutter

The progressive advent of artificial intelligence machines may represent both an opportunity or a threat. In order to have an idea of what is coming we propose a model that simulate a Human-AI ecosystem. In particular we consider systems…

人机交互 · 计算机科学 2022-10-12 Pierluigi Contucci , János Kertész , Godwin Osabutey