English
Related papers

Related papers: Understanding Agent Incentives using Causal Influe…

200 papers

We present a framework for analysing agent incentives using causal influence diagrams. We establish that a well-known criterion for value of information is complete. We propose a new graphical criterion for value of control, establishing…

Artificial Intelligence · Computer Science 2021-03-17 Tom Everitt , Ryan Carey , Eric Langlois , Pedro A Ortega , Shane Legg

We introduce three concepts that describe an agent's incentives: response incentives indicate which variables in the environment, such as sensitive demographic information, affect the decision under the optimal policy. Instrumental control…

Artificial Intelligence · Computer Science 2025-06-24 Ryan Carey , Eric Langlois , Chris van Merwijk , Shane Legg , Tom Everitt

Causal models of agents have been used to analyse the safety aspects of machine learning systems. But identifying agents is non-trivial -- often the causal model is just assumed by the modeler without much justification -- and modelling…

Artificial Intelligence · Computer Science 2022-08-25 Zachary Kenton , Ramana Kumar , Sebastian Farquhar , Jonathan Richens , Matt MacDermott , Tom Everitt

We focus on how individual behavior that complies with social norms interferes with performance-based incentive mechanisms in organizations with multiple distributed decision-making agents. We model social norms to emerge from interactions…

General Economics · Economics 2021-02-25 Ravshanbek Khodzhimatov , Stephan Leitner , Friederike Wall

This paper investigates causal influences between agents linked by a social graph and interacting over time. In particular, the work examines the dynamics of social learning models and distributed decision-making protocols, and derives…

Social and Information Networks · Computer Science 2026-05-19 Mert Kayaalp , Ali H. Sayed

We present a general framework for training safe agents whose naive incentives are unsafe. As an example, manipulative or deceptive behaviour can improve rewards but should be avoided. Most approaches fail here: agents maximize expected…

Artificial Intelligence · Computer Science 2022-04-22 Sebastian Farquhar , Ryan Carey , Tom Everitt

Unlike reinforcement learning (RL) agents, humans remain capable multitaskers in changing environments. In spite of only experiencing the world through their own observations and interactions, people know how to balance focusing on tasks…

Artificial Intelligence · Computer Science 2024-07-02 Rishav Bhagat , Jonathan Balloch , Zhiyu Lin , Julia Kim , Mark Riedl

Transparency and explainability are important features that responsible autonomous vehicles should possess, particularly when interacting with humans, and causal reasoning offers a strong basis to provide these qualities. However, even if…

Artificial Intelligence · Computer Science 2025-11-18 Rhys Howard , Nick Hawes , Lars Kunze

Causal reasoning has been an indispensable capability for humans and other intelligent animals to interact with the physical world. In this work, we propose to endow an artificial agent with the capability of causal reasoning for completing…

Machine Learning · Computer Science 2019-10-07 Suraj Nair , Yuke Zhu , Silvio Savarese , Li Fei-Fei

Critical sectors of human society are progressing toward the adoption of powerful artificial intelligence (AI) agents, which are trained individually on behalf of self-interested principals but deployed in a shared environment. Short of…

Multiagent Systems · Computer Science 2021-12-22 Jiachen Yang , Ethan Wang , Rakshit Trivedi , Tuo Zhao , Hongyuan Zha

We study systems of interacting reinforced stochastic processes, where agents' decisions evolve under reinforcement, network-mediated interactions, and environmental influences. In competitive environments with irreducible networks, we…

Probability · Mathematics 2025-09-18 Michele Aleandri , Paolo Dai Pra , Ida Germana Minelli

Algorithms are often used to produce decision-making rules that classify or evaluate individuals. When these individuals have incentives to be classified a certain way, they may behave strategically to influence their outcomes. We develop a…

Machine Learning · Computer Science 2019-08-02 Jon Kleinberg , Manish Raghavan

This work considers a repeated principal-agent bandit game, where the principal can only interact with her environment through the agent. The principal and the agent have misaligned objectives and the choice of action is only left to the…

We study strategic classification in binary decision-making settings where agents can modify their features in order to improve their classification outcomes. Importantly, our work considers the causal structure across different features,…

Computer Science and Game Theory · Computer Science 2025-02-11 Valia Efthymiou , Chara Podimata , Diptangshu Sen , Juba Ziani

In this paper, we consider a general distributed system with multiple agents who select and then implement actions in the system. The system has an operator with a centralized objective. The agents, on the other hand, are selfinterested and…

Computer Science and Game Theory · Computer Science 2020-01-15 Donya Ghavidel , Pratyush Chakraborty , Enrique Baeyens , Vijay Gupta , Pramod P. Khargonekar

The real world is awash with multi-agent problems that require collective action by self-interested agents, from the routing of packets across a computer network to the management of irrigation systems. Such systems have local incentives…

Multiagent Systems · Computer Science 2021-02-16 Michiel A. Bakker , Richard Everett , Laura Weidinger , Iason Gabriel , William S. Isaac , Joel Z. Leibo , Edward Hughes

Can humans get arbitrarily capable reinforcement learning (RL) agents to do their bidding? Or will sufficiently capable RL agents always find ways to bypass their intended objectives by shortcutting their reward signal? This question…

Artificial Intelligence · Computer Science 2021-03-29 Tom Everitt , Marcus Hutter , Ramana Kumar , Victoria Krakovna

In many settings, machine learning models may be used to inform decisions that impact individuals or entities who interact with the model. Such entities, or agents, may game model decisions by manipulating their inputs to the model to…

Machine Learning · Computer Science 2024-12-04 Trenton Chang , Lindsay Warrenburg , Sae-Hwan Park , Ravi B. Parikh , Maggie Makar , Jenna Wiens

To regulate a social system comprised of self-interested agents, economic incentives are often required to induce a desirable outcome. This incentive design problem naturally possesses a bilevel structure, in which a designer modifies the…

Computer Science and Game Theory · Computer Science 2022-10-14 Boyi Liu , Jiayang Li , Zhuoran Yang , Hoi-To Wai , Mingyi Hong , Yu Marco Nie , Zhaoran Wang

Neural nets are powerful function approximators, but the behavior of a given neural net, once trained, cannot be easily modified. We wish, however, for people to be able to influence neural agents' actions despite the agents never training…

Machine Learning · Computer Science 2022-02-01 Mycal Tucker , William Kuhl , Khizer Shahid , Seth Karten , Katia Sycara , Julie Shah
‹ Prev 1 2 3 10 Next ›