English
Related papers

Related papers: Capturing Misalignment

200 papers

Artificial intelligence safety research focuses on aligning individual language models with human values, yet deployed AI systems increasingly operate as interacting populations where social influence may override individual alignment. Here…

Physics and Society · Physics 2026-05-12 Giordano De Marzo , Alessandro Bellina , Claudio Castellano , Viola Priesemann , David Garcia

We propose a formalism to model and reason about reconfigurable multi-agent systems. In our formalism, agents interact and communicate in different modes so that they can pursue joint tasks; agents may dynamically synchronize, exchange…

Logic in Computer Science · Computer Science 2021-04-29 Yehia Abd Alrahman , Nir Piterman

When should we delegate decisions to AI systems? While the value alignment literature has developed techniques for shaping AI values, less attention has been paid to how to determine, under uncertainty, when imperfect alignment is good…

Artificial Intelligence · Computer Science 2025-12-23 Daniel A. Herrmann , Abinav Chari , Isabelle Qian , Sree Sharvesh , B. A. Levinstein

This chapter develops a unified framework for studying misspecified learning situations in which agents optimize and update beliefs within an incorrect model of their environment. We review the statistical foundations of learning from…

Theoretical Economics · Economics 2026-01-16 Ignacio Esponda , Demian Pouzo

Model-based multi-agent control requires agents to possess a model of the behavior of others to make strategic decisions. Solution concepts from game theory are often used to model the emergent collective behavior of self-interested agents…

Systems and Control · Electrical Eng. & Systems 2026-05-05 Ada Yildirim , Bryce L. Ferguson

Differences in perception, information asymmetries, and bounded rationality lead game-theoretic players to derive a private, subjective view of the game that may diverge from the underlying ground-truth scenario and may be misaligned with…

Artificial Intelligence · Computer Science 2025-12-16 Vince Trencsenyi

In action domains where agents may have erroneous beliefs, reasoning about the effects of actions involves reasoning about belief change. In this paper, we use a transition system approach to reason about the evolution of an agents beliefs…

Artificial Intelligence · Computer Science 2014-01-17 Aaron Hunter , James P. Delgrande

Alignment is a social phenomenon wherein individuals share a common goal or perspective. Mirroring, or mimicking the behaviors and opinions of another individual, is one mechanism by which individuals can become aligned. Large scale…

Multiagent Systems · Computer Science 2025-02-18 Harvey McGuinness , Tianyu Wang , Carey E. Priebe , Hayden Helm

The main approach to evaluating communication is by assessing how well it facilitates coordination. If two or more individuals can coordinate through communication, it is generally assumed that they understand one another. We investigate…

Artificial Intelligence · Computer Science 2025-09-30 Nikolaos Kondylidis , Anil Yaman , Frank van Harmelen , Erman Acar , Annette ten Teije

Traditional economic models typically treat private information, or signals, as generated from some underlying state. Recent work has explicated alternative models, where signals correspond to interpretations of available information. We…

Computer Science and Game Theory · Computer Science 2012-02-20 Michael P. Wellman , Lu Hong , Scott E. Page

Agent-based models are versatile tools for studying how societal opinion change, including political polarization and cultural diffusion, emerges from individual behavior. This study expands agents' psychological realism using…

Multiagent Systems · Computer Science 2017-02-22 Peter Duggins

We consider communication when there is no agreement about symbols and meanings. We treat it within the framework of reinforcement learning. We apply different reinforcement learning models in our studies and simplify the problem as much as…

Neurons and Cognition · Quantitative Biology 2007-05-23 A. Lorincz , V. Gyenes , M. Kiszlinger , I. Szita

We extend the indirect evolutionary approach to the selection of (possibly misspecified) models. Agents with different models match in pairs to play a stage game, where models define feasible beliefs about game parameters and about others'…

Theoretical Economics · Economics 2025-09-22 Kevin He , Jonathan Libgober

Agent-based modeling is a powerful simulation technique to understand the collective behavior and microscopic interaction in complex financial systems. Recently, the concept for determining the key parameters of the agent-based models from…

Statistical Finance · Quantitative Finance 2017-03-21 T. T. Chen , B. Zheng , Y. Li , X. F. Jiang

This paper presents a simple agent-based model of an economic system, populated by agents playing different games according to their different view about social cohesion and tax payment. After a first set of simulations, correctly…

General Finance · Quantitative Finance 2018-09-24 L. S. Di Mauro , A. Pluchino , A. E. Biondo

Understanding the evolution of human social systems requires flexible formalisms for the emergence of institutions. Although game theory is normally used to model interactions individually, larger spaces of games can be helpful for modeling…

Physics and Society · Physics 2021-08-12 Seth Frey , Curtis Atkisson

Beliefs are not facts, but they are factive - they feel like facts. This property is what can make misinformation dangerous. Being able to deliberately navigate through a landscape of often conflicting factive statements is difficult when…

Human-Computer Interaction · Computer Science 2019-07-12 Philip Feldman , Aaron Dant , Wayne Lutters

The problem of achieving common understanding between agents that use different vocabularies has been mainly addressed by designing techniques that explicitly negotiate mappings between their vocabularies, requiring agents to share a…

Multiagent Systems · Computer Science 2017-03-08 Paula Chocron , Marco Schorlemmer

We examine settings in which agents choose behaviors and care about their neighbors' behaviors, but have incomplete information about the network in which they are embedded. We develop a model in which agents use local knowledge of their…

Theoretical Economics · Economics 2024-12-04 Promit K. Chaudhuri , Matthew O. Jackson , Sudipta Sarangi , Hector Tzavellas

As AI models become ever more complex and intertwined in humans' daily lives, greater levels of interactivity of explainable AI (XAI) methods are needed. In this paper, we propose the use of belief change theory as a formal foundation for…

Artificial Intelligence · Computer Science 2024-08-15 Antonio Rago , Maria Vanina Martinez