English
Related papers

Related papers: Responsibility in Extensive Form Games

200 papers

AI evaluation is undergoing a structural change. Large language models (LLMs) are increasingly deployed as systems that act over time through tools, environments, users, and other agents, while many evaluation practices still inherit…

Artificial Intelligence · Computer Science 2026-05-19 Keyang Xuan , Peiyang Song , Pan Lu , Pengrui Han , Wenkai Li , Zhenyu Zhang , Zexue He , Wenyue Hua , Manling Li , Jiaxuan You , Adrian Weller , Yizhong Wang , Jiaxin Pei

Counterfactual explanations are a common tool to explain artificial intelligence models. For Reinforcement Learning (RL) agents, they answer "Why not?" or "What if?" questions by illustrating what minimal change to a state is needed such…

Machine Learning · Computer Science 2023-02-27 Tobias Huber , Maximilian Demmler , Silvan Mertes , Matthew L. Olson , Elisabeth André

Being able to reason about how one's behaviour can affect the behaviour of others is a core skill required of intelligent driving agents. Despite this, the state of the art struggles to meet the need of agents to discover causal links…

Robotics · Computer Science 2024-03-07 Rhys Howard , Lars Kunze

Reciprocity is an important feature of human social interaction and underpins our cooperative nature. What is more, simple forms of reciprocity have proved remarkably resilient in matrix game social dilemmas. Most famously, the tit-for-tat…

Multiagent Systems · Computer Science 2019-03-20 Tom Eccles , Edward Hughes , János Kramár , Steven Wheelwright , Joel Z. Leibo

This paper studies lying in a novel context. Previous work has focused on situations in which people are either fully aware of the economic consequences of all available actions (e.g., die-under-cup paradigm), or they are uncertain, but…

Physics and Society · Physics 2019-08-05 Hélène Barcelo , Valerio Capraro

This paper argues that the finite horizon paradox, where game theory contradicts intuition, stems from the limitations of standard number systems in modelling the cognitive perception of infinity. To address this issue, we propose a new…

Computer Science and Game Theory · Computer Science 2025-10-10 Kiri Sakahara , Takashi Sato

Interactive constraint systems often suffer from infeasibility (no solution) due to conflicting user constraints. A common approach to recover infeasibility is to eliminate the constraints that cause the conflicts in the system. This…

Artificial Intelligence · Computer Science 2022-04-08 Sharmi Dev Gupta , Begum Genc , Barry O'Sullivan

Originating in psychology, $\textit{Theory of Mind}$ (ToM) has attracted significant attention across multiple research communities, especially logic, economics, and robotics. Most psychological work does not aim at formalizing those…

Artificial Intelligence · Computer Science 2025-12-01 Fengming Zhu , Yuxin Pan , Xiaomeng Zhu , Fangzhen Lin

It is increasingly possible for real-world agents, such as software-based agents or human institutions, to view the internal programming of other such agents that they interact with. For instance, a company can read the bylaws of another…

Computer Science and Game Theory · Computer Science 2022-08-16 Andrew Critch , Michael Dennis , Stuart Russell

Explainability is emerging as a key requirement for autonomous systems. While many works have focused on what constitutes a valid explanation, few have considered formalizing explainability as a system property. In this work, we approach…

Logic in Computer Science · Computer Science 2025-10-21 Bernd Finkbeiner , Julian Siber

Explainable models in Artificial Intelligence are often employed to ensure transparency and accountability of AI systems. The fidelity of the explanations are dependent upon the algorithms used as well as on the fidelity of the data. Many…

Machine Learning · Computer Science 2019-07-31 Muhammad Aurangzeb Ahmad , Carly Eckert , Ankur Teredesai

Achieving human-AI alignment in complex multi-agent games is crucial for creating trustworthy AI agents that enhance gameplay. We propose a method to evaluate this alignment using an interpretable task-sets framework, focusing on high-level…

Artificial Intelligence · Computer Science 2024-06-21 Sugandha Sharma , Guy Davidson , Khimya Khetarpal , Anssi Kanervisto , Udit Arora , Katja Hofmann , Ida Momennejad

Explainability and its emerging counterpart contestability have become important normative and design principles for trustworthy AI as they enable users and subjects to understand and challenge AI decisions. However, realizing these…

Computers and Society · Computer Science 2025-08-15 Timothée Schmude , Mireia Yurrita , Kars Alfrink , Thomas Le Goff , Sebastian Tschiatschek , Tiphaine Viard

Although deep reinforcement learning agents have produced impressive results in many domains, their decision making is difficult to explain to humans. To address this problem, past work has mainly focused on explaining why an action was…

Machine Learning · Computer Science 2019-10-01 Matthew L. Olson , Lawrence Neal , Fuxin Li , Weng-Keen Wong

Counterfactual reasoning requires predicting how alternative events, contrary to what actually happened, might have resulted in different outcomes. Despite being considered a necessary component of AI-complete systems, few resources have…

Computation and Language · Computer Science 2019-09-13 Lianhui Qin , Antoine Bosselut , Ari Holtzman , Chandra Bhagavatula , Elizabeth Clark , Yejin Choi

Moses & Nachum ([7]) identify conceptual flaws in Bacharach's generalization ([3]) of Aumann's seminal "agreeing to disagree" result ([1]). Essentially, Bacharach's framework requires agents' decision functions to be defined over events…

Computer Science and Game Theory · Computer Science 2013-10-28 Bassel Tarbush

In a satisficing equilibrium each agent $i$ plays one of her top $k_i$ actions in response to the actions of the other agents. Our concept unifies models of bounded rationality and yields predictions that differ from canonical solution…

Theoretical Economics · Economics 2026-04-27 Bary S. R. Pradelski , Bassel Tarbush

In recent years, explainability in machine learning has gained importance. In this context, counterfactual explanation (CE), which is an explanation method that uses examples, has attracted attention. However, it has been pointed out that…

Machine Learning · Computer Science 2025-02-04 Keita Kinjo

People often interact repeatedly: with relatives, through file sharing, in politics, etc. Many such interactions are reciprocal: reacting to the actions of the other. In order to facilitate decisions regarding reciprocal interactions, we…

Computer Science and Game Theory · Computer Science 2016-03-01 Gleb Polevoy , Mathijs de Weerdt , Catholijn Jonker

Conventional AI evaluation approaches concentrated within the AI stack exhibit systemic limitations for exploring, navigating and resolving the human and societal factors that play out in real world deployment such as in education, finance,…

‹ Prev 1 8 9 10 Next ›