中文
相关论文

相关论文: InfoChess: A Game of Adversarial Inference and a L…

200 篇论文

We study a continuous-time stochastic Stackelberg game in which a leader seeks to accomplish a primary objective while inferring a hidden parameter of a rational follower. The follower solves an entropy-regularized tracking problem and…

最优化与控制 · 数学 2025-10-08 Ruimeng Hu , Daniel Ralston , Xu Yang , Haosheng Zhou

Learning to cooperate with friends and compete with foes is a key component of multi-agent reinforcement learning. Typically to do so, one requires access to either a model of or interaction with the other agent(s). Here we show how to…

人工智能 · 计算机科学 2019-01-03 DJ Strouse , Max Kleiman-Weiner , Josh Tenenbaum , Matt Botvinick , David Schwab

In this paper we study the problem of information sharing among rational self-interested agents as a dynamic game of asymmetric information. We assume that the agents imperfectly observe a Markov chain and they are called to decide whether…

计算机科学与博弈论 · 计算机科学 2021-03-30 Konstantinos Ntemos , George Pikramenos , Nicholas Kalouptsidis

In imperfect information games, the game state is generally not fully observable to players. Therefore, good gameplay requires policies that deal with the different information that is hidden from each player. To combat this, effective…

人工智能 · 计算机科学 2024-07-15 Timo Bertram , Johannes Fürnkranz , Martin Müller

We present a general framework for evolutionary learning to emergent unbiased state representation without any supervision. Evolutionary frameworks such as self-play converge to bad local optima in case of multi-agent reinforcement learning…

机器学习 · 统计学 2023-02-03 Shohei Ohsawa

Motivated by applications in cyber security, we develop a simple game model for describing how a learning agent's private information influences an observing agent's inference process. The model describes a situation in which one of the…

计算机科学与博弈论 · 计算机科学 2019-09-16 Erik Miehling , Roy Dong , Cédric Langbort , Tamer Başar

Incentive design deals with interaction between a principal and an agent where the former can shape the latter's utility through a policy commitment. It is well known that the principal faces an information rent when dealing with an agent…

计算机科学与博弈论 · 计算机科学 2025-09-03 Raj Kiriti Velicheti , Subhonmesh Bose , Tamer Başar

Conventional noncooperative game theory hypothesizes that the joint strategy of a set of players in a game must satisfy an "equilibrium concept". All other joint strategies are considered impossible; the only issue is what equilibrium…

适应与自组织系统 · 物理学 2007-05-23 David H. Wolpert

Reinforcement learning has enabled agents to solve challenging tasks in unknown environments. However, manually crafting reward functions can be time consuming, expensive, and error prone to human error. Competing objectives have been…

机器学习 · 计算机科学 2021-02-11 Brendon Matusch , Jimmy Ba , Danijar Hafner

Dynamic game theory is an increasingly popular tool for modeling multi-agent, e.g. human-robot, interactions. Game-theoretic models presume that each agent wishes to minimize a private cost function that depends on others' actions. These…

机器人学 · 计算机科学 2025-10-17 Cade Armstrong , Ryan Park , Xinjie Liu , Kushagra Gupta , David Fridovich-Keil

This paper studies a Stackelberg game wherein a sender (leader) attempts to shape the information of a less informed receiver (follower) who in turn takes an action that determines the payoff for both players. The sender chooses signals to…

计算机科学与博弈论 · 计算机科学 2022-10-07 Reema Deori , Ankur A. Kulkarni

Deception is a technique to mislead human or computer systems by manipulating beliefs and information. Successful deception is characterized by the information-asymmetric, dynamic, and strategic behaviors of the deceiver and the deceivee.…

密码学与安全 · 计算机科学 2018-10-02 Tao Zhang , Quanyan zhu

In many real-world settings agents engage in strategic interactions with multiple opposing agents who can employ a wide variety of strategies. The standard approach for designing agents for such settings is to compute or approximate a…

计算机科学与博弈论 · 计算机科学 2024-07-30 Sam Ganzfried , Kevin A. Wang , Max Chiswick

The goal of agents in multi-agent environments is to maximize total reward against the opposing agents that are encountered. Following a game-theoretic solution concept, such as Nash equilibrium, may obtain a strong performance in some…

计算机科学与博弈论 · 计算机科学 2026-01-05 Sam Ganzfried

The basis of the method proposed in this article is the idea that information is one of the most important factors in strategic decisions, including decisions in computer chess and other strategy games. The model proposed in this article…

人工智能 · 计算机科学 2015-03-19 Alexandru Godescu

Many multi-agent interaction scenarios can be naturally modeled as noncooperative games, where each agent's decisions depend on others' future actions. However, deploying game-theoretic planners for autonomous decision-making requires a…

机器学习 · 计算机科学 2026-01-05 Yash Jain , Xinjie Liu , Lasse Peters , David Fridovich-Keil , Ufuk Topcu

We present an approach for systematically anticipating the actions and policies employed by \emph{oblivious} environments in concurrent stochastic games, while maximizing a reward function. Our main contribution lies in the synthesis of a…

人工智能 · 计算机科学 2024-09-19 Shadi Tasdighi Kalat , Sriram Sankaranarayanan , Ashutosh Trivedi

This paper considers a non-cooperative game in which competing users sharing a frequency-selective interference channel selfishly optimize their power allocation in order to improve their achievable rates. Previously, it was shown that a…

计算机科学与博弈论 · 计算机科学 2008-11-04 Yi Su , Mihaela van der Schaar

In this paper, we present a conceptual model game to examine the dynamics of asymmetric interactions in games with imperfect information. The game involves two agents with starkly contrasting capabilities: one agent can take actions but has…

多智能体系统 · 计算机科学 2025-01-09 Fabian Farestam , Dilian Gurov

This paper presents a potential game approach for distributed cooperative selection of informative sensors, when the goal is to maximize the mutual information between the measurement variables and the quantities of interest. It is proved…

系统与控制 · 计算机科学 2014-03-05 Han-Lim Choi , Su-Jin Lee
‹ 上一页 1 2 3 10 下一页 ›