中文
相关论文

相关论文: Integrating Sequential Hypothesis Testing into Adv…

200 篇论文

Agents rarely act in isolation -- their behavioral history, in particular, is public to others. We seek a non-asymptotic understanding of how a leader agent should shape this history to its maximal advantage, knowing that follower agent(s)…

计算机科学与博弈论 · 计算机科学 2019-05-29 Vidya Muthukumar , Anant Sahai

Security challenges accompany the efficiency. The pervasive integration of information and communications technologies (ICTs) makes cyber-physical systems vulnerable to targeted attacks that are deceptive, persistent, adaptive and…

计算机科学与博弈论 · 计算机科学 2018-09-11 Linan Huang , Quanyan Zhu

Deception is helpful for agents masking their intentions from an observer. We consider a team of agents deceiving their supervisor. The supervisor defines nominal behavior for the agents via reference policies, but the agents share an…

最优化与控制 · 数学 2024-10-28 Caleb Probine , Mustafa O. Karabag , Ufuk Topcu

Stackelberg equilibrium is a solution concept that describes optimal strategies to commit: Player 1 (the leader) first commits to a strategy that is publicly announced, then Player 2 (the follower) plays a best response to the leader's…

计算机科学与博弈论 · 计算机科学 2021-11-04 Aditya Aradhye , Branislav Bošanský , Michael Hlaváček

In a Stackelberg game, a leader commits to a randomized strategy, and a follower chooses their best strategy in response. We consider an extension of a standard Stackelberg game, called a discrete-time dynamic Stackelberg game, that has an…

计算机科学与博弈论 · 计算机科学 2022-02-11 Niklas Lauffer , Mahsa Ghasemi , Abolfazl Hashemi , Yagiz Savas , Ufuk Topcu

Consider a 4-player version of Matching Pennies where a team of three players competes against the Devil. Each player simultaneously says "Heads" or "Tails". The team wins if all four choices match; otherwise the Devil wins. If all team…

计算机科学与博弈论 · 计算机科学 2026-05-14 Léonard Brice , Thomas A. Henzinger , K. S. Thejaswini

This paper is concerned with a linear-quadratic partially observed Stackelberg stochastic differential game with correlated state and observation noises, where the diffusion coefficient does not contain the control variable and the control…

最优化与控制 · 数学 2021-05-25 Yueyang Zheng , Jingtao Shi

Effectively predicting intent and behavior requires inferring leadership in multi-agent interactions. Dynamic games provide an expressive theoretical framework for modeling these interactions. Employing this framework, we propose a novel…

多智能体系统 · 计算机科学 2024-04-10 Hamzah Khan , David Fridovich-Keil

In the problem of active sequential hypothesis testing (ASHT), a learner seeks to identify the true hypothesis from among a known set of hypotheses. The learner is given a set of actions and knows the random distribution of the outcome of…

机器学习 · 计算机科学 2021-10-08 Kyra Gan , Su Jia , Andrew Li

Automated verification techniques for stochastic games allow formal reasoning about systems that feature competitive or collaborative behaviour among rational agents in uncertain or probabilistic settings. Existing tools and techniques…

计算机科学中的逻辑 · 计算机科学 2020-09-01 Marta Kwiatkowska , Gethin Norman , David Parker , Gabriel Santos

We study a setting where a group of agents, each receiving partially informative private observations, seek to collaboratively learn the true state (among a set of hypotheses) that explains their joint observation profiles over time. To…

系统与控制 · 计算机科学 2019-03-15 Aritra Mitra , John A. Richards , Shreyas Sundaram

In dynamic noncooperative games, each player makes conjectures about other players' reactions before choosing a strategy. However, resulting equilibria may be multiple and do not always lead to desirable outcomes. These issues are typically…

计算机科学与博弈论 · 计算机科学 2025-11-24 Francesco Morri , Hélène Le Cadre , David Salas , Didier Aussel

This paper is concerned with an overlapping information linear-quadratic (LQ) Stackelberg stochastic differential game with two leaders and two followers, where the diffusion terms of the state equation contain both the control and state…

最优化与控制 · 数学 2024-01-17 Yu Si , Jingtao Shi

This work addresses competitive resource allocation in a sequential setting, where two players allocate resources across objects or locations of shared interest. Departing from the simultaneous Colonel Blotto game, our framework introduces…

计算机科学与博弈论 · 计算机科学 2025-04-24 Omkar Thakoor , Rajgopal Kannan , Victor Prasanna

The emerging social network platforms enable users to share their own opinions, as well as to exchange opinions with others. However, adversarial network perturbation, where malicious users intentionally spread their extreme opinions,…

计算机与社会 · 计算机科学 2023-04-26 Yuejiang Li , Zhanjiang Chen , H. Vicky Zhao

Sequential reasoning is a complex human ability, with extensive previous research focusing on gaming AI in a single continuous game, round-based decision makings extending to a sequence of games remain less explored. Counter-Strike: Global…

人工智能 · 计算机科学 2020-08-13 Yilei Zeng , Deren Lei , Beichen Li , Gangrong Jiang , Emilio Ferrara , Michael Zyda

Adversarial deep learning is to train robust DNNs against adversarial attacks, which is one of the major research focuses of deep learning. Game theory has been used to answer some of the basic questions about adversarial deep learning such…

机器学习 · 计算机科学 2022-07-19 Xiao-Shan Gao , Shuang Liu , Lijia Yu

We apply a Bayesian agent-based framework inspired by QBism to iterations of two quantum games, the CHSH game and the quantum prisoners' dilemma. In each two-player game, players hold beliefs about an amount of shared entanglement and about…

量子物理 · 物理学 2026-04-28 John B. DeBrota , Peter J. Love

This work studies Stackelberg network interdiction games -- an important class of games in which a defender first allocates (randomized) defense resources to a set of critical nodes on a graph while an adversary chooses its path to attack…

最优化与控制 · 数学 2023-01-31 Tien Mai , Avinandan Bose , Arunesh Sinha , Thanh H. Nguyen

Model-based reinforcement learning (MBRL) has recently gained immense interest due to its potential for sample efficiency and ability to incorporate off-policy data. However, designing stable and efficient MBRL algorithms using rich…

机器学习 · 计算机科学 2021-03-12 Aravind Rajeswaran , Igor Mordatch , Vikash Kumar