中文
相关论文

相关论文: The Stackelberg Equilibrium for One-sided Zero-sum…

200 篇论文

A wide variety of goals could cause an AI to disable its off switch because "you can't fetch the coffee if you're dead" (Russell 2019). Prior theoretical work on this shutdown problem assumes that humans know everything that AIs do. In…

计算机科学与博弈论 · 计算机科学 2024-12-10 Andrew Garber , Rohan Subramani , Linus Luu , Mark Bedaywi , Stuart Russell , Scott Emmons

Establishing the existence of Nash equilibria for partially observed stochastic dynamic games is known to be quite challenging, with the difficulties stemming from the noisy nature of the measurements available to individual players…

系统与控制 · 计算机科学 2018-06-06 Naci Saldi , Tamer Basar , Maxim Raginsky

This paper is concerned with a linear-quadratic (LQ) leader-follower differential game with mixed deterministic and stochastic controls. In the game, the follower is a random controller which means that the follower can choose adapted…

最优化与控制 · 数学 2025-09-26 Jingtao Shi , Guangchen Wang

Stackelberg games have been widely used to model interactive decision-making problems in a variety of domains such as energy systems, transportation, cybersecurity, and human-robot interaction. However, existing algorithms for solving…

最优化与控制 · 数学 2023-03-14 Yansong Li , Shuo Han

In this technical note, we consider the linear-quadratic time-inconsistent mean-field type leader-follower Stackelberg differential game with an adapted open-loop information structure. The objective functionals of the leader and the…

最优化与控制 · 数学 2019-11-12 Jun Moon , Hyun Jong Yang

In this work, we use a Stackelberg infinite discrete-time dynamic game model to study the optimal supply schedule and the optimal demand response under a market-driven dynamic price. A two-layer optimization framework is established. At the…

系统与控制 · 电气工程与系统科学 2019-11-19 Yunhan Huang

The paper presents a new method for approximating Strong Stackelberg Equilibrium in general-sum sequential games with imperfect information and perfect recall. The proposed approach is generic as it does not rely on any specific properties…

计算机科学与博弈论 · 计算机科学 2022-08-16 Jan Karwowski , Jacek Mańdziuk

State-of-the-art methods for solving 2-player zero-sum imperfect information games rely on linear programming or regret minimization, though not on dynamic programming (DP) or heuristic search (HS), while the latter are often at the core of…

人工智能 · 计算机科学 2022-10-27 Aurélien Delage , Olivier Buffet , Jilles S. Dibangoye , Abdallah Saffidine

This paper investigates strategic interactions within a three party deception security game involving a defender, an insider, and external attackers. We propose a robust deception mechanism where the leader manipulates game parameters…

计算机科学与博弈论 · 计算机科学 2026-04-06 Xiaoyu Xin , Gehui Xu , Yiguang Hong

Adversarial decision-making in partially observable multi-agent systems requires sophisticated strategies for both deception and counter-deception. This paper presents a sequential hypothesis testing (SHT)-driven framework that captures the…

最优化与控制 · 数学 2026-04-14 Haosheng Zhou , Daniel Ralston , Xu Yang , Ruimeng Hu

This paper studies the problem of multi-step manipulative attacks in Stackelberg security games, in which a clever attacker attempts to orchestrate its attacks over multiple time steps to mislead the defender's learning of the attacker's…

人工智能 · 计算机科学 2022-03-02 Thanh H. Nguyen , Arunesh Sinha

This paper studies a stochastic game theoretic approach to security and intrusion detection in communication and computer networks. Specifically, an Attacker and a Defender take part in a two-player game over a network of nodes whose…

密码学与安全 · 计算机科学 2010-03-15 Kien C. Nguyen , Tansu Alpcan , Tamer Basar

In 1996, Mallozzi and Morgan [33] proposed a new model for Stackelberg games which we refer here to as the Bayesian approach. The leader has only partial information about how followers select their reaction among possibly multiple optimal…

最优化与控制 · 数学 2023-05-12 David Salas , Anton Svensson

Min-max optimization problems (i.e., min-max games) have attracted a great deal of attention recently as their applicability to a wide range of machine learning problems has become evident. In this paper, we study min-max games with…

计算机科学与博弈论 · 计算机科学 2022-08-23 Denizalp Goktas , Amy Greenwald

Optimizing strategic decisions (a.k.a. computing equilibrium) is key to the success of many non-cooperative multi-agent applications. However, in many real-world situations, we may face the exact opposite of this game-theoretic problem --…

计算机科学与博弈论 · 计算机科学 2022-10-05 Jibang Wu , Weiran Shen , Fei Fang , Haifeng Xu

In a Stackelberg game, a leader commits to a randomized strategy, and a follower chooses their best strategy in response. We consider an extension of a standard Stackelberg game, called a discrete-time dynamic Stackelberg game, that has an…

计算机科学与博弈论 · 计算机科学 2022-02-11 Niklas Lauffer , Mahsa Ghasemi , Abolfazl Hashemi , Yagiz Savas , Ufuk Topcu

This paper introduces a differentially private (DP) mechanism to protect the information exchanged during the coordination of sequential and interdependent markets. This coordination represents a classic Stackelberg game and relies on the…

最优化与控制 · 数学 2020-04-20 Ferdinando Fioretto , Lesia Mitridati , Pascal Van Hentenryck

Inverse game theory is utilized to infer the cost functions of all players based on game outcomes. However, existing inverse game theory methods do not consider the learner as an active participant in the game, which could significantly…

计算机科学与博弈论 · 计算机科学 2025-10-20 Jianguo Chen , Jinlong Lei , Biqiang Mu , Yiguang Hong , Hongsheng Qi

Interdicting a criminal with limited police resources is a challenging task as the criminal changes location over time. The size of the large transportation network further adds to the difficulty of this scenario. To tackle this issue, we…

人工智能 · 计算机科学 2026-04-08 Sukanya Samanta , Kei Kimura , Makoto Yokoo , Palash Dey

We compute equilibrium strategies in multi-stage games with continuous signal and action spaces as they are widely used in the management sciences and economics. Examples include sequential sales via auctions, multi-stage elimination…

计算机科学与博弈论 · 计算机科学 2024-07-23 Fabian R. Pieroth , Nils Kohring , Martin Bichler