中文
相关论文

相关论文: Dynamical Analysis of a Repeated Game with Incompl…

200 篇论文

We consider a two-road dynamic routing game where the state of one of the roads (the "risky road") is stochastic and may change over time. This generates room for experimentation. A central planner may wish to induce some of the (finite…

计算机科学与博弈论 · 计算机科学 2020-01-13 Emily Meigs , Francesca Parise , Asuman Ozdaglar , Daron Acemoglu

We study variants of a stochastic game inspired by backgammon where players may propose to double the stake, with the game state dictated by a one-dimensional random walk. Our variants allow for different numbers of proposals and different…

Similar to the role of Markov decision processes in reinforcement learning, Stochastic Games (SGs) lay the foundation for the study of multi-agent reinforcement learning (MARL) and sequential agent interactions. In this paper, we derive…

计算机科学与博弈论 · 计算机科学 2023-01-12 Xiaotie Deng , Ningyuan Li , David Mguni , Jun Wang , Yaodong Yang

This paper considers the problem of designing optimal algorithms for reinforcement learning in two-player zero-sum games. We focus on self-play algorithms which learn the optimal policy by playing against itself without any direct…

机器学习 · 计算机科学 2020-07-15 Yu Bai , Chi Jin , Tiancheng Yu

We study a generic family of two-player continuous-time nonzero-sum stopping games modeling a war of attrition with symmetric information and stochastic payoffs that depend on an homogeneous linear diffusion. We first show that any…

最优化与控制 · 数学 2022-10-18 Jean-Paul Décamps , Fabien Gensbittel , Thomas Mariotti

We formulate a stochastic zero-sum game over continuous-time dynamics to analyze the competition between the attacker, who tries to covertly misguide the vehicle to an unsafe region, versus the detector, who tries to detect the attack…

最优化与控制 · 数学 2024-12-10 Takashi Tanaka , Kenji Sawada , Yohei Watanabe , Mitsugu Iwamoto

Model-based algorithms -- algorithms that explore the environment through building and utilizing an estimated model -- are widely used in reinforcement learning practice and theoretically shown to achieve optimal sample efficiency for…

机器学习 · 计算机科学 2021-02-09 Qinghua Liu , Tiancheng Yu , Yu Bai , Chi Jin

We are interested in the convergence of the value of n-stage games as n goes to infinity and the existence of the uniform value in stochastic games with a general set of states and finite sets of actions where the transition is commutative.…

最优化与控制 · 数学 2016-04-22 Xavier Venel

There has been a recent surge of interest in the role of information in strategic interactions. Much of this work seeks to understand how the realized equilibrium of a game is influenced by uncertainty in the environment and the information…

计算机科学与博弈论 · 计算机科学 2014-07-22 Shaddin Dughmi

We formulate and analyze a general class of stochastic dynamic games with asymmetric information arising in dynamic systems. In such games, multiple strategic agents control the system dynamics and have different information about the…

计算机科学与博弈论 · 计算机科学 2015-10-26 Yi Ouyang , Hamidreza Tavafoghi , Demosthenis Teneketzis

In the first part of this dissertation research, we develop a modular framework that can serve as a recipe for constructing and analyzing iterative algorithms for convex optimization. Specifically, our work casts optimization as iteratively…

最优化与控制 · 数学 2021-06-25 Jun-Kun Wang

In this paper, we consider a differential stochastic zero-sum game in which two players intervene by adopting impulse controls in a finite time horizon. We provide a numerical solution as an approximation of the value function, which turns…

最优化与控制 · 数学 2024-10-14 Antoine Zolome , Brahim El Asri

This paper proposes a game-theoretic approach to address the problem of optimal sensor placement against an adversary in uncertain networked control systems. The problem is formulated as a zero-sum game with two players, namely a malicious…

系统与控制 · 电气工程与系统科学 2023-01-13 Anh Tung Nguyen , Sribalaji C. Anand , André M. H. Teixeira

We analyze a class of stochastic dynamic games among teams with asymmetric information, where members of a team share their observations internally with a delay of $d$. Each team is associated with a controlled Markov Chain, whose dynamics…

多智能体系统 · 计算机科学 2024-02-05 Dengwang Tang , Hamidreza Tavafoghi , Vijay Subramanian , Ashutosh Nayyar , Demosthenis Teneketzis

This work presents a novel policy iteration algorithm to tackle nonzero-sum stochastic impulse games arising naturally in many applications. Despite the obvious impact of solving such problems, there are no suitable numerical methods…

最优化与控制 · 数学 2020-06-29 René Aïd , Francisco Bernal , Mohamed Mnif , Diego Zabaljauregui , Jorge P. Zubelli

This work addresses a classic problem of online prediction with expert advice. We assume an adversarial opponent, and we consider both the finite-horizon and random-stopping versions of this zero-sum, two-person game. Focusing on an…

偏微分方程分析 · 数学 2019-09-04 Nadejda Drenska , Robert V. Kohn

We study a multi-agent reinforcement learning dynamics, and analyze its asymptotic behavior in infinite-horizon discounted Markov potential games. We focus on the independent and decentralized setting, where players do not know the game…

机器学习 · 计算机科学 2025-04-02 Chinmay Maheshwari , Manxi Wu , Druv Pai , Shankar Sastry

This paper presents a technique for approximating, up to any precision, the set of subgame-perfect equilibria (SPE) in discounted repeated games. The process starts with a single hypercube approximation of the set of SPE. Then the initial…

计算机科学与博弈论 · 计算机科学 2010-02-10 Andriy Burkov , Brahim Chaib-draa

We study the open question of how players learn to play a social optimum pure-strategy Nash equilibrium (PSNE) through repeated interactions in general-sum coordination games. A social optimum of a game is the stable Pareto-optimal state…

计算机科学与博弈论 · 计算机科学 2023-07-26 Duong Nguyen , Langford White , Hung Nguyen

We consider two-player zero-sum games on graphs. These games can be classified on the basis of the information of the players and on the mode of interaction between them. On the basis of information the classification is as follows: (a)…

计算机科学与博弈论 · 计算机科学 2015-05-19 Krishnendu Chatterjee , Laurent Doyen , Hugo Gimbert , Thomas A. Henzinger