中文
相关论文

相关论文: Learning Stationary Nash Equilibrium Policies in $…

200 篇论文

We study a subclass of $n$-player stochastic games, namely, stochastic games with independent chains and unknown transition matrices. In this class of games, players control their own internal Markov chains whose transitions do not depend…

计算机科学与博弈论 · 计算机科学 2023-12-05 Tiancheng Qin , S. Rasoul Etesami

We study the performance of the gradient play algorithm for stochastic games (SGs), where each agent tries to maximize its own total discounted reward by making decisions independently based on current state information which is shared…

机器学习 · 计算机科学 2023-12-08 Runyu Zhang , Zhaolin Ren , Na Li

A strategy profile in a multi-player game is a Nash equilibrium if no player can unilaterally deviate to achieve a strictly better payoff. A profile is an $\epsilon$-Nash equilibrium if no player can gain more than $\epsilon$ by…

计算机科学与博弈论 · 计算机科学 2026-01-27 Ali Asadi , Léonard Brice , Krishnendu Chatterjee , K. S. Thejaswini

We study the complexity of computing stationary Nash equilibrium (NE) in n-player infinite-horizon general-sum stochastic games. We focus on the problem of computing NE in such stochastic games when each player is restricted to choosing a…

计算机科学与博弈论 · 计算机科学 2022-11-30 Yujia Jin , Vidya Muthukumar , Aaron Sidford

In multi-agent autonomous systems, deception is a fundamental concept which characterizes the exploitation of unbalanced information to mislead victims into choosing oblivious actions. This effectively alters the system's long term…

系统与控制 · 电气工程与系统科学 2025-08-27 Michael Tang , Miroslav Krstic , Jorge Poveda

We formulate two-party policy competition as a two-player non-cooperative game, generalizing Lin et al.'s work (2021). Each party selects a real-valued policy vector as its strategy from a compact subset of Euclidean space, and a voter's…

计算机科学与博弈论 · 计算机科学 2026-03-19 Chuang-Chieh Lin , Chi-Jen Lu , Po-An Chen , Chih-Chieh Hung

Recent extensions to dynamic games of the well-known fictitious play learning procedure in static games were proved to globally converge to stationary Nash equilibria in two important classes of dynamic games (zero-sum and…

计算机科学与博弈论 · 计算机科学 2022-07-08 Lucas Baudin , Rida Laraki

Motivated by the scarcity of accurate payoff feedback in practical applications of game theory, we examine a class of learning dynamics where players adjust their choices based on past payoff observations that are subject to noise and…

最优化与控制 · 数学 2016-06-03 Mario Bravo , Panayotis Mertikopoulos

Game theory is a very profound study on distributed decision-making behavior and has been extensively developed by many scholars. However, many existing works rely on certain strict assumptions such as knowing the opponent's private…

计算机科学与博弈论 · 计算机科学 2020-04-21 Kuo Chun Tsai , Zhu Han

This paper deals with N-person nonzero-sum discrete-time Markov games under a probability criterion, in which the transition probabilities and reward functions are allowed to vary with time. Differing from the existing works on the expected…

概率论 · 数学 2025-05-16 Xin Guo , Xin Wen

This paper investigates online stochastic aggregative games subject to local set constraints and time-varying coupled inequality constraints, where each player possesses a time-varying expectation-valued cost function relying on not only…

最优化与控制 · 数学 2025-11-18 Kaixin Du , Min Meng

In this paper, we study a subclass of n-player stochastic games, in which each player has their own internal state controlled only by their own action and their objective is a common goal called team variance which measures the total…

最优化与控制 · 数学 2025-07-31 Li Xia

Consider a strongly monotone game where the players' utility functions include a reward function and a linear term for each dimension, with coefficients that are controlled by the manager. Gradient play converges to a unique Nash…

多智能体系统 · 计算机科学 2026-02-25 Siddharth Chandak , Ilai Bistritz , Nicholas Bambos

We study Nash equilibrium learning in partially observable Markov games (POMGs), a multi-agent reinforcement learning framework in which agents cannot fully observe the underlying state. Prior work in this setting relies on centralization…

计算机科学与博弈论 · 计算机科学 2026-05-08 Philip Jordan , Maryam Kamgarpour

We consider two classes of constrained finite state-action stochastic games. First, we consider a two player nonzero sum single controller constrained stochastic game with both average and discounted cost criterion. We consider the same…

最优化与控制 · 数学 2012-06-11 Vikas Vikram Singh , N. Hemachandra

In this paper, we apply the idea of fictitious play to design deep neural networks (DNNs), and develop deep learning theory and algorithms for computing the Nash equilibrium of asymmetric $N$-player non-zero-sum stochastic differential…

最优化与控制 · 数学 2020-09-07 Ruimeng Hu

In this work, we study stochastic non-cooperative games, where only noisy black-box function evaluations are available to estimate the cost function for each player. Since each player's cost function depends on both its own decision…

计算机科学与博弈论 · 计算机科学 2025-11-18 Haidong Li , Anzhi Sheng , Yijie Peng , Long Wang

In this paper, we consider two-player zero-sum matrix and stochastic games and develop learning dynamics that are payoff-based, convergent, rational, and symmetric between the two players. Specifically, the learning dynamics for matrix…

机器学习 · 计算机科学 2024-09-06 Zaiwei Chen , Kaiqing Zhang , Eric Mazumdar , Asuman Ozdaglar , Adam Wierman

Learning problems commonly exhibit an interesting feedback mechanism wherein the population data reacts to competing decision makers' actions. This paper formulates a new game theoretic framework for this phenomenon, called "multi-player…

计算机科学与博弈论 · 计算机科学 2022-04-08 Adhyyan Narang , Evan Faulkner , Dmitriy Drusvyatskiy , Maryam Fazel , Lillian J. Ratliff

In this paper we propose and analyze a class of $N$-player stochastic games that include finite fuel stochastic games as a special case. We first derive sufficient conditions for the Nash equilibrium (NE) in the form of a verification…

数理金融 · 定量金融 2021-10-26 Xin Guo , Wenpin Tang , Renyuan Xu
‹ 上一页 1 2 3 10 下一页 ›