中文
相关论文

相关论文: Last-iterate Convergence for Symmetric, General-su…

200 篇论文

Do boundedly rational players learn to choose equilibrium strategies as they play a game repeatedly? A large literature in behavioral game theory has proposed and experimentally tested various learning algorithms, but a comparative analysis…

经济学 · 定量金融 2021-09-03 Marco Pangallo , James Sanders , Tobias Galla , Doyne Farmer

We consider payoff-based learning of a generalized Nash equilibrium (GNE) in multi-agent systems. Our focus is on games with jointly convex constraints of a linear structure and strongly monotone pseudo-gradients. We present a convergent…

最优化与控制 · 数学 2025-07-18 Tatiana Tatarenko , Maryam Kamgarpour

This paper examines the convergence of no-regret learning in games with continuous action sets. For concreteness, we focus on learning via "dual averaging", a widely used class of no-regret learning schemes where players take small steps…

最优化与控制 · 数学 2018-01-17 Panayotis Mertikopoulos , Zhengyuan Zhou

This paper presents a learning dynamic with almost sure convergence guarantee for any stochastic game with turn-based controllers (on state transitions) as long as stage-payoffs induce a zero-sum or identical-interest game. Stage-payoffs…

计算机科学与博弈论 · 计算机科学 2023-10-11 Muhammed O. Sayin

We study two-player (zero-sum) concurrent mean-payoff games played on a finite-state graph. We focus on the important sub-class of ergodic games where all states are visited infinitely often with probability 1. The algorithmic study of…

计算机科学与博弈论 · 计算机科学 2014-04-24 Krishnendu Chatterjee , Rasmus Ibsen-Jensen

Evolutionary game theory has been a successful tool to combine classical game theory with learning-dynamical descriptions in multiagent systems. Provided some symmetric structures of interacting players, many studies have been focused on…

人工智能 · 计算机科学 2022-06-23 Xinyu Zhang , Peng Peng , Yushan Zhou , Haifeng Wang , Wenxin Li

This paper presents a universal representation of symmetric (permutation-invariant) functions with multidimensional variable-size variables. These representations help justify approximation methods that aggregate information from each…

综合经济学 · 经济学 2025-05-22 Takeshi Fukasawa

We are interested in the convergence of the value of n-stage games as n goes to infinity and the existence of the uniform value in stochastic games with a general set of states and finite sets of actions where the transition is commutative.…

最优化与控制 · 数学 2016-04-22 Xavier Venel

We study the convergence of Optimistic Gradient Descent Ascent in unconstrained bilinear games. In a first part, we consider the zero-sum case and extend previous results by Daskalakis et al. in 2018, Liang and Stokes in 2019, and others:…

最优化与控制 · 数学 2022-11-24 Étienne de Montbrun , Jérôme Renault

We obtain global, non-asymptotic convergence guarantees for independent learning algorithms in competitive reinforcement learning settings with two agents (i.e., zero-sum stochastic games). We consider an episodic setting where in each…

机器学习 · 计算机科学 2021-01-13 Constantinos Daskalakis , Dylan J. Foster , Noah Golowich

A partial differential equation is derived, describing the replicator dynamics with mutations of games with a continuous strategy space. This equation is then applied to continuous versions of symmetric 2x2 games, such as the Prisoners…

适应与自组织系统 · 物理学 2007-05-23 M. Ruijgrok , T. W. Ruijgrok

This paper proves several Tauberian theorems for general iterations of operators, and provides two applications to zero-sum stochastic games where the total payoff is a weighted sum of the stage payoffs. The first application is to provide…

最优化与控制 · 数学 2016-09-09 Bruno Ziliotto

The predominant paradigm in evolutionary game theory and more generally online learning in games is based on a clear distinction between a population of dynamic agents that interact given a fixed, static game. In this paper, we move away…

计算机科学与博弈论 · 计算机科学 2020-12-16 Stratis Skoulakis , Tanner Fiez , Ryann Sim , Georgios Piliouras , Lillian Ratliff

We introduce novel multi-agent interaction models of entropic spatially inhomogeneous evolutionary undisclosed games and their quasi-static limits. These evolutions vastly generalize first and second order dynamics. Besides the…

最优化与控制 · 数学 2022-03-10 Mauro Bonafini , Massimo Fornasier , Bernhard Schmitzer

Economic ensembles can be modeled as networks of interacting agents whose be-haviors are described in terms of game theory. The evolutionary paradigm has been applied to two-person games to discover strategies in this context.…

凝聚态物理 · 物理学 2007-05-23 Wan Ahmad Tajuddin Wan Abdullah

Game theory is playing more and more important roles in understanding complex systems and in investigating intelligent machines with various uncertainties. As a starting point, we consider the classical two-player zero-sum linear-quadratic…

最优化与控制 · 数学 2022-04-20 Nian Liu , Lei Guo

We consider two-player games played on weighted directed graphs with mean-payoff and total-payoff objectives, two classical quantitative objectives. While for single-dimensional games the complexity and memory bounds for both objectives…

计算机科学与博弈论 · 计算机科学 2014-11-04 Krishnendu Chatterjee , Laurent Doyen , Mickael Randour , Jean-François Raskin

The Multiplicative Weights Update (MWU) method is a ubiquitous meta-algorithm that works as follows: A distribution is maintained on a certain set, and at each step the probability assigned to element $\gamma$ is multiplied by $(1 -\epsilon…

计算机科学与博弈论 · 计算机科学 2017-03-06 Gerasimos Palaiopanos , Ioannis Panageas , Georgios Piliouras

We consider 2-player stochastic games with perfectly observed actions, and study the limit, as the discount factor goes to one, of the equilibrium payoffs set. In the usual setup where current states are observed by the players, we show…

最优化与控制 · 数学 2014-12-11 Jérôme Renault , Bruno Ziliotto

We study agents competing against each other in a repeated network zero-sum game while applying the multiplicative weights update (MWU) algorithm with fixed learning rates. In our implementation, agents select their strategies…

计算机科学与博弈论 · 计算机科学 2021-10-06 James P. Bailey , Sai Ganesh Nagarajan , Georgios Piliouras