中文
相关论文

相关论文: Last-iterate Convergence for Symmetric, General-su…

200 篇论文

Evolutionary game theory studies populations that change in response to an underlying game. Often, the functional form relating outcome to player attributes or strategy is complex, preventing mathematical progress. In this work, we…

计算机科学与博弈论 · 计算机科学 2025-11-25 Pablo Lechon-Alonso , Andrew Dennehy , Ruizheng Bai , Nicolas Sanchez , Derek K. Wise , David Sewell , David Rosenbluth , Alexander Strang

We study infinite-horizon discounted two-player zero-sum Markov games, and develop a decentralized algorithm that provably converges to the set of Nash equilibria under self-play. Our algorithm is based on running an Optimistic Gradient…

机器学习 · 计算机科学 2021-07-08 Chen-Yu Wei , Chung-Wei Lee , Mengxiao Zhang , Haipeng Luo

This paper investigates the discrete-time asynchronous games in which noncooperative agents seek to minimize their individual cost functions. Building on the assumption of partial asynchronism, i.e., each agent updates at least once within…

最优化与控制 · 数学 2025-08-13 Zifan Wang , Xinlei Yi , Michael M. Zavlanos , Karl H. Johansson

Quantitative games are two-player zero-sum games played on directed weighted graphs. Total-payoff games (that can be seen as a refinement of the well-studied mean-payoff games) are the variant where the payoff of a play is computed as the…

计算机科学与博弈论 · 计算机科学 2015-07-15 Thomas Brihaye , Gilles Geeraerts , Axel Haddad , Benjamin Monmege

Iterated coopetitive games capture the situation when one must efficiently balance between cooperation and competition with the other agents over time in order to win the game (e.g., to become the player with highest total utility).…

计算机科学与博弈论 · 计算机科学 2022-03-11 Shivakumar Mahesh , Nicholas Bishop , Le Cong Dinh , Long Tran-Thanh

We consider two-player random extensive form games where the payoffs at the leaves are independently drawn uniformly at random from a given feasible set C. We study the asymptotic distribution of the subgame perfect equilibrium outcome for…

计算机科学与博弈论 · 计算机科学 2015-09-09 Itai Arieli , Yakov Babichenko

We consider two classes of constrained finite state-action stochastic games. First, we consider a two player nonzero sum single controller constrained stochastic game with both average and discounted cost criterion. We consider the same…

最优化与控制 · 数学 2012-06-11 Vikas Vikram Singh , N. Hemachandra

Scale-invariance in games has recently emerged as a widely valued desirable property. Yet, almost all fast convergence guarantees in learning in games require prior knowledge of the utility scale. To address this, we develop learning…

计算机科学与博弈论 · 计算机科学 2026-02-13 Taira Tsuchiya , Haipeng Luo , Shinji Ito

This paper proposes Mutation-Driven Multiplicative Weights Update (M2WU) for learning an equilibrium in two-player zero-sum normal-form games and proves that it exhibits the last-iterate convergence property in both full and noisy feedback…

计算机科学与博弈论 · 计算机科学 2023-05-29 Kenshi Abe , Kaito Ariu , Mitsuki Sakamoto , Kentaro Toyoshima , Atsushi Iwasaki

This paper studies the convergence of the Optimistic Multiplicative Weights Update algorithm (OMWU) in two player zero-sum games. Recent works have identified instances on which the last-iterate of OMWU can converge arbitrarily slowly, but…

计算机科学与博弈论 · 计算机科学 2026-05-14 John Lazarsfeld , Anas Barakat , Georgios Piliouras , Antonios Varvitsiotis , Andre Wibisono

Most of the literature on learning in games has focused on the restrictive setting where the underlying repeated game does not change over time. Much less is known about the convergence of no-regret learning algorithms in dynamic multiagent…

机器学习 · 计算机科学 2023-10-19 Ioannis Anagnostides , Ioannis Panageas , Gabriele Farina , Tuomas Sandholm

We study two-player zero-sum stochastic games, and propose a form of independent learning dynamics called Doubly Smoothed Best-Response dynamics, which integrates a discrete and doubly smoothed variant of the best-response dynamics into…

计算机科学与博弈论 · 计算机科学 2023-03-07 Zaiwei Chen , Kaiqing Zhang , Eric Mazumdar , Asuman Ozdaglar , Adam Wierman

In this paper, we study nonzero-sum separable games, which are continuous games whose payoffs take a sum-of-products form. Included in this subclass are all finite games and polynomial games. We investigate the structure of equilibria in…

计算机科学与博弈论 · 计算机科学 2010-04-26 Noah D. Stein , Asuman Ozdaglar , Pablo A. Parrilo

This paper studies a class of dynamic Stackelberg games under open-loop information structure with constrained linear agent dynamics and quadratic utility functions. We show two important properties for this class of dynamic Stackelberg…

最优化与控制 · 数学 2016-08-09 Sen Li , Wei Zhang , Jianming Lian , Karanjit Kalsi

Understanding how agents coordinate or compete from limited behavioral data is central to modeling strategic interactions in traffic, robotics, and other multi-agent systems. In this work, we investigate the following complementary…

计算机科学与博弈论 · 计算机科学 2026-01-16 Daniela Aguirre Salazar , Firas Moatemri , Tatiana Tatarenko

We consider 2-players, 2-values minimization games where the players' costs take on two values, $a,b$, $a<b$. The players play mixed strategies and their costs are evaluated by unimodal valuations. This broad class of valuations includes…

计算机科学与博弈论 · 计算机科学 2020-09-10 Chryssis Georgiou , Marios Mavronicolas , Burkhard Monien

Zero-sum stochastic games generalize the notion of Markov Decision Processes (i.e. controlled Markov chains, or stochastic dynamic programming) to the 2-player competitive case : two players jointly control the evolution of a state…

最优化与控制 · 数学 2019-05-17 Jérôme Renault

Evolutionary $2 \times 2$ games are studied with players located on a square lattice. During the evolution the randomly chosen neighboring players try to maximize their collective income by adopting a random strategy pair with a probability…

种群与进化 · 定量生物学 2010-08-23 Gyorgy Szabo , Attila Szolnoki , Melinda Varga , Livia Hanusovszky

Pursuit-evasion scenarios appear widely in robotics, security domains, and many other real-world situations. We focus on two-player pursuit-evasion games with concurrent moves, infinite horizon, and discounted rewards. We assume that the…

计算机科学与博弈论 · 计算机科学 2016-08-05 Karel Horák , Branislav Bošanský

This paper gives a complete analysis of worst-case equilibria for various versions of weighted congestion games with two players and affine cost functions. The results are exact price of anarchy bounds which are parametric in the weights of…

计算机科学与博弈论 · 计算机科学 2022-03-04 Joran van den Bosse , Marc Uetz , Matthias Walter