中文
相关论文

相关论文: Convergence Rate of Learning a Strongly Variationa…

200 篇论文

We consider the problem of learning stable matchings with unknown preferences in a decentralized and uncoordinated manner, where "decentralized" means that players make decisions individually without the influence of a central platform, and…

计算机科学与博弈论 · 计算机科学 2024-08-16 S. Rasoul Etesami , R. Srikant

We consider two classes of constrained finite state-action stochastic games. First, we consider a two player nonzero sum single controller constrained stochastic game with both average and discounted cost criterion. We consider the same…

最优化与控制 · 数学 2012-06-11 Vikas Vikram Singh , N. Hemachandra

Reinforcement learning from self-play has recently reported many successes. Self-play, where the agents compete with themselves, is often used to generate training data for iterative policy improvement. In previous work, heuristic rules are…

机器学习 · 计算机科学 2020-09-15 Yuanyi Zhong , Yuan Zhou , Jian Peng

Motivated by Generative Adversarial Networks, we study the computation of Nash equilibrium in concave network zero-sum games (NZSGs), a multiplayer generalization of two-player zero-sum games first proposed with linear payoffs. Extending…

机器学习 · 计算机科学 2020-07-13 Amit Kadan , Hu Fu

We study best-response type learning dynamics for zero-sum polymatrix games under two information settings. The two settings are distinguished by the type of information that each player has about the game and their opponents' strategy. The…

最优化与控制 · 数学 2025-08-13 Fathima Zarin Faizal , Asuman Ozdaglar , Martin J. Wainwright

We examine the long-run behavior of multi-agent online learning in games that evolve over time. Specifically, we focus on a wide class of policies based on mirror descent, and we show that the induced sequence of play (a) converges to Nash…

计算机科学与博弈论 · 计算机科学 2022-08-11 Benoit Duvocelle , Panayotis Mertikopoulos , Mathias Staudigl , Dries Vermeulen

We analyze best response dynamics for finding a Nash equilibrium of an infinite horizon zero-sum stochastic linear quadratic dynamic game (LQDG) with partial and asymmetric information. We derive explicit expressions for each player's best…

系统与控制 · 电气工程与系统科学 2025-09-03 Yuxiang Guan , Iman Shames , Tyler H. Summers

We introduce a new class of population games that we call monotropic; these are games characterized by the presence of a unique globally neutrally stable Nash equilibrium. Monotropic games generalize strictly concave potential games and…

计算机科学与博弈论 · 计算机科学 2014-04-22 Ioannis Avramopoulos

We characterize Nash equilibrium by postulating coherent behavior across varying games. Nash equilibrium is the only solution concept that satisfies the following axioms: (i) strictly dominant actions are played with positive probability,…

理论经济学 · 经济学 2024-07-02 Florian Brandl , Felix Brandt

This paper considers a class of reinforcement-learning that belongs to the family of Learning Automata and provides a stochastic-stability analysis in strategic-form games. For this class of dynamics, convergence to pure Nash equilibria has…

计算机科学与博弈论 · 计算机科学 2017-02-28 Georgios C. Chasparis

This paper considers the problem of designing optimal algorithms for reinforcement learning in two-player zero-sum games. We focus on self-play algorithms which learn the optimal policy by playing against itself without any direct…

机器学习 · 计算机科学 2020-07-15 Yu Bai , Chi Jin , Tiancheng Yu

We formulate a general framework for competitive gradient-based learning that encompasses a wide breadth of multi-agent learning algorithms, and analyze the limiting behavior of competitive gradient-based learning algorithms using dynamical…

机器学习 · 计算机科学 2020-02-21 Eric Mazumdar , Lillian J. Ratliff , S. Shankar Sastry

The study of learning in games has thus far focused primarily on normal form games. In contrast, our understanding of learning in extensive form games (EFGs) and particularly in EFGs with many agents lags far behind, despite them being…

计算机科学与博弈论 · 计算机科学 2022-07-19 Georgios Piliouras , Lillian Ratliff , Ryann Sim , Stratis Skoulakis

We introduce a novel class of Nash equilibrium seeking dynamics for non-cooperative games with a finite number of players, where the convergence to the Nash equilibrium is bounded by a KL function with a settling time that can be upper…

最优化与控制 · 数学 2020-12-25 Jorge I. Poveda , Miroslav Krstic , Tamer Basar

We discuss a natural game of competition and solve the corresponding mean field game with \emph{common noise} when agents' rewards are \emph{rank dependent}. We use this solution to provide an approximate Nash equilibrium for the finite…

概率论 · 数学 2016-10-18 Erhan Bayraktar , Yuchong Zhang

This paper is concerned with non-zero sum differential games of mean-field stochastic differential equations with partial information and convex control domain. First, applying the classical convex variations, we obtain stochastic maximum…

最优化与控制 · 数学 2016-01-11 Hua Xiao , Shuaiqi Zhang

We give a simple proof of the well-known result that the marginal strategies of a coarse correlated equilibrium form a Nash equilibrium in two-player zero-sum games. A corollary of this fact is that no-external-regret learning algorithms…

计算机科学与博弈论 · 计算机科学 2023-10-18 Revan MacQueen

This paper examines the convergence behaviour of simultaneous best-response dynamics in random potential games. We provide a theoretical result showing that, for two-player games with sufficiently many actions, the dynamics converge quickly…

计算机科学与博弈论 · 计算机科学 2025-05-19 Galit Ashkenazi-Golan , Domenico Mergoni Cecchelli , Edward Plumb

In this paper, we address the inverse problem in the case of linear-quadratic discrete-time dynamic non-cooperative games. Given feedback laws of players that are known to be a Nash equilibrium pair for a discrete-time linear system, we…

最优化与控制 · 数学 2024-07-19 Emin Martirosyan , Ming Cao

A new class of projected dynamical systems of third order is investigated for quasi (parametric) variational inequalities in which the convex set in the classical variational inequality also depends upon the solution explicitly or…

最优化与控制 · 数学 2025-01-10 Oday Hazaimah