中文
相关论文

相关论文: Mutation-Driven Follow the Regularized Leader for …

200 篇论文

We present tools for the analysis of Follow-The-Regularized-Leader (FTRL), Dual Averaging, and Mirror Descent algorithms when the regularizer (equivalently, prox-function or learning rate schedule) is chosen adaptively based on the data.…

机器学习 · 计算机科学 2015-11-10 H. Brendan McMahan

We introduce Cautious Optimism, a framework for substantially faster regularized learning in general games. Cautious Optimism, as a variant of Optimism, adaptively controls the learning pace in a dynamic, non-monotone manner to accelerate…

机器学习 · 计算机科学 2025-11-17 Ashkan Soleymani , Georgios Piliouras , Gabriele Farina

Multi-leader multi-follower games are a class of hierarchical games in which a collection of leaders compete in a Nash game constrained by the equilibrium conditions of another Nash game amongst the followers. The resulting equilibrium…

最优化与控制 · 数学 2014-08-27 Ankur A. Kulkarni , Uday V. Shanbhag

We study the emergence of chaotic behavior of Follow-the-Regularized Leader (FoReL) dynamics in games. We focus on the effects of increasing the population size or the scale of costs in congestion games, and generalize recent results on…

计算机科学与博弈论 · 计算机科学 2022-01-28 Jakub Bielawski , Thiparat Chotibut , Fryderyk Falniowski , Grzegorz Kosiorowski , Michał Misiurewicz , Georgios Piliouras

Trust region methods are widely applied in single-agent reinforcement learning problems due to their monotonic performance-improvement guarantee at every iteration. Nonetheless, when applied in multi-agent settings, the guarantee of trust…

多智能体系统 · 计算机科学 2021-06-15 Ying Wen , Hui Chen , Yaodong Yang , Zheng Tian , Minne Li , Xu Chen , Jun Wang

No-regret learning has a long history of being closely connected to game theory. Recent works have devised uncoupled no-regret learning dynamics that, when adopted by all the players in normal-form games, converge to various equilibrium…

计算机科学与博弈论 · 计算机科学 2024-04-24 Weichao Mao , Haoran Qiu , Chen Wang , Hubertus Franke , Zbigniew Kalbarczyk , Tamer Başar

We study online learning problems in which the learner has extra knowledge about the adversary's behaviour, i.e., in game-theoretic settings where opponents typically follow some no-external regret learning algorithms. Under this…

机器学习 · 计算机科学 2023-02-15 Le Cong Dinh , Tri-Dung Nguyen , Alain Zemkoho , Long Tran-Thanh

We consider a bilevel optimization problem having a single leader and multiple followers. The followers choose their strategies simultaneously, and are assumed to converge to a Nash equilibrium strategy profile. We begin by providing a…

最优化与控制 · 数学 2025-03-20 Parin Chaipunya , Thirumulanathan D , Joydeep Dutta

We investigate the strategic surplus obtainable against a Follow-the-Regularized-Leader (FTRL) learner with constant step size $\eta$ in $n\times m$ two-player zero-sum games played over $T$ rounds against a clairvoyant optimizer. In…

计算机科学与博弈论 · 计算机科学 2026-05-25 Yiheng Su , Emmanouil-Vasileios Vlatakis-Gkaragkounis

Last-iterate convergence of learning dynamics in games has attracted significant recent attention. In two-player zero-sum games with bandit feedback, where only the loss of the selected action pair is observed, Fiegel et al. (2025) show a…

机器学习 · 计算机科学 2026-05-12 Soumita Hait , Ping Li , Haipeng Luo , Mengxiao Zhang

Aligning large language models (LLMs) with human preferences has proven effective for enhancing model capabilities, yet standard preference modeling using the Bradley-Terry model assumes transitivity, overlooking the inherent complexity of…

机器学习 · 计算机科学 2026-01-05 Shulun Chen , Runlong Zhou , Zihan Zhang , Maryam Fazel , Simon S. Du

Mean Field Games (MFGs) have the ability to handle large-scale multi-agent systems, but learning Nash equilibria in MFGs remains a challenging task. In this paper, we propose a deep reinforcement learning (DRL) algorithm that achieves…

计算机科学与博弈论 · 计算机科学 2024-03-07 Zida Wu , Mathieu Lauriere , Samuel Jia Cong Chua , Matthieu Geist , Olivier Pietquin , Ankur Mehta

A mean-field game (MFG) seeks the Nash Equilibrium of a game involving a continuum of players, where the Nash Equilibrium corresponds to a fixed point of the best-response mapping. However, simple fixed-point iterations do not always…

最优化与控制 · 数学 2025-07-15 Jiajia Yu , Xiuyuan Cheng , Jian-Guo Liu , Hongkai Zhao

In this paper we investigate the Follow the Regularized Leader dynamics in sequential imperfect information games (IIG). We generalize existing results of Poincar\'e recurrence from normal-form games to zero-sum two-player imperfect…

To model complex real-world systems, such as traders in stock markets, or the dissemination of contagious diseases, graphon mean-field games (GMFG) have been proposed to model many agents. Despite the empirical success, our understanding of…

计算机科学与博弈论 · 计算机科学 2024-10-14 Jing Dong , Baoxiang Wang , Yaoliang Yu

The finitely repeated Prisoners' Dilemma is a good illustration of the discrepancy between the strategic behaviour suggested by a game-theoretic analysis and the behaviour often observed among human players, where cooperation is maintained…

种群与进化 · 定量生物学 2013-01-10 Kristian Lindgren , Vilhelm Verendel

In this work, we introduce the concept of non-negative weighted regret, an extension of non-negative regret \cite{anagnostides2022last} in games. Investigating games with non-negative weighted regret helps us to understand games with…

计算机科学与博弈论 · 计算机科学 2025-05-22 Nanxiang Zhou , Jing Dong , Baoxiang Wang

The understanding of a dynamical system's properties can be significantly advanced by establishing it as a Hamiltonian system and then systematically exploring its inherent symmetries. By formulating agents' strategies and cumulative…

计算机科学与博弈论 · 计算机科学 2025-12-30 Toshihiro Ota , Yuma Fujimoto

We present a novel variant of fictitious play dynamics combining classical fictitious play with Q-learning for stochastic games and analyze its convergence properties in two-player zero-sum stochastic games. Our dynamics involves players…

计算机科学与博弈论 · 计算机科学 2022-06-03 Muhammed O. Sayin , Francesca Parise , Asuman Ozdaglar

In the Lasry--Lions framework, Mean-Field Games (MFGs) model interactions among an infinite number of agents. However, existing algorithms either require strict monotonicity or only guarantee the convergence of averaged iterates, as in…

计算机科学与博弈论 · 计算机科学 2025-10-28 Noboru Isobe , Kenshi Abe , Kaito Ariu