English
Related papers

Related papers: Last-iterate Convergence in Regularized Graphon Me…

200 papers

We study online learning and equilibrium computation in games with polyhedral decision sets, a property shared by both normal-form games and extensive-form games (EFGs), when the learning agent is restricted to using a best-response oracle.…

Computer Science and Game Theory · Computer Science 2023-12-07 Darshan Chakrabarti , Gabriele Farina , Christian Kroer

In this paper, the optimal convergence rate $O\left(N^{-1/2}\right)$ (where $N$ is the total number of iterations performed by the algorithm), without the presence of a logarithmic factor, is proved for mirror descent algorithms with…

Optimization and Control · Mathematics 2025-06-04 Mohammad Alkousa , Fedor Stonyakin , Asmaa Abdo , Mohammad Alcheikh

Mean Field Games (MFG) are the class of games with a very large number of agents and the standard equilibrium concept is a Mean Field Equilibrium (MFE). Algorithms for learning MFE in dynamic MFGs are unknown in general. Our focus is on an…

Optimization and Control · Mathematics 2021-02-02 Kiyeob Lee , Desik Rengarajan , Dileep Kalathil , Srinivas Shakkottai

Reinforcement Learning from Human Feedback (RLHF) has been highly successful in aligning large language models with human preferences. While prevalent methods like DPO have demonstrated strong performance, they frame interactions with the…

Machine Learning · Computer Science 2025-05-27 Yongtao Wu , Luca Viano , Yihang Chen , Zhenyu Zhu , Kimon Antonakopoulos , Quanquan Gu , Volkan Cevher

Regret matching (RM) -- and its modern variants -- is a foundational online algorithm that has been at the heart of many AI breakthrough results in solving benchmark zero-sum games, such as poker. Yet, surprisingly little is known so far in…

Computer Science and Game Theory · Computer Science 2025-11-18 Ioannis Anagnostides , Emanuel Tewolde , Brian Hu Zhang , Ioannis Panageas , Vincent Conitzer , Tuomas Sandholm

Offline multi-agent reinforcement learning in general-sum settings is challenged by the distribution shift between logged datasets and target equilibrium policies. While standard methods rely on manual pessimistic penalties, we demonstrate…

Machine Learning · Computer Science 2026-05-19 Claire Chen , Yuheng Zhang

In this paper we study a fully discrete Semi-Lagrangian approximation of a second order Mean Field Game system, which can be degenerate. We prove that the resulting scheme is well posed and, if the state dimension is equals to one, we prove…

Numerical Analysis · Mathematics 2014-04-24 Elisabetta Carlini , Francisco José Silva Álvarez

In this paper, we provide exponential rates of convergence to the interior Nash equilibrium for continuous-time dual-space game dynamics such as mirror descent (MD) and actor-critic (AC). We perform our analysis in $N$-player continuous…

Optimization and Control · Mathematics 2022-02-04 Bolin Gao , Lacra Pavel

The approximation of mixed Nash equilibria (MNE) for zero-sum games with mean-field interacting players has recently raised much interest in machine learning. In this paper we propose a mean-field gradient descent dynamics for finding the…

Optimization and Control · Mathematics 2025-05-13 Yulong Lu , Pierre Monmarché

We design and analyze minimax-optimal algorithms for online linear optimization games where the player's choice is unconstrained. The player strives to minimize regret, the difference between his loss and the loss of a post-hoc benchmark…

Machine Learning · Computer Science 2013-02-12 H. Brendan McMahan

The monotone variational inequality is a central problem in mathematical programming that unifies and generalizes many important settings such as smooth convex optimization, two-player zero-sum games, convex-concave saddle point problems,…

Optimization and Control · Mathematics 2022-05-17 Yang Cai , Argyris Oikonomou , Weiqiang Zheng

Counterfactual Regret Minimization (CFR) and its variants developed based upon Regret Matching (RM) have been considered to be the best method to solve incomplete information extensive form games. In addition to RM and CFR, Fictitious Play…

Computer Science and Game Theory · Computer Science 2023-11-14 Qi Ju

Here, we prove the existence of solutions to first-order mean-field games (MFGs) arising in optimal switching. First, we use the penalization method to construct approximate solutions. Then, we prove uniform estimates for the penalized…

Analysis of PDEs · Mathematics 2016-10-04 Diogo A. Gomes , Stefania Patrizi

Motivated by numerical challenges in first-order mean field games (MFGs) and the weak noise theory for the Kardar-Parisi-Zhang equation, we consider the problem of vanishing viscosity approximations for MFGs. We provide the first results on…

Analysis of PDEs · Mathematics 2023-04-04 Wenpin Tang , Yuming Paul Zhang

This paper considers no-regret learning for repeated continuous-kernel games with lossy bandit feedback. Since it is difficult to give the explicit model of the utility functions in dynamic environments, the players' action can only be…

Machine Learning · Computer Science 2022-05-17 Wenting Liu , Jinlong Lei , Peng Yi , Yiguang Hong

The emergence of the graphon theory of large networks and their infinite limits has enabled the formulation of a theory of the centralized control of dynamical systems distributed on asymptotically infinite networks (Gao and Caines, IEEE…

Optimization and Control · Mathematics 2021-12-30 Peter E. Caines , Minyi Huang

Online learning algorithms are fast, memory-efficient, easy to implement, and applicable to many prediction problems, including classification, regression, and ranking. Several online algorithms were proposed in the past few decades, some…

Machine Learning · Computer Science 2015-07-03 Francesco Orabona , Koby Crammer , Nicolò Cesa-Bianchi

This work studies an algorithm, which we call magnetic mirror descent, that is inspired by mirror descent and the non-Euclidean proximal gradient algorithm. Our contribution is demonstrating the virtues of magnetic mirror descent as both an…

In this paper we study the smooth convex-concave saddle point problem. Specifically, we analyze the last iterate convergence properties of the Extragradient (EG) algorithm. It is well known that the ergodic (averaged) iterates of EG…

Machine Learning · Computer Science 2020-07-08 Noah Golowich , Sarath Pattathil , Constantinos Daskalakis , Asuman Ozdaglar

Motivated by applications in Game Theory, Optimization, and Generative Adversarial Networks, recent work of Daskalakis et al \cite{DISZ17} and follow-up work of Liang and Stokes \cite{LiangS18} have established that a variant of the widely…

Optimization and Control · Mathematics 2025-09-30 Constantinos Daskalakis , Ioannis Panageas
‹ Prev 1 8 9 10 Next ›