English
Related papers

Related papers: Theoretical and Practical Advances on Smoothing fo…

200 papers

Extensive-form games (EFGs) provide a powerful framework for modeling sequential decision making, capturing strategic interaction under imperfect information, chance events, and temporal structure. Most positive algorithmic and theoretical…

Computer Science and Game Theory · Computer Science 2026-05-26 Rui Zheng , Ryann Sim , Antonios Varvitsiotis

In this paper, we present exploitability descent, a new algorithm to compute approximate equilibria in two-player zero-sum extensive-form games with imperfect information, by direct policy optimization against worst-case opponents. We prove…

Artificial Intelligence · Computer Science 2020-06-15 Edward Lockhart , Marc Lanctot , Julien Pérolat , Jean-Baptiste Lespiau , Dustin Morrill , Finbarr Timbers , Karl Tuyls

The paper is concerned with distributed learning in large-scale games. The well-known fictitious play (FP) algorithm is addressed, which, despite theoretical convergence results, might be impractical to implement in large-scale settings due…

Optimization and Control · Mathematics 2016-11-17 Brian Swenson , Soummya Kar , Joao Xavier

Recent breakthrough results by Dagan, Daskalakis, Fishelson and Golowich [2023] and Peng and Rubinstein [2023] established an efficient algorithm attaining at most $\epsilon$ swap regret over extensive-form strategy spaces of dimension $N$…

Computer Science and Game Theory · Computer Science 2025-02-14 Brian Hu Zhang , Ioannis Anagnostides , Gabriele Farina , Tuomas Sandholm

Gradient-based algorithms have shown great promise in solving large (two-player) zero-sum games. However, their success has been mostly confined to the low-precision regime since the number of iterations grows polynomially in $1/\epsilon$,…

Computer Science and Game Theory · Computer Science 2024-10-30 Ioannis Anagnostides , Tuomas Sandholm

Counterfactual regret minimization (CFR) is a family of algorithms for effectively solving imperfect-information games. It decomposes the total regret into counterfactual regrets, utilizing local regret minimization algorithms, such as…

Machine Learning · Computer Science 2024-05-15 Hang Xu , Kai Li , Bingyun Liu , Haobo Fu , Qiang Fu , Junliang Xing , Jian Cheng

We study the problem of finding optimal correlated equilibria of various sorts in extensive-form games: normal-form coarse correlated equilibrium (NFCCE), extensive-form coarse correlated equilibrium (EFCCE), and extensive-form correlated…

Computer Science and Game Theory · Computer Science 2025-01-28 Brian Zhang , Gabriele Farina , Andrea Celli , Tuomas Sandholm

Cost-efficient compressive sensing is challenging when facing large-scale data, {\em i.e.}, data with large sizes. Conventional compressive sensing methods for large-scale data will suffer from low computational efficiency and massive…

Data Structures and Algorithms · Computer Science 2016-03-18 Sung-Hsien Hsieh , Chun-Shien Lu , Soo-Chang Pei

We develop both first and second order numerical optimization methods to solve non-smooth optimization problems featuring a shared sparsity penalty, constrained by differential equations with uncertainty. To alleviate the curse of…

Optimization and Control · Mathematics 2025-09-18 Harbir Antil , Sergey Dolgov , Akwum Onwunta

We address the challenge of exploration in reinforcement learning (RL) when the agent operates in an unknown environment with sparse or no rewards. In this work, we study the maximum entropy exploration problem of two different types. The…

Policy gradient methods have become a staple of any single-agent reinforcement learning toolbox, due to their combination of desirable properties: iterate convergence, efficient use of stochastic trajectory feedback, and theoretically-sound…

Computer Science and Game Theory · Computer Science 2025-07-10 Mingyang Liu , Gabriele Farina , Asuman Ozdaglar

Force-directed approach is one of the most widely used methods in graph drawing research. There are two main problems with the traditional force-directed algorithms. First, there is no mature theory to ensure the convergence of iteration…

Computational Geometry · Computer Science 2018-03-12 Yong-Xian Wang , Zheng-Hua Wang

We consider the use of no-regret algorithms to compute equilibria for particular classes of convex-concave games. While standard regret bounds would lead to convergence rates on the order of $O(T^{-1/2})$, recent work \citep{RS13,SALS15}…

Machine Learning · Computer Science 2018-05-18 Jacob Abernethy , Kevin A. Lai , Kfir Y. Levy , Jun-Kun Wang

Continuous DR-submodular functions are a class of functions that satisfy the Diminishing Returns (DR) property, which implies that they are concave along non-negative directions. Existing works have studied monotone continuous DR-submodular…

Machine Learning · Computer Science 2022-05-31 Omid Sadeghi , Maryam Fazel

Imperfect-Information Extensive-Form Games (IIEFGs) is a prevalent model for real-world games involving imperfect information and sequential plays. The Extensive-Form Correlated Equilibrium (EFCE) has been proposed as a natural solution…

Machine Learning · Computer Science 2022-05-17 Ziang Song , Song Mei , Yu Bai

We study the existence of classical solutions to a broad class of local, first order, forward-backward Extended Mean Field Games systems, that includes standard Mean Field Games, Mean Field Games with congestion, and mean field type control…

Analysis of PDEs · Mathematics 2023-01-12 Sebastian Munoz

In this work, we establish near-linear and strong convergence for a natural first-order iterative algorithm that simulates Von Neumann's Alternating Projections method in zero-sum games. First, we provide a precise analysis of Optimistic…

Optimization and Control · Mathematics 2021-08-18 Ioannis Anagnostides , Paolo Penna

In this paper, we propose a passivity-based methodology for analysis and design of reinforcement learning in multi-agent finite games. Starting from a known exponentially-discounted reinforcement learning scheme, we show that convergence to…

Optimization and Control · Mathematics 2024-10-30 Bolin Gao , Lacra Pavel

The theory of first-order mean field type differential games examines the systems of infinitely many identical agents interacting via some external media under assumption that each agent is controlled by two players. We study the…

Optimization and Control · Mathematics 2020-11-24 Yurii Averboukh

By incorporating regret minimization, double oracle methods have demonstrated rapid convergence to Nash Equilibrium (NE) in normal-form games and extensive-form games, through algorithms such as online double oracle (ODO) and extensive-form…

Computer Science and Game Theory · Computer Science 2023-07-14 Xiaohang Tang , Le Cong Dinh , Stephen Marcus McAleer , Yaodong Yang