中文
相关论文

相关论文: Theoretical and Practical Advances on Smoothing fo…

200 篇论文

In this paper, we establish efficient and uncoupled learning dynamics so that, when employed by all players in multiplayer perfect-recall imperfect-information extensive-form games, the trigger regret of each player grows as $O(\log T)$…

计算机科学与博弈论 · 计算机科学 2023-09-20 Ioannis Anagnostides , Gabriele Farina , Tuomas Sandholm

In this paper, we introduce the first algorithmic framework for Blackwell approachability on the sequence-form polytope, the class of convex polytopes capturing the strategies of players in extensive-form games (EFGs). This leads to a new…

计算机科学与博弈论 · 计算机科学 2024-03-08 Darshan Chakrabarti , Julien Grand-Clément , Christian Kroer

Iterated regret minimization has been introduced recently by J.Y. Halpern and R. Pass in classical strategic games. For many games of interest, this new solution concept provides solutions that are judged more reasonable than solutions…

计算机科学与博弈论 · 计算机科学 2015-05-18 Emmanuel Filiot , Tristan Le Gall , Jean-François Raskin

The entropic fictitious play (EFP) is a recently proposed algorithm that minimizes the sum of a convex functional and entropy in the space of measures -- such an objective naturally arises in the optimization of a two-layer neural network…

机器学习 · 统计学 2023-03-07 Atsushi Nitanda , Kazusato Oko , Denny Wu , Nobuhito Takenouchi , Taiji Suzuki

Here, we consider one-dimensional forward-forward mean-field games (MFGs) with congestion, which were introduced to approximate stationary MFGs. We use methods from the theory of conservation laws to examine the qualitative properties of…

偏微分方程分析 · 数学 2017-03-30 Diogo Gomes , Marc Sedjro

Extensive-form games with imperfect recall are an important game-theoretic model that allows a compact representation of strategies in dynamic strategic interactions. Practical use of imperfect recall games is limited due to negative…

计算机科学与博弈论 · 计算机科学 2017-05-25 Branislav Bosansky , Jiri Cermak , Karel Horak , Michal Pechoucek

Feedback delays are inevitable in real-world multi-agent learning. They are known to severely degrade performance, and the convergence rate under delayed feedback is still unclear, even for bilinear games. This paper derives the rate of…

机器学习 · 计算机科学 2026-02-20 Yuma Fujimoto , Kenshi Abe , Kaito Ariu

We focus on the design of algorithms for finding equilibria in 2-player zero-sum games. Although it is well known that such problems can be solved by a single linear program, there has been a surge of interest in recent years for simpler…

计算机科学与博弈论 · 计算机科学 2025-02-03 Michail Fasoulakis , Evangelos Markakis , Giorgos Roussakis , Christodoulos Santorinaios

The CFR framework has been a powerful tool for solving large-scale extensive-form games in practice. However, the theoretical rate at which past CFR-based algorithms converge to the Nash equilibrium is on the order of $O(T^{-1/2})$, where…

计算机科学与博弈论 · 计算机科学 2019-02-14 Gabriele Farina , Christian Kroer , Noam Brown , Tuomas Sandholm

Regret minimization is a powerful method for finding Nash equilibria in Normal-Form Games (NFGs) and Extensive-Form Games (EFGs), but it typically guarantees convergence only for the average strategy. However, computing the average strategy…

计算机科学与博弈论 · 计算机科学 2025-09-18 Hang Ren , Yulin Wu , Shuhan Qi , Jiajia Zhang , Xiaozhen Sun , Tianzi Ma , Xuan Wang

For two classes of Mean Field Game systems we study the convergence of solutions as the interest rate in the cost functional becomes very large, modeling agents caring only about a very short time-horizon, and the cost of the control…

最优化与控制 · 数学 2020-04-10 Martino Bardi , Pierre Cardaliaguet

Discounted-sum games provide a formal model for the study of reinforcement learning, where the agent is enticed to get rewards early since later rewards are discounted. When the agent interacts with the environment, she may regret her…

计算机科学与博弈论 · 计算机科学 2018-11-20 Michaël Cadilhac , Guillermo A. Pérez , Marie van den Bogaard

Counterfactual regret minimization (CFR) is the most popular algorithm on solving two-player zero-sum extensive games with imperfect information and achieves state-of-the-art performance in practice. However, the performance of CFR is not…

机器学习 · 计算机科学 2018-12-27 Yichi Zhou , Tongzheng Ren , Jialian Li , Dong Yan , Jun Zhu

It was recently established that for convex optimization problems with sparse optimal solutions (be it entry-wise sparsity or matrix rank-wise sparsity) it is possible to design first-order methods with linear convergence rates that depend…

最优化与控制 · 数学 2026-03-20 Dan Garber

Scale-invariance in games has recently emerged as a widely valued desirable property. Yet, almost all fast convergence guarantees in learning in games require prior knowledge of the utility scale. To address this, we develop learning…

计算机科学与博弈论 · 计算机科学 2026-02-13 Taira Tsuchiya , Haipeng Luo , Shinji Ito

A set of accelerated first order algorithms with memory are proposed for minimising strongly convex functions. The algorithms are differentiated by their use of the iterate history for the gradient step. The increased convergence rate of…

最优化与控制 · 数学 2018-08-31 Ross Drummond , Stephen Duncan

Tree-form sequential decision making (TFSDM) extends classical one-shot decision making by modeling tree-form interactions between an agent and a potentially adversarial environment. It captures the online decision-making problems that each…

计算机科学与博弈论 · 计算机科学 2021-03-09 Gabriele Farina , Robin Schmucker , Tuomas Sandholm

The practical scalability of many optimization algorithms for large extensive-form games is often limited by the games' huge payoff matrices. To ameliorate the issue, Zhang and Sandholm (2020) recently proposed a sparsification technique…

计算机科学与博弈论 · 计算机科学 2021-12-08 Gabriele Farina , Tuomas Sandholm

This paper proposes a Smoothing Accelerated Proximal Gradient Method with Extrapolation Term (SAPGM) for nonsmooth multiobjective optimization. By combining the smoothing methods and the accelerated algorithm for multiobjective optimization…

最优化与控制 · 数学 2024-10-21 Chengzhi Huang

In a recent paper, Bubeck, Lee, and Singh introduced a new first order method for minimizing smooth strongly convex functions. Their geometric descent algorithm, largely inspired by the ellipsoid method, enjoys the optimal linear rate of…

最优化与控制 · 数学 2017-03-02 Dmitriy Drusvyatskiy , Maryam Fazel , Scott Roy