中文
相关论文

相关论文: Sync Pure Counterfactual Regret Minimization in In…

200 篇论文

Regret minimization is a powerful tool for solving large-scale problems; it was recently used in breakthrough results for large-scale extensive-form game solving. This was achieved by composing simplex regret minimizers into an overall…

机器学习 · 计算机科学 2019-02-19 Gabriele Farina , Christian Kroer , Tuomas Sandholm

No-regret learning dynamics play a central role in game theory, enabling decentralized convergence to equilibrium for concepts such as Coarse Correlated Equilibrium (CCE) or Correlated Equilibrium (CE). In this work, we improve the…

计算机科学与博弈论 · 计算机科学 2025-11-05 Asrin Efe Yorulmaz , Tamer Başar

Extensive-form games provide a versatile framework for modeling interactions of multiple agents subjected to imperfect observations and stochastic events. In recent years, two paradigms, policy space response oracles (PSRO) and…

计算机科学与博弈论 · 计算机科学 2022-04-12 Xinrun Wang , Jakub Cerny , Shuxin Li , Chang Yang , Zhuyun Yin , Hau Chan , Bo An

In this paper, we establish efficient and uncoupled learning dynamics so that, when employed by all players in multiplayer perfect-recall imperfect-information extensive-form games, the trigger regret of each player grows as $O(\log T)$…

计算机科学与博弈论 · 计算机科学 2023-09-20 Ioannis Anagnostides , Gabriele Farina , Tuomas Sandholm

A major open question in algorithmic game theory is whether normal-form correlated equilibria (NFCE) can be computed efficiently in succinct games such as extensive-form games [DFF+25,6PR24,FP23,HvS08,VSF08,PR08]. Motivated by this…

计算机科学与博弈论 · 计算机科学 2025-07-16 Vincent Cheval , Florian Horn , Soumyajit Paul , Mahsa Shirmohammadi

In the context of multi-player, general-sum games, there is an increasing interest in solution concepts modeling some form of communication among players, since they can lead to socially better outcomes with respect to Nash equilibria, and…

计算机科学与博弈论 · 计算机科学 2019-10-15 Andrea Celli , Alberto Marchesi , Tommaso Bianchi , Nicola Gatti

Extensive-form games (EFGs) provide a powerful framework for modeling sequential decision making, capturing strategic interaction under imperfect information, chance events, and temporal structure. Most positive algorithmic and theoretical…

计算机科学与博弈论 · 计算机科学 2026-05-26 Rui Zheng , Ryann Sim , Antonios Varvitsiotis

The existence of simple, uncoupled no-regret dynamics that converge to correlated equilibria in normal-form games is a celebrated result in the theory of multi-agent systems. Specifically, it has been known for more than 20 years that when…

计算机科学与博弈论 · 计算机科学 2022-09-05 Andrea Celli , Alberto Marchesi , Gabriele Farina , Nicola Gatti

We propose a novel online learning method for minimizing regret in large extensive-form games. The approach learns a function approximator online to estimate the regret for choosing a particular action. A no-regret algorithm uses these…

人工智能 · 计算机科学 2015-01-05 Kevin Waugh , Dustin Morrill , J. Andrew Bagnell , Michael Bowling

We study the problem of finding optimal correlated equilibria of various sorts in extensive-form games: normal-form coarse correlated equilibrium (NFCCE), extensive-form coarse correlated equilibrium (EFCCE), and extensive-form correlated…

计算机科学与博弈论 · 计算机科学 2025-01-28 Brian Zhang , Gabriele Farina , Andrea Celli , Tuomas Sandholm

This work investigates the ambient potential identification problem in inverse Mean-Field Games (MFGs), where the goal is to recover the unknown potential from the value function at equilibrium. We propose a simple yet effective iterative…

最优化与控制 · 数学 2025-10-14 Jiajia Yu , Jian-Guo Liu , Hongkai Zhao

Perfect Bayesian Equilibrium (PBE) is a refinement of the Nash equilibrium for imperfect-information extensive-form games (EFGs) that enforces consistency between the two components of a solution: agents' strategy profile describing their…

计算机科学与博弈论 · 计算机科学 2026-02-23 Christine Konicki , Mithun Chakraborty , Michael P. Wellman

Modeling strategic conflict from a game theoretical perspective involves dealing with epistemic uncertainty. Payoff uncertainty models are typically restricted to simple probability models due to computational restrictions. Recent…

计算机科学与博弈论 · 计算机科学 2019-05-13 Juan Leni , John Levine , John Quigley

Fictitious play is an algorithm for computing Nash equilibria of matrix games. Recently, machine learning variants of fictitious play have been successfully applied to complicated real-world games. This paper presents a simple modification…

计算机科学与博弈论 · 计算机科学 2022-12-21 Alex Cloud , Albert Wang , Wesley Kerr

In game theory, imperfect-recall decision problems model situations in which an agent forgets information it held before. They encompass games such as the ``absentminded driver'' and team games with limited communication. In this paper, we…

计算机科学与博弈论 · 计算机科学 2026-02-18 Emanuel Tewolde , Brian Hu Zhang , Ioannis Anagnostides , Tuomas Sandholm , Vincent Conitzer

Blackwell approachability is a framework for reasoning about repeated games with vector-valued payoffs. We introduce predictive Blackwell approachability, where an estimate of the next payoff vector is given, and the decision maker tries to…

计算机科学与博弈论 · 计算机科学 2021-03-09 Gabriele Farina , Christian Kroer , Tuomas Sandholm

While the topic of mean-field games (MFGs) has a relatively long history, heretofore there has been limited work concerning algorithms for the computation of equilibrium control policies. In this paper, we develop a computable policy…

系统与控制 · 电气工程与系统科学 2020-04-07 Muhammad Aneeq uz Zaman , Kaiqing Zhang , Erik Miehling , Tamer Başar

We consider regret minimization in repeated games with non-convex loss functions. Minimizing the standard notion of regret is computationally intractable. Thus, we define a natural notion of regret which permits efficient optimization and…

机器学习 · 计算机科学 2017-11-06 Elad Hazan , Karan Singh , Cyril Zhang

Current approximate Coarse Correlated Equilibria (CCE) algorithms struggle with equilibrium approximation for games in large stochastic environments but are theoretically guaranteed to converge to a strong solution concept. In contrast,…

机器学习 · 计算机科学 2024-12-04 Ryan Yu , Mateusz Nowak , Qintong Xie , Michelle Yilin Feng , Peter Chin

Previous partial permutation synchronization (PPS) algorithms, which are commonly used for multi-object matching, often involve computation-intensive and memory-demanding matrix operations. These operations become intractable for large…

计算机视觉与模式识别 · 计算机科学 2022-04-01 Shaohan Li , Yunpeng Shi , Gilad Lerman