中文
相关论文

相关论文: Value Functions for Depth-Limited Solving in Zero-…

200 篇论文

Computational equilibrium finding in large zero-sum extensive-form imperfect-information games has led to significant recent AI breakthroughs. The fastest algorithms for the problem are new forms of counterfactual regret minimization [Brown…

计算机科学与博弈论 · 计算机科学 2020-07-01 Brian Hu Zhang , Tuomas Sandholm

This paper uses value functions to characterize the pure-strategy subgame-perfect equilibria of an arbitrary, possibly infinite-horizon game. It specifies the game's extensive form as a pentaform (Streufert 2023p, arXiv:2107.10801v4), which…

理论经济学 · 经济学 2023-03-14 Peter A. Streufert

We propose a generic mechanism for incentivizing behavior in an arbitrary finite game using payments. Doing so is trivial if the mechanism is allowed to observe all actions taken in the game, as this allows it to simply punish those agents…

计算机科学与博弈论 · 计算机科学 2023-04-05 Nikolaj I. Schwartzbach

Regret minimization has proved to be a versatile tool for tree-form sequential decision making and extensive-form games. In large two-player zero-sum imperfect-information games, modern extensions of counterfactual regret minimization (CFR)…

计算机科学与博弈论 · 计算机科学 2021-03-09 Gabriele Farina , Tuomas Sandholm

In imperfect-information games, the optimal strategy in a subgame may depend on the strategy in other, unreached subgames. Thus a subgame cannot be solved in isolation and must instead consider the strategy for the entire game as a whole,…

人工智能 · 计算机科学 2017-11-20 Noam Brown , Tuomas Sandholm

We study zero-sum differential games with state constraints and one-sided information, where the informed player (Player 1) has a categorical payoff type unknown to the uninformed player (Player 2). The goal of Player 1 is to minimize his…

计算机科学与博弈论 · 计算机科学 2024-06-05 Mukesh Ghimire , Lei Zhang , Zhe Xu , Yi Ren

We present a functional framework for automated mechanism design based on a two-stage game model of strategic interaction between the designer and the mechanism participants, and apply it to several classes of two-player infinite games of…

计算机科学与博弈论 · 计算机科学 2012-06-26 Yevgeniy Vorobeychik , Daniel Reeves , Michael P. Wellman

In this paper, we consider the infinite-horizon reach-avoid zero-sum game problem, where the goal is to find a set in the state space, referred to as the reach-avoid set, such that the system starting at a state therein could be controlled…

系统与控制 · 电气工程与系统科学 2024-09-19 Jingqi Li , Donggun Lee , Somayeh Sojoudi , Claire J. Tomlin

Imperfect Information Games (IIGs) offer robust models for scenarios where decision-makers face uncertainty or lack complete information. Counterfactual Regret Minimization (CFR) has been one of the most successful family of algorithms for…

机器学习 · 计算机科学 2025-11-12 Jiayu Chen , Zhekai Wang , Vaneet Aggarwal

No-regret learning has emerged as a powerful tool for solving extensive-form games. This was facilitated by the counterfactual-regret minimization (CFR) framework, which relies on the instantiation of regret minimizers for simplexes at each…

计算机科学与博弈论 · 计算机科学 2017-11-10 Gabriele Farina , Christian Kroer , Tuomas Sandholm

In this work, we propose, for the first time, a reinforcement learning framework specifically designed for zero-sum linear-quadratic stochastic differential games. This approach offers a generalized solution for scenarios in which accurate…

最优化与控制 · 数学 2026-02-10 Yiyuan Wang

We study two-player zero-sum repeated games with incomplete information on one side, where the payoff function is tail measurable (and not necessarily the long-run average payoff). We show that the maxmin value equals the concavification of…

最优化与控制 · 数学 2025-12-02 Gil Bar Castellon Koltun , Ehud Lehrer , Eilon Solan

We introduce a simple extensive-form algorithm for finding equilibria of two-player, zero-sum games. The algorithm is realization equivalent to a generalized form of Fictitious Play. We compare its performance to that of a similar…

计算机科学与博弈论 · 计算机科学 2023-10-17 Tim P. Schulze

In their seminal work, Nayyar et al. (2013) showed that imperfect information can be abstracted away from common-payoff games by having players publicly announce their policies as they play. This insight underpins sound solvers and…

计算机科学与博弈论 · 计算机科学 2023-08-02 Samuel Sokota , Ryan D'Orazio , Chun Kai Ling , David J. Wu , J. Zico Kolter , Noam Brown

A general model for zero-sum stochastic games with asymmetric information is considered. In this model, each player's information at each time can be divided into a common information part and a private information part. Under certain…

系统与控制 · 电气工程与系统科学 2019-12-25 Dhruva Kartik , Ashutosh Nayyar

An imperfect-information game is a type of game with asymmetric information. It is more common in life than perfect-information game. Artificial intelligence (AI) in imperfect-information games, such like poker, has made considerable…

人工智能 · 计算机科学 2024-05-29 Qibin Zhou , Dongdong Bai , Junge Zhang , Fuqing Duan , Kaiqi Huang

We introduce a new approach for computing optimal equilibria via learning in games. It applies to extensive-form settings with any number of players, including mechanism design, information design, and solution concepts such as correlated,…

A common goal in the areas of secure information flow and privacy is to build effective defenses against unwanted leakage of information. To this end, one must be able to reason about potential attacks and their interplay with possible…

密码学与安全 · 计算机科学 2022-04-15 Mário S. Alvim , Konstantinos Chatzikokolakis , Yusuke Kawamoto , Catuscia Palamidessi

In many real-world scenarios, a team of agents coordinate with each other to compete against an opponent. The challenge of solving this type of game is that the team's joint action space grows exponentially with the number of agents, which…

人工智能 · 计算机科学 2021-05-19 Shuxin Li , Youzhi Zhang , Xinrun Wang , Wanqi Xue , Bo An

Counterfactual regret minimization (CFR) is the most popular algorithm on solving two-player zero-sum extensive games with imperfect information and achieves state-of-the-art performance in practice. However, the performance of CFR is not…

机器学习 · 计算机科学 2018-12-27 Yichi Zhou , Tongzheng Ren , Jialian Li , Dong Yan , Jun Zhu