中文
相关论文

相关论文: The Value of Recall in Extensive-Form Games

200 篇论文

A valuation for a player in a game in extensive form is an assignment of numeric values to the players moves. The valuation reflects the desirability moves. We assume a myopic player, who chooses a move with the highest valuation.…

机器学习 · 计算机科学 2007-05-23 Philippe Jehiel , Dov Samet

We investigate optimal decision making under imperfect recall, that is, when an agent forgets information it once held before. An example is the absentminded driver game, as well as team games in which the members have limited communication…

计算机科学与博弈论 · 计算机科学 2024-06-25 Emanuel Tewolde , Brian Hu Zhang , Caspar Oesterheld , Manolis Zampetakis , Tuomas Sandholm , Paul W. Goldberg , Vincent Conitzer

Extensive-form games (EFGs) model finite sequential interactions between players. The amount of memory required to represent these games is the main bottleneck of algorithms for computing optimal strategies and the size of these strategies…

计算机科学与博弈论 · 计算机科学 2020-04-16 Jiri Cermak , Viliam Lisy , Branislav Bosansky

Counterfactual Regret Minimization (CFR) is an efficient no-regret learning algorithm for decision problems modeled as extensive games. CFR's regret bounds depend on the requirement of perfect recall: players always remember information…

计算机科学与博弈论 · 计算机科学 2012-05-04 Marc Lanctot , Richard Gibson , Neil Burch , Martin Zinkevich , Michael Bowling

In game theory, imperfect-recall decision problems model situations in which an agent forgets information it held before. They encompass games such as the ``absentminded driver'' and team games with limited communication. In this paper, we…

计算机科学与博弈论 · 计算机科学 2026-02-18 Emanuel Tewolde , Brian Hu Zhang , Ioannis Anagnostides , Tuomas Sandholm , Vincent Conitzer

Extensive-form games constitute the standard representation scheme for games with a temporal component. But do all extensive-form games correspond to protocols that we can implement in the real world? We often rule out games with imperfect…

计算机科学与博弈论 · 计算机科学 2015-02-12 Sune K. Jakobsen , Troels B. Sørensen , Vincent Conitzer

Imperfect recall games represent dynamic interactions where players forget previously known information, such as a history of played actions. The importance of imperfect recall games stems from allowing a concise representation of…

计算机科学与博弈论 · 计算机科学 2017-05-25 Jiri Cermak , Branislav Bosansky , Michal Pechoucek

The paper proposes a natural measure space of zero-sum perfect information games with upper semicontinuous payoffs. Each game is specified by the game tree, and by the assignment of the active player and of the capacity to each node of the…

计算机科学与博弈论 · 计算机科学 2021-04-22 János Flesch , Arkadi Predtetchinski , Ville Suomala

Extensive-form games (EFGs) provide a powerful framework for modeling sequential decision making, capturing strategic interaction under imperfect information, chance events, and temporal structure. Most positive algorithmic and theoretical…

计算机科学与博弈论 · 计算机科学 2026-05-26 Rui Zheng , Ryann Sim , Antonios Varvitsiotis

Extensive-form games with imperfect recall are an important game-theoretic model that allows a compact representation of strategies in dynamic strategic interactions. Practical use of imperfect recall games is limited due to negative…

计算机科学与博弈论 · 计算机科学 2017-05-25 Branislav Bosansky , Jiri Cermak , Karel Horak , Michal Pechoucek

We consider a two-player zero-sum game with integral payoff and with incomplete information on one side, where the payoff is chosen among a continuous set of possible payoffs. We prove that the value function of this game is solution of an…

概率论 · 数学 2012-02-23 Pierre Cardaliaguet , Catherine Rainer

Experience replay enables off-policy reinforcement learning (RL) agents to utilize past experiences to maximize the cumulative reward. Prioritized experience replay that weighs experiences by the magnitude of their temporal-difference error…

机器学习 · 计算机科学 2021-02-08 Ang A. Li , Zongqing Lu , Chenglin Miao

In this paper, we establish efficient and uncoupled learning dynamics so that, when employed by all players in multiplayer perfect-recall imperfect-information extensive-form games, the trigger regret of each player grows as $O(\log T)$…

计算机科学与博弈论 · 计算机科学 2023-09-20 Ioannis Anagnostides , Gabriele Farina , Tuomas Sandholm

In games with imperfect recall, players may forget the sequence of decisions they made in the past. When players also forget whether they have already encountered their current decision point, they are said to be absent-minded. Solving…

计算机科学与博弈论 · 计算机科学 2025-02-20 Hugo Gimbert , Soumyajit Paul , B. Srivathsan

We study single-player extensive-form games with imperfect recall, such as the Sleeping Beauty problem or the Absentminded Driver game. For such games, two natural equilibrium concepts have been proposed as alternative solution concepts to…

计算机科学与博弈论 · 计算机科学 2023-05-30 Emanuel Tewolde , Caspar Oesterheld , Vincent Conitzer , Paul W. Goldberg

We study continuity properties of stochastic game problems with respect to various topologies on information structures, defined as probability measures characterizing a game. We will establish continuity properties of the value function…

最优化与控制 · 数学 2022-11-02 Ian Hogeboom-Burr , Serdar Yüksel

We provide a formal definition of depth-limited games together with an accessible and rigorous explanation of the underlying concepts, both of which were previously missing in imperfect-information games. The definition works for an…

人工智能 · 计算机科学 2022-03-25 Vojtěch Kovařík , Dominik Seitz , Viliam Lisý , Jan Rudolf , Shuo Sun , Karel Ha

We consider repeated zero-sum games with incomplete information on the side of Player 2 with the total payoff given by the non-normalized sum of stage gains. In the classical examples the value $V_N$ of such an $N$-stage game is of the…

计算机科学与博弈论 · 计算机科学 2016-09-14 Fedor Sandomirskiy

We extend Kuhn's Theorem to games of the extensive form with unawareness. We prove that if a game of the extensive form with unawareness has perfect recall, then for each mixed strategy there is an equivalent behavior strategy. We show that…

理论经济学 · 经济学 2026-04-17 Ki Vin Foo , Burkhard C. Schipper

In some games, additional information hurts a player, e.g., in games with first-mover advantage, the second-mover is hurt by seeing the first-mover's move. What properties of a game determine whether it has such negative "value of…

计算机科学与博弈论 · 计算机科学 2015-02-02 Nils Bertschinger , David H. Wolpert , Eckehard Olbrich , Juergen Jost
‹ 上一页 1 2 3 10 下一页 ›