中文
相关论文

相关论文: An optimal MOO strategy

200 篇论文

In the Penney-Ante game, Player I chooses a head/tail string of a predetermined length $n\ge3$. Player II, upon seeing Player I's choice, chooses another head/tail string of the same length. A coin is then tossed repeatedly and the player…

组合数学 · 数学 2021-07-16 Reed Phillips , A. J. Hildebrand

In two-player games on graphs, the players move a token through a graph to produce an infinite path, which determines the winner of the game. Such games are central in formal methods since they model the interaction between a…

计算机科学与博弈论 · 计算机科学 2023-06-22 Milad Aghajohari , Guy Avni , Thomas A. Henzinger

In the context of strategic games, we provide an axiomatic proof of the statement Common knowledge of rationality implies that the players will choose only strategies that survive the iterated elimination of strictly dominated strategies.…

计算机科学与博弈论 · 计算机科学 2010-06-28 Jonathan A. Zvesper , Krzysztof R. Apt

Matchmaking connects multiple players to participate in online player-versus-player games. Current matchmaking systems depend on a single core strategy: create fair games at all times. These systems pair similarly skilled players on the…

社会与信息网络 · 计算机科学 2018-06-27 Zhengxing Chen , Su Xue , John Kolen , Navid Aghdaie , Kazi A. Zaman , Yizhou Sun , Magy Seif El-Nasr

We present Self-Play Preference Optimization (SPO), an algorithm for reinforcement learning from human feedback. Our approach is minimalist in that it does not require training a reward model nor unstable adversarial training and is…

机器学习 · 计算机科学 2024-06-14 Gokul Swamy , Christoph Dann , Rahul Kidambi , Zhiwei Steven Wu , Alekh Agarwal

This paper investigates a class of games with large strategy spaces, motivated by challenges in AI alignment and language games. We introduce the hidden game problem, where for each player, an unknown subset of strategies consistently…

人工智能 · 计算机科学 2025-10-07 Gon Buzaglo , Noah Golowich , Elad Hazan

Repeated games have provided an explanation how mutual cooperation can be achieved even if defection is more favorable in a one-shot game in prisoner's dilemma situation. Recently found zero-determinant strategies have substantially been…

计算机科学与博弈论 · 计算机科学 2021-05-27 Masahiko Ueda

We consider a card guessing game with complete feedback. An ordered deck of $n$ cards labeled $1$ up to $n$ is shelf-shuffled exactly one time. One after the other a single card is drawn from the shuffled deck. The guesser makes has guess…

组合数学 · 数学 2026-02-24 Markus Kuba

Consider a two-player game repeated N times. Player 1 can choose between two styles (for interpretability, offensive and defensive), whereas Player 2 uses a single fixed style. Let X N\,:= \#wins -\#losses for Player 1 after N games, and…

计算机科学与博弈论 · 计算机科学 2026-04-20 Jonatha ANSELMI , Bruno Gaujal

We consider concurrent games played by two-players on a finite-state graph, where in every round the players simultaneously choose a move, and the current state along with the joint moves determine the successor state. We study a…

计算机科学与博弈论 · 计算机科学 2014-09-19 Krishnendu Chatterjee , Rasmus Ibsen-Jensen

We define a class of zero-sum games with combinatorial structure, where the best response problem of one player is to maximize a submodular function. For example, this class includes security games played on networks, as well as the problem…

计算机科学与博弈论 · 计算机科学 2017-12-04 Bryan Wilder

We consider the problem of maximizing the probability of hitting a strategically chosen hidden virtual network by placing a wiretap on a single link of a communication network. This can be seen as a two-player win-lose (zero-sum) game that…

计算机科学与博弈论 · 计算机科学 2009-10-04 Haris Aziz , Oded Lachish , Mike Paterson , Rahul Savani

Weighted timed games are two-player zero-sum games played in a timed automaton equipped with integer weights. We consider optimal reachability objectives, in which one of the players, that we call Min, wants to reach a target location while…

计算机科学与博弈论 · 计算机科学 2025-03-05 Benjamin Monmege , Julie Parreaux , Pierre-Alain Reynier

An unknown positive number of items arrive at independent uniformly distributed times in the interval [0,1] to a selector, whose task is to pick online the last one. We show that under the assumption of an adversary determining the number…

计算机科学与博弈论 · 计算机科学 2011-04-18 Johan Wästlund

We define memory-efficient certificates for $\mu$-calculus model checking problems based on the well-known correspondence of the $\mu$-calculus model checking with winning certain parity games. Winning strategies can independently checked,…

计算机科学中的逻辑 · 计算机科学 2014-01-09 Martin Hofmann , Harald Ruess

In this article, we prove the completeness of the following game search algorithms: unbounded best-first minimax with completion and descent with completion, i.e. we show that, with enough time, they find the best game strategy. We then…

计算机科学与博弈论 · 计算机科学 2021-09-21 Quentin Cohen-Solal

In numerous positional games the identity of the winner is easily determined. In this case one of the more interesting questions is not {\em who} wins but rather {\em how fast} can one win. These type of problems were studied earlier for…

组合数学 · 数学 2008-06-03 Dan Hefetz , Michael Krivelevich , Miloš Stojaković , Tibor Szabó

We apply several quantization schemes to simple versions of the Chinos game. Classically, for two players with one coin each, there is a symmetric stable strategy that allows each player to win half of the times on average. A partial…

量子物理 · 物理学 2009-11-07 F. Guinea , M. A. Martin-Delgado

We show that there is an $m=2n+o(n)$, such that, in the Maker-Breaker game played on $\Z^d$ where Maker needs to put at least $m$ of his marks consecutively in one of $n$ given winning directions, Breaker can force a draw using a pairing…

组合数学 · 数学 2010-06-01 Padmini Mukkamala , Dömötör Pálvölgyi

Adversarial training, a special case of multi-objective optimization, is an increasingly prevalent machine learning technique: some of its most notable applications include GAN-based generative modeling and self-play techniques in…