中文
相关论文

相关论文: Strategy Iteration using Non-Deterministic Strateg…

200 篇论文

We study the limiting behavior of the mixed strategies that result from optimal no-regret learning strategies in a repeated game setting where the stage game is any 2 by 2 competitive game. We consider optimal no-regret algorithms that are…

计算机科学与博弈论 · 计算机科学 2022-03-03 Vidya Muthukumar , Soham Phade , Anant Sahai

Priced timed games (PTGs) are two-player zero-sum games played on the infinite graph of configurations of priced timed automata where two players take turns to choose transitions in order to optimize cost to reach target states. Bouyer et…

计算机科学与博弈论 · 计算机科学 2020-02-18 Thomas Brihaye , Gilles Geeraerts , Shankara Narayanan Krishna , Lakshmi Manasa , Benjamin Monmege , Ashutosh Trivedi

The iterated prisoner's dilemma is a game that produces many counter-intuitive and complex behaviors in a social environment, based on very simple basic rules. It illustrates that cooperation can be a good thing even in a competitive world,…

计算机科学与博弈论 · 计算机科学 2020-09-07 Robert Prentner

The Kelly or proportional allocation mechanism is a simple and efficient auction-based scheme that distributes an infinitely divisible resource proportionally to the agents bids. When agents are aware of the allocation rule, their…

计算机科学与博弈论 · 计算机科学 2026-03-27 Younes Ben Mazziane , Cleque-Marlain Mboulou Moutoubi , Eitan Altman , Francesco De Pellegrini

We study a sequence of independent one-shot non-cooperative games where agents play equilibria determined by a tunable mechanism. Observing only equilibrium decisions, without parametric or distributional knowledge of utilities, we aim to…

计算机科学与博弈论 · 计算机科学 2025-11-10 Luke Snow , Vikram Krishnamurthy

We study a sequential decision-making model where a set of items is repeatedly matched to the same set of agents over multiple rounds. The objective is to determine a sequence of matchings that either maximizes the utility of the least…

计算机科学与博弈论 · 计算机科学 2025-10-07 Eugene Lim , Tzeh Yuan Neoh , Nicholas Teh

In communication systems where users share common resources, users' selfish behavior usually results in suboptimal resource utilization. There have been extensive works that model communication systems with selfish users as one-shot games…

信息论 · 计算机科学 2011-11-11 Yuanzhang Xiao , Jaeok Park , Mihaela van der Schaar

Infinite games where several players seek to coordinate under imperfect information are deemed to be undecidable, unless the information is hierarchically ordered among the players. We identify a class of games for which joint winning…

计算机科学与博弈论 · 计算机科学 2015-07-29 Dietmar Berwanger , Anup Basil Mathew

We consider a setting in which a principal gets to choose which game from some given set is played by a group of agents. The principal would like to choose a game that favors one of the players, the social preferences of the players, or the…

计算机科学与博弈论 · 计算机科学 2025-11-27 Caspar Oesterheld , Vincent Conitzer

The performance of two pivoting algorithms, due to Lemke and Cottle and Dantzig, is studied on linear complementarity problems (LCPs) that arise from infinite games, such as parity, average-reward, and discounted games. The algorithms have…

计算机科学与博弈论 · 计算机科学 2020-01-16 John Fearnley , Marcin Jurdziński , Rahul Savani

Progress-measure lifting algorithms for solving parity games have the best worst-case asymptotic runtime, but are limited by their asymmetric nature, and known from the work of Czerwi\'nski et al. (2018) to be subject to a matching…

计算机科学中的逻辑 · 计算机科学 2020-10-19 Marcin Jurdziński , Rémi Morvan , Pierre Ohlmann , K. S. Thejaswini

Evolutionary game theory provides a mathematical foundation for cross-disciplinary fertilization, especially for integrating ideas from artificial intelligence and game theory. Such integration offers a transparent and rigorous approach to…

物理与社会 · 物理学 2023-05-31 Xingru Chen , Feng Fu

We consider game theory from the perspective of quantum algorithms. Strategies in classical game theory are either pure (deterministic) or mixed (probabilistic). We introduce these basic ideas in the context of a simple example, closely…

量子物理 · 物理学 2009-10-31 David A. Meyer

A formula is presented for designing zero-determinant(ZD) strategies of general finite games, which have $n$ players and players can have different numbers of strategies. To this end, using semi-tensor product (STP) of matrices, the profile…

计算机科学与博弈论 · 计算机科学 2021-07-08 Daizhan Cheng

We show that the problem of checking if a given nondeterministic parity automaton simulates another given nondeterministic parity automaton is NP-hard. We then adapt the techniques used for this result to show that the problem of checking…

形式语言与自动机理论 · 计算机科学 2026-05-28 Keya Prakash

Simple stochastic games are turn-based 2.5-player zero-sum graph games with a reachability objective. The problem is to compute the winning probability as well as the optimal strategies of both players. In this paper, we compare the three…

计算机科学与博弈论 · 计算机科学 2020-09-24 Jan Křetínský , Emanuel Ramneantu , Alexander Slivinskiy , Maximilian Weininger

The process of revising (or constructing) a policy at execution time -- known as decision-time planning -- has been key to achieving superhuman performance in perfect-information games like chess and Go. A recent line of work has extended…

人工智能 · 计算机科学 2024-05-14 Samuel Sokota , Gabriele Farina , David J. Wu , Hengyuan Hu , Kevin A. Wang , J. Zico Kolter , Noam Brown

Markovian processes have long been used to model stochastic environments. Reinforcement learning has emerged as a framework to solve sequential planning and decision-making problems in such environments. In recent years, attempts were made…

人工智能 · 计算机科学 2014-01-17 Mahdi Milani Fard , Joelle Pineau

We consider a model for repeated stochastic matching where compatibility is probabilistic, is realized the first time agents are matched, and persists in the future. Such a model has applications in the gig economy, kidney exchange, and…

计算机科学与博弈论 · 计算机科学 2021-06-15 Mobin Y. Jeloudar , Irene Lo , Tristan Pollner , Amin Saberi

In the context of multiplayer games, the parallel repetition problem can be phrased as follows: given a game $G$ with optimal winning probability $1-\alpha$ and its repeated version $G^n$ (in which $n$ games are played together, in…

量子物理 · 物理学 2025-06-09 Rotem Arnon , Renato Renner , Thomas Vidick