中文
相关论文

相关论文: Regularity of the minmax value and equilibria in m…

200 篇论文

We use techniques from the statistical mechanics of disordered systems to analyse the properties of Nash equilibria of bimatrix games with large random payoff matrices. By means of an annealed bound, we calculate their number and analyse…

无序系统与神经网络 · 物理学 2009-10-31 Johannes Berg , Martin Weigt

In the early 1950s Lloyd Shapley proposed an ordinal and set-valued solution concept for zero-sum games called \emph{weak saddle}. We show that all weak saddles of a given zero-sum game are interchangeable and equivalent. As a consequence,…

计算机科学与博弈论 · 计算机科学 2018-04-19 Felix Brandt , Markus Brill , Warut Suksompong

We study infinite two-player games where one of the players is unsure about the set of moves available to the other player. In particular, the set of moves of the other player is a strict superset of what she assumes it to be. We explore…

计算机科学与博弈论 · 计算机科学 2013-03-05 Nicholas Asher , Soumya Paul

Game-theoretic upper expectations are joint (global) probability models that mathematically describe the behaviour of uncertain processes in terms of supermartingales; capital processes corresponding to available betting strategies.…

概率论 · 数学 2021-07-14 Natan T'Joens , Jasper De Bock , Gert de Cooman

Box-simplex games are a family of bilinear minimax objectives which encapsulate graph-structured problems such as maximum flow [She17], optimal transport [JST19], and bipartite matching [AJJ+22]. We develop efficient near-linear time,…

数据结构与算法 · 计算机科学 2022-06-15 Arun Jambulapati , Yujia Jin , Aaron Sidford , Kevin Tian

Consider a set of agents who play a network game repeatedly. Agents may not know the network. They may even be unaware that they are interacting with other agents in a network. Possibly, they just understand that their payoffs depend on an…

理论经济学 · 经济学 2022-07-26 Pierpaolo Battigalli , Fabrizio Panebianco , Paolo Pin

We examine a patient player's behavior when he can build reputations in front of a sequence of myopic opponents. With positive probability, the patient player is a commitment type who plays his Stackelberg action in every period. We…

理论经济学 · 经济学 2021-02-11 Yingkai Li , Harry Pei

We consider N-player and mean field games in continuous time over a finite horizon, where the position of each agent belongs to {-1,1}. If there is uniqueness of mean field game solutions, e.g. under monotonicity assumptions, then the…

最优化与控制 · 数学 2019-02-06 Alekos Cecchin , Paolo Dai Pra , Markus Fischer , Guglielmo Pelino

Two-player games on graphs is central in many problems in formal verification and program analysis such as synthesis and verification of open systems. In this work we consider solving recursive game graphs (or pushdown game graphs) that can…

计算机科学中的逻辑 · 计算机科学 2016-05-17 Krishnendu Chatterjee , Yaron Velner

The game of Cops and Robber is traditionally played on a finite graph. The purpose of this note is to introduce and analyze the game that is played on an arbitrary geodesic space. The game is defined in such a way that it preserves the…

组合数学 · 数学 2021-12-07 Bojan Mohar

The paper is concerned with two-person games with saddle point. We investigate the limits of value functions for long-time-average payoff, discounted average payoff, and the payoff that follows a probability density. Most of our assumptions…

最优化与控制 · 数学 2015-01-29 Dmitry Khlopin

We introduce and study incentive equilibria for multi-player meanpayoff games. Incentive equilibria generalise well-studied solution concepts such as Nash equilibria and leader equilibria (also known as Stackelberg equilibria). Recall that…

计算机科学与博弈论 · 计算机科学 2015-11-03 Anshul Gupta , M. S. Krishna Deepak , Bharath Kumar Padarthi , Sven Schewe , Ashutosh Trivedi

We investigate a class of reinforcement learning dynamics where players adjust their strategies based on their actions' cumulative payoffs over time - specifically, by playing mixed strategies that maximize their expected cumulative payoff…

最优化与控制 · 数学 2016-02-10 Panayotis Mertikopoulos , William H. Sandholm

We study graphs and two-player games in which rewards are assigned to states, and the goal of the players is to satisfy or dissatisfy certain property of the generated outcome, given as a mean payoff property. Since the notion of…

计算机科学中的逻辑 · 计算机科学 2016-04-22 Tomáš Brázdil , Vojtěch Forejt , Antonín Kučera , Petr Novotný

A key challenge of evolutionary game theory and multi-agent learning is to characterize the limit behavior of game dynamics. Whereas convergence is often a property of learning algorithms in games satisfying a particular reward structure…

计算机科学与博弈论 · 计算机科学 2022-02-09 Aleksander Czechowski , Georgios Piliouras

Maximal (also miner) extractable value, or MEV, usually refers to the value that privileged players can extract by strategically ordering, censoring, and placing transactions in a blockchain. Each blockchain network, which we refer to as a…

计算机科学与博弈论 · 计算机科学 2022-08-30 Bruno Mazorra , Michael Reynolds , Vanesa Daza

Nash equilibrium} (NE) can be stated as a formal theorem on a multilinear form, free of game theory terminology. On the other hand, inspired by this formalism, we state and prove a {\it multilinear minimax theorem}, a generalization of von…

计算机科学与博弈论 · 计算机科学 2024-01-01 Bahman Kalantari

Recent applications that arise in machine learning have surged significant interest in solving min-max saddle point games. This problem has been extensively studied in the convex-concave regime for which a global equilibrium solution can be…

最优化与控制 · 数学 2019-11-01 Maher Nouiehed , Maziar Sanjabi , Tianjian Huang , Jason D. Lee , Meisam Razaviyayn

In increasingly different contexts, it happens that a human player has to interact with artificial players who make decisions following decision-making algorithms. How should the human player play against these algorithms to maximize his…

计算机科学与博弈论 · 计算机科学 2022-02-22 Maurizio D 'Andrea

We design and analyze minimax-optimal algorithms for online linear optimization games where the player's choice is unconstrained. The player strives to minimize regret, the difference between his loss and the loss of a post-hoc benchmark…

机器学习 · 计算机科学 2013-02-12 H. Brendan McMahan