中文
相关论文

相关论文: Anytime-Constrained Equilibria in Polynomial Time

200 篇论文

We study multiplayer turn-based timed games with reachability objectives. In particular, we are interested in the notion of subgame perfect equilibrium (SPE). We prove that deciding the constrained existence of an SPE in this setting is…

计算机科学与博弈论 · 计算机科学 2020-06-19 Thomas Brihaye , Aline Goeminne

This paper focuses on a class of continuous-time controlled Markov chains with time-inconsistent and distribution-dependent cost functional (in some appropriate sense). A new definition of time-inconsistent distribution-dependent…

最优化与控制 · 数学 2019-09-26 Hongwei Mei , George Yin

Markov games model interactions among multiple players in a stochastic, dynamic environment. Each player in a Markov game maximizes its expected total discounted reward, which depends upon the policies of the other players. We formulate a…

计算机科学与博弈论 · 计算机科学 2023-09-11 Shenghui Chen , Yue Yu , David Fridovich-Keil , Ufuk Topcu

We study a finite-horizon two-person zero-sum risk-sensitive stochastic game for continuous-time Markov chains and Borel state and action spaces, in which payoff rates, transition rates and terminal reward functions are allowed to be…

最优化与控制 · 数学 2021-03-09 Junyu Zhang , Xianping Guo , Li Xia

We study the problem of finding Stackelberg equilibria in games with a massive number of players. So far, the only known game instances in which the problem is solved in polynomial time are some particular congestion games. However, a…

计算机科学与博弈论 · 计算机科学 2019-05-31 Alberto Marchesi , Matteo Castiglioni , Nicola Gatti

We consider the problem of planning with participation constraints introduced in [Zhang et al., 2022]. In this problem, a principal chooses actions in a Markov decision process, resulting in separate utilities for the principal and the…

计算机科学与博弈论 · 计算机科学 2022-05-17 Hanrui Zhang , Yu Cheng , Vincent Conitzer

Infinite-duration games with disturbances extend the classical framework of infinite-duration games, which captures the reactive synthesis problem, with a discrete measure of resilience against non-antagonistic external influence. This…

计算机科学与博弈论 · 计算机科学 2020-07-09 Daniel Neider , Patrick Totzke , Martin Zimmermann

We study the global convergence of policy optimization for finding the Nash equilibria (NE) in zero-sum linear quadratic (LQ) games. To this end, we first investigate the landscape of LQ games, viewing it as a nonconvex-nonconcave…

机器学习 · 计算机科学 2021-02-12 Kaiqing Zhang , Zhuoran Yang , Tamer Başar

A game-theoretic framework for time-inconsistent stopping problems where the time-inconsistency is due to the consideration of a non-linear function of an expected reward is developed. A class of mixed strategy stopping times that allows…

最优化与控制 · 数学 2020-01-23 Sören Christensen , Kristoffer Lindensjö

We study the existence and computation of Nash equilibria in concave games where the players' admissible strategies are subject to shared coupling constraints. Under playerwise concavity of constraints, we prove existence of Nash…

计算机科学与博弈论 · 计算机科学 2026-02-09 Philip Jordan , Maryam Kamgarpour

In optimal stopping problems, a Markov structure guarantees Markovian optimal stopping times (first exit times). Surprisingly, there is no analogous result for Markovian stopping games once randomization is required. This paper addresses…

概率论 · 数学 2024-08-02 Sören Christensen , Boy Schultz

In this work we study of competitive situations among users of a set of global resources. More precisely we study the effect of cost policies used by these resources in the convergence time to a pure Nash equilibrium. The work is divided in…

计算机科学与博弈论 · 计算机科学 2011-03-28 Vissarion Fisikopoulos

In stochastic dynamic environments, team Markov games have emerged as a versatile paradigm for studying sequential decision-making problems of fully cooperative multi-agent systems. However, the optimality of the derived policies is usually…

最优化与控制 · 数学 2022-05-03 Feng Huang , Ming Cao , Long Wang

In game theory, mechanism design is concerned with the design of incentives so that a desired outcome of the game can be achieved. In this paper, we study the design of incentives so that a desirable equilibrium is obtained, for instance,…

计算机科学与博弈论 · 计算机科学 2021-06-21 Julian Gutierrez , Muhammad Najib , Giuseppe Perelli , Michael Wooldridge

We analyse the computational complexity of finding Nash equilibria in turn-based stochastic multiplayer games with omega-regular objectives. We show that restricting the search space to equilibria whose payoffs fall into a certain interval…

计算机科学与博弈论 · 计算机科学 2015-07-01 Michael Ummels , Dominik Wojtczak

We study the problem of finding robust equilibria in multiplayer concurrent games with mean payoff objectives. A $(k,t)$-robust equilibrium is a strategy profile such that no coalition of size $k$ can improve the payoff of one its member by…

计算机科学与博弈论 · 计算机科学 2016-02-02 Romain Brenguier

We formulate two-party policy competition as a two-player non-cooperative game, generalizing Lin et al.'s work (2021). Each party selects a real-valued policy vector as its strategy from a compact subset of Euclidean space, and a voter's…

计算机科学与博弈论 · 计算机科学 2026-03-19 Chuang-Chieh Lin , Chi-Jen Lu , Po-An Chen , Chih-Chieh Hung

One-clock priced timed games is a class of two-player, zero-sum, continuous-time games that was defined and thoroughly studied in previous works. We show that one-clock priced timed games can be solved in time m 12^n n^(O(1)), where n is…

计算机科学与博弈论 · 计算机科学 2013-01-15 Thomas Dueholm Hansen , Rasmus Ibsen-Jensen , Peter Bro Miltersen

There are only limited classes of multi-player stochastic games in which independent learning is guaranteed to converge to a Nash equilibrium. Markov potential games are a key example of such classes. Prior work has outlined sets of…

计算机科学与博弈论 · 计算机科学 2024-05-15 Fatemeh Fardno , Seyed Majid Zahedi

Distributed Nash equilibrium seeking of aggregative games is investigated and a continuous-time algorithm is proposed. The algorithm is designed by virtue of projected gradient play dynamics and distributed average tracking dynamics, and is…

最优化与控制 · 数学 2021-12-07 Shu Liang , Peng Yi , Yiguang Hong , Kaixiang Peng