中文
相关论文

相关论文: Using Strategy Improvement to Stay Alive

200 篇论文

Synthesis of bulletproof strategies in imperfect information scenarios is a notoriously hard problem. In this paper, we suggest that it is sometimes a viable alternative to aim at "reasonably good" strategies instead. This makes sense not…

多智能体系统 · 计算机科学 2023-10-26 Wojciech Jamroga , Damian Kurpiewski

This work investigates the ambient potential identification problem in inverse Mean-Field Games (MFGs), where the goal is to recover the unknown potential from the value function at equilibrium. We propose a simple yet effective iterative…

最优化与控制 · 数学 2025-10-14 Jiajia Yu , Jian-Guo Liu , Hongkai Zhao

Parity games are abstract infinite-round games that take an important role in formal verification. In the basic setting, these games are two-player, turn-based, and played under perfect information on directed graphs, whose nodes are…

计算机科学与博弈论 · 计算机科学 2019-10-31 Antonio Di Stasio , Aniello Murano , Giuseppe Perelli , Moshe Y. Vardi

We consider multiple-environment Markov decision processes (MEMDP), which consist of a finite set of MDPs over the same state space, representing different scenarios of transition structure and probability. The value of a strategy is the…

计算机科学中的逻辑 · 计算机科学 2025-04-23 Krishnendu Chatterjee , Laurent Doyen , Jean-François Raskin , Ocan Sankur

In common-interest stochastic games all players receive an identical payoff. Players participating in such games must learn to coordinate with each other in order to receive the highest-possible value. A number of reinforcement learning…

人工智能 · 计算机科学 2011-06-28 R. I. Brafman , M. Tennenholtz

Motivated by recent developments in mean-field games in ecology, in this paper we introduce a connection between the best response dynamics in evolutionary game theory, the minimization of the highest income of a game, and minimizing…

最优化与控制 · 数学 2024-11-13 Dante Kalise , Alessio Oliviero , Domènec Ruiz-Balet

Autonomous drone racing pushes the boundaries of high-speed motion planning and multi-agent strategic decision-making. Success in this domain requires drones not only to navigate at their limits but also to anticipate and counteract…

机器人学 · 计算机科学 2026-02-09 Andrei-Carlo Papuc , Lasse Peters , Sihao Sun , Laura Ferranti , Javier Alonso-Mora

We present a new tool for the study of multiplayer stochastic games, namely the modified game, which is a normal-form game that depends on the discount factor, the initial state, and for every player a partition of the set of states and a…

概率论 · 数学 2017-03-14 Eilon Solan

Bargaining can be used to resolve mixed-motive games in multi-agent systems. Although there is an abundance of negotiation strategies implemented in automated negotiating agents, most agents are based on single fixed strategies, while it is…

多智能体系统 · 计算机科学 2022-12-21 Bram M. Renting , Holger H. Hoos , Catholijn M. Jonker

We study \emph{partial-information} two-player turn-based games on graphs with omega-regular objectives, when the partial-information player has \emph{limited memory}. Such games are a natural formalization for reactive synthesis when the…

形式语言与自动机理论 · 计算机科学 2020-02-19 Dhananjay Raju , Rüdiger Ehlers , Ufuk Topcu

Markov decision processes (MDPs) are widely used in modeling decision making problems in stochastic environments. However, precise specification of the reward functions in MDPs is often very difficult. Recent approaches have focused on…

人工智能 · 计算机科学 2012-02-20 Eunsoo Oh , Kee-Eung Kim

We introduce a mean field game with rank-based reward: competing agents optimize their effort to achieve a goal, are ranked according to their completion time, and paid a reward based on their relative rank. First, we propose a tractable…

最优化与控制 · 数学 2017-08-07 Marcel Nutz , Yuchong Zhang

We study the problem of repeated play in a zero-sum game in which the payoff matrix may change, in a possibly adversarial fashion, on each round; we call these Online Matrix Games. Finding the Nash Equilibrium (NE) of a two player zero-sum…

机器学习 · 计算机科学 2020-04-06 Adrian Rivera Cardoso , Jacob Abernethy , He Wang , Huan Xu

Entropy regularization has been extensively adopted to improve the efficiency, the stability, and the convergence of algorithms in reinforcement learning. This paper analyzes both quantitatively and qualitatively the impact of entropy…

最优化与控制 · 数学 2021-12-10 Xin Guo , Renyuan Xu , Thaleia Zariphopoulou

Multi-agent robust reinforcement learning, also known as multi-player robust Markov games (RMGs), is a crucial framework for modeling competitive interactions under environmental uncertainties, with wide applications in multi-agent systems.…

机器学习 · 计算机科学 2024-12-31 Yuchen Jiao , Gen Li

Value iteration is a well-known method of solving Markov Decision Processes (MDPs) that is simple to implement and boasts strong theoretical convergence guarantees. However, the computational cost of value iteration quickly becomes…

机器学习 · 计算机科学 2021-07-26 Guanting Chen , Johann Demetrio Gaebler , Matt Peng , Chunlin Sun , Yinyu Ye

We suggest a new algorithm for two-person zero-sum undiscounted stochastic games focusing on stationary strategies. Given a positive real $\epsilon$, let us call a stochastic game $\epsilon$-ergodic, if its values from any two initial…

计算机科学与博弈论 · 计算机科学 2015-08-17 Endre Boros , Khaled Elbassioni , Vladimir Gurvich , Kazuhisa Makino

When reasoning about the strategic capabilities of an agent, it is important to consider the nature of its adversaries. In the particular context of controller synthesis for quantitative specifications, the usual problem is to devise a…

计算机科学与博弈论 · 计算机科学 2014-04-04 Véronique Bruyère , Emmanuel Filiot , Mickael Randour , Jean-François Raskin

Each of two players, by turns, rolls a dice several times accumulating the successive scores until he decides to stop, or he rolls an ace. When stopping, the accumulated turn score is added to the player account and the dice is given to his…

概率论 · 数学 2009-12-31 Fabian Crocce , Ernesto Mordecki

Financial markets are often driven by latent factors which traders cannot observe. Here, we address an algorithmic trading problem with collections of heterogeneous agents who aim to perform optimal execution or statistical arbitrage, where…

数理金融 · 定量金融 2019-04-02 Philippe Casgrain , Sebastian Jaimungal
‹ 上一页 1 8 9 10 下一页 ›