中文
相关论文

相关论文: On Bellman's Optimality Principle for zs-POSGs

200 篇论文

Motivated by uncertain parameters encountered in Markov decision processes (MDPs) and stochastic games, we study the effect of parameter uncertainty on Bellman operator-based algorithms under a set-based framework. Specifically, we first…

计算机科学与博弈论 · 计算机科学 2021-12-14 Sarah H. Q. Li , Assalé , Adjé , Pierre-Loïc Garoche , Behçet Açıkmeşe

The study of convex optimization has historically been concerned with worst-case convergence rates. The development of the Optimized Gradient Method (OGM), due to \citet{drori2012PerformanceOF,Kim2016optimal}, marked a major milestone in…

最优化与控制 · 数学 2026-04-21 Benjamin Grimmer , Kevin Shu , Alex L. Wang

We study offline learning in KL-regularized two-player zero-sum games, where policies are optimized with respect to a fixed reference policy through KL regularization. Prior work relies on pessimistic value estimation to handle distribution…

计算机科学与博弈论 · 计算机科学 2026-05-11 Yuheng Zhang , Claire Chen , Nan Jiang

We consider a two-person trading game in continuous time whereby each player chooses a constant rebalancing rule $b$ that he must adhere to over $[0,t]$. If $V_t(b)$ denotes the final wealth of the rebalancing rule $b$, then Player 1 (the…

投资组合管理 · 定量金融 2022-10-24 Alex Garivaltis

We investigate the optimal reinsurance problem under the criterion of maximizing the expected utility of terminal wealth when the insurance company has restricted information on the loss process. We propose a risk model with claim arrival…

数理金融 · 定量金融 2020-05-15 Matteo Brachetta , Claudia Ceci

We analyse the computational complexity of finding Nash equilibria in turn-based stochastic multiplayer games with omega-regular objectives. We show that restricting the search space to equilibria whose payoffs fall into a certain interval…

计算机科学与博弈论 · 计算机科学 2015-07-01 Michael Ummels , Dominik Wojtczak

We introduce a new unified framework for modelling both decision problems and finite games based on quantifiers and selection functions. We show that the canonical utility maximisation is one special case of a quantifier and that our more…

计算机科学中的逻辑 · 计算机科学 2014-09-29 Jules Hedges , Paulo Oliva , Evguenia Winschel , Viktor Winschel , Philipp Zahn

This article introduces a class of $Nash$ games among $Stackelberg$ players ($NASPs$), namely, a class of simultaneous non-cooperative games where the players solve sequential Stackelberg games. Specifically, each player solves a…

计算机科学与博弈论 · 计算机科学 2025-03-04 Margarida Carvalho , Gabriele Dragotto , Felipe Feijoo , Andrea Lodi , Sriram Sankaranarayanan

We study single-player extensive-form games with imperfect recall, such as the Sleeping Beauty problem or the Absentminded Driver game. For such games, two natural equilibrium concepts have been proposed as alternative solution concepts to…

计算机科学与博弈论 · 计算机科学 2023-05-30 Emanuel Tewolde , Caspar Oesterheld , Vincent Conitzer , Paul W. Goldberg

This paper investigates the convergence time of log-linear learning to an $\epsilon$-efficient Nash equilibrium in potential games, where an efficient Nash equilibrium is defined as the maximizer of the potential function. Previous…

多智能体系统 · 计算机科学 2026-01-13 Anna Maddux , Reda Ouhamma , Maryam Kamgarpour

A zero-sum two-person Perfect Information Semi-Markov game (PISMG) under limiting ratio average payoff has a value and both the maximiser and the minimiser have optimal pure semi-stationary strategies. We arrive at the result by first…

计算机科学与博弈论 · 计算机科学 2023-02-15 S. Sinha , K. G. Bakshi

We examine global non-asymptotic convergence properties of policy gradient methods for multi-agent reinforcement learning (RL) problems in Markov potential games (MPG). To learn a Nash equilibrium of an MPG in which the size of state space…

机器学习 · 计算机科学 2022-08-08 Dongsheng Ding , Chen-Yu Wei , Kaiqing Zhang , Mihailo R. Jovanović

Markov games model interactions among multiple players in a stochastic, dynamic environment. Each player in a Markov game maximizes its expected total discounted reward, which depends upon the policies of the other players. We formulate a…

计算机科学与博弈论 · 计算机科学 2023-09-11 Shenghui Chen , Yue Yu , David Fridovich-Keil , Ufuk Topcu

Congestion games are attractive because they can model many concrete situations where some competing entities interact through the use of some shared resources, and also because they always admit pure Nash equilibria which correspond to the…

计算机科学与博弈论 · 计算机科学 2024-08-22 Vittorio Bilò , Angelo Fanelli , Laurent Gourvès , Christos Tsoufis , Cosimo Vinci

In this paper, we present an optimal control problem for stochastic differential games under Markov regime-switching forward-backward stochastic differential equations with jumps and partial information. First, we prove a sufficient maximum…

最优化与控制 · 数学 2014-10-14 Olivier Menoukeu Pamen , Romual Herve Momeya

In this paper, zero-sum mean-field type games (ZSMFTG) with linear dynamics and quadratic cost are studied under infinite-horizon discounted utility function. ZSMFTG are a class of games in which two decision makers whose utilities sum to…

最优化与控制 · 数学 2020-09-02 René Carmona , Kenza Hamidouche , Mathieu Laurière , Zongjun Tan

We consider the problem of computing mixed Nash equilibria of two-player zero-sum games with continuous sets of pure strategies and with first-order access to the payoff function. This problem arises for example in game-theory-inspired…

最优化与控制 · 数学 2025-09-04 Guillaume Wang , Lénaïc Chizat

We address payoff-based decentralized learning in infinite-horizon zero-sum Markov games. In this setting, each player makes decisions based solely on received rewards, without observing the opponent's strategy or actions nor sharing…

计算机科学与博弈论 · 计算机科学 2025-02-11 Reda Ouhamma , Maryam Kamgarpour

We present a new approach to solving games with a countably or uncountably infinite number of players. Such games are often used to model multiagent systems with a large number of agents. The latter are frequently encountered in economics,…

计算机科学与博弈论 · 计算机科学 2025-01-17 Carlos Martin , Tuomas Sandholm

Modern random access mechanisms combine packet repetitions with multi-user detection mechanisms at the receiver to maximize the throughput and reliability in massive Internet of Things (IoT) scenarios. However, optimizing the access policy,…

‹ 上一页 1 8 9 10 下一页 ›