中文
相关论文

相关论文: Policy Optimization for Linear-Quadratic Zero-Sum …

200 篇论文

This paper studies a large population dynamic game involving nonlinear stochastic dynamical systems with agents of the following mixed types: (i) a major agent, and (ii) a population of $N$ minor agents where $N$ is very large. The major…

最优化与控制 · 数学 2013-06-07 Mojtaba Nourian , Peter E. Caines

Stochastic games are an important class of problems that generalize Markov decision processes to game theoretic scenarios. We consider finite state two-player zero-sum stochastic games over an infinite time horizon with discounted rewards.…

最优化与控制 · 数学 2008-06-17 Parikshit Shah , Pablo A. Parrilo

We consider a class of mean field games in which the agents interact through both their states and controls, and we focus on situations in which a generic agent tries to adjust her speed (control) to an average speed (the average is made in…

偏微分方程分析 · 数学 2020-03-10 Y Achdou , Z Kobeissi

This paper investigates value function approximation in the context of zero-sum Markov games, which can be viewed as a generalization of the Markov decision process (MDP) framework to the two-agent case. We generalize error bounds from MDPs…

人工智能 · 计算机科学 2013-01-07 Michail Lagoudakis , Ron Parr

We study policy optimization in Stackelberg mean field games (MFGs), a hierarchical framework for modeling the strategic interaction between a single leader and an infinitely large population of homogeneous followers. The objective can be…

机器学习 · 计算机科学 2025-11-27 Sihan Zeng , Benjamin Patrick Evans , Sujay Bhatt , Leo Ardon , Sumitra Ganesh , Alec Koppel

This paper studies an optimal investment-consumption problem for competitive agents with exponential or power utilities and a common finite time horizon. Each agent regards the average of habit formation and wealth from all peers as…

最优化与控制 · 数学 2024-05-06 Zongxia Liang , Keyu Zhang

This paper studies policy optimization algorithms for multi-agent reinforcement learning. We begin by proposing an algorithm framework for two-player zero-sum Markov Games in the full-information setting, where each iteration consists of a…

机器学习 · 计算机科学 2022-07-26 Runyu Zhang , Qinghua Liu , Huan Wang , Caiming Xiong , Na Li , Yu Bai

Competitive games involving thousands or even millions of players are prevalent in real-world contexts, such as transportation, communications, and computer networks. However, learning in these large-scale multi-agent environments presents…

最优化与控制 · 数学 2025-02-04 Batuhan Yardim , Semih Cayci , Niao He

In this paper, we consider a finite horizon, non-stationary, mean field games (MFG) with a large population of homogeneous players, sequentially making strategic decisions, where each player is affected by other players through an aggregate…

系统与控制 · 电气工程与系统科学 2020-04-07 Rajesh K Mishra , Deepanshu Vasal , Sriram Vishwanath

This paper proposes a new mathematical paradigm to analyze discrete-time mean-field games. It is shown that finding Nash equilibrium solutions for a general class of discrete-time mean-field games is equivalent to solving an optimization…

最优化与控制 · 数学 2023-08-29 Xin Guo , Anran Hu , Junzi Zhang

Nonzero-sum stochastic differential games with impulse controls offer a realistic and far-reaching modelling framework for applications within finance, energy markets, and other areas, but the difficulty in solving such problems has…

数值分析 · 数学 2020-06-29 Diego Zabaljauregui

Zero-sum stochastic games generalize the notion of Markov Decision Processes (i.e. controlled Markov chains, or stochastic dynamic programming) to the 2-player competitive case : two players jointly control the evolution of a state…

最优化与控制 · 数学 2019-05-17 Jérôme Renault

We consider stochastic differential games with $N$ nearly identical players, linear-Gaussian dynamics, and infinite horizon discounted quadratic cost. Admissible controls are feedbacks for which the system is ergodic. We first study the…

偏微分方程分析 · 数学 2014-03-18 Fabio S. Priuli

Mean-field games (MFGs) study the Nash equilibrium of systems with a continuum of interacting agents, which can be formulated as the fixed-point of optimal control problems. They provide a unified framework for a variety of applications,…

机器学习 · 统计学 2025-12-02 Jiajia Yu , Junghwan Lee , Yao Xie , Xiuyuan Cheng

Semi-Markov model is one of the most general models for stochastic dynamic systems. This paper deals with a two-person zero-sum game for semi-Markov processes. We focus on the expected discounted payoff criterion with state-action-dependent…

计算机科学与博弈论 · 计算机科学 2021-03-09 Zhihui Yu , Xianping Guo , Li Xia

The recent mean field game (MFG) formalism facilitates otherwise intractable computation of approximate Nash equilibria in many-agent settings. In this paper, we consider discrete-time finite MFGs subject to finite-horizon objectives. We…

多智能体系统 · 计算机科学 2022-07-11 Kai Cui , Heinz Koeppl

We provide an in-depth study of Nash equilibria in multi-objective normal form games (MONFGs), i.e., normal form games with vectorial payoffs. Taking a utility-based approach, we assume that each player's utility can be modelled with a…

计算机科学与博弈论 · 计算机科学 2022-07-19 Willem Röpke , Diederik M. Roijers , Ann Nowé , Roxana Rădulescu

Stochastic games generalize Markov decision processes (MDPs) to a multiagent setting by allowing the state transitions to depend jointly on all player actions, and having rewards determined by multiplayer matrix games at each state. We…

计算机科学与博弈论 · 计算机科学 2013-01-18 Michael Kearns , Yishay Mansour , Satinder Singh

This work presents a novel policy iteration algorithm to tackle nonzero-sum stochastic impulse games arising naturally in many applications. Despite the obvious impact of solving such problems, there are no suitable numerical methods…

最优化与控制 · 数学 2020-06-29 René Aïd , Francisco Bernal , Mohamed Mnif , Diego Zabaljauregui , Jorge P. Zubelli

This paper is concerned with a new class of mean-field games which involve a finite number of agents. Necessary and sufficient conditions are obtained for the existence of the decentralized open-loop Nash equilibrium in terms of…

最优化与控制 · 数学 2022-06-14 Bing-Chang Wang , Huanshui Zhang , Minyue Fu , Yong Liang