中文
相关论文

相关论文: Policy Optimization for Linear-Quadratic Zero-Sum …

200 篇论文

In this work, we propose, for the first time, a reinforcement learning framework specifically designed for zero-sum linear-quadratic stochastic differential games. This approach offers a generalized solution for scenarios in which accurate…

最优化与控制 · 数学 2026-02-10 Yiyuan Wang

In this paper, we consider the problem of optimization and learning for constrained and multi-objective Markov decision processes, for both discounted rewards and expected average rewards. We formulate the problems as zero-sum games where…

最优化与控制 · 数学 2021-03-05 Ather Gattami , Qinbo Bai , Vaneet Agarwal

We study an $N$-player and a mean field exponential utility game. Each player manages two stocks; one is driven by an individual shock and the other is driven by a common shock. Moreover, each player is concerned not only with her own…

最优化与控制 · 数学 2020-07-17 Guanxing Fu , Xizhi Su , Chao Zhou

We study how to synthesize a robust and safe policy for autonomous systems under signal temporal logic (STL) tasks in adversarial settings against unknown dynamic agents. To ensure the worst-case STL satisfaction, we propose STLGame, a…

机器人学 · 计算机科学 2024-12-03 Shuo Yang , Hongrui Zheng , Cristian-Ioan Vasile , George Pappas , Rahul Mangharam

In this paper we formulate and solve a mean-field game described by a linear stochastic dynamics and a quadratic or exponential-quadratic cost functional for each generic player. The optimal strategies for the players are given explicitly…

最优化与控制 · 数学 2014-12-02 Djehiche Boualem , Tembine Hamidou

We investigate stochastic utility maximization games under relative performance concerns in both finite-agent and infinite-agent (graphon) settings. An incomplete market model is considered where agents with power (CRRA) utility functions…

最优化与控制 · 数学 2024-12-05 Zongxia Liang , Keyu Zhang , Yaqi Zhuang

Learning by experience in Multi-Agent Systems (MAS) is a difficult and exciting task, due to the lack of stationarity of the environment, whose dynamics evolves as the population learns. In order to design scalable algorithms for systems…

最优化与控制 · 数学 2020-02-24 Romuald Elie , Julien Pérolat , Mathieu Laurière , Matthieu Geist , Olivier Pietquin

This work studies the behaviors of two large-population teams competing in a discrete environment. The team-level interactions are modeled as a zero-sum game while the agent dynamics within each team is formulated as a collaborative…

系统与控制 · 电气工程与系统科学 2024-02-26 Yue Guan , Mohammad Afshari , Panagiotis Tsiotras

In recent years, state-of-the-art game-playing agents often involve policies that are trained in self-playing processes where Monte Carlo tree search (MCTS) algorithms and trained policies iteratively improve each other. The strongest…

机器学习 · 计算机科学 2019-05-16 Dennis J. N. J. Soemers , Éric Piette , Matthew Stephenson , Cameron Browne

Zero-sum Markov Stackelberg games can be used to model myriad problems, in domains ranging from economics to human robot interaction. In this paper, we develop policy gradient methods that solve these games in continuous state and action…

计算机科学与博弈论 · 计算机科学 2024-01-24 Denizalp Goktas , Arjun Prakash , Amy Greenwald

Zero-sum games are natural, if informal, analogues of closed physical systems where no energy/utility can enter or exit. This analogy can be extended even further if we consider zero-sum network (polymatrix) games where multiple agents…

计算机科学与博弈论 · 计算机科学 2019-03-06 James P. Bailey , Georgios Piliouras

We introduce a new class of games called the networked common goods game (NCGG), which generalizes the well-known common goods game. We focus on a fairly general subclass of the game where each agent's utility functions are the same across…

计算机科学与博弈论 · 计算机科学 2015-05-18 Jinsong Tan

We propose a single-level numerical approach to solve Stackelberg mean field game (MFG) problems. In Stackelberg MFG, an infinite population of agents play a non-cooperative game and choose their controls to optimize their individual…

最优化与控制 · 数学 2024-04-24 Gokce Dayanikli , Mathieu Lauriere

Empirically derived continuum models of collective behavior among large populations of dynamic agents are a subject of intense study in several fields, including biology, engineering and finance. We formulate and study a mean-field game…

适应与自组织系统 · 物理学 2018-06-22 Piyush Grover , Kaivalya Bakshi , Evangelos A. Theodorou

Mean-field theory has been extensively explored in decision analysis of {large-scale} (LS) systems but traditionally in ``pure" cooperative or competitive settings. This leads to the so-called mean-field game (MG) or mean-field team (MT).…

最优化与控制 · 数学 2023-06-30 Huang Jianhui , Qiu Zhenghong , Wang Shujun , Wu Zhen

Finding Nash equilibria in two-player zero-sum continuous games is a central problem in machine learning, e.g. for training both GANs and robust models. The existence of pure Nash equilibria requires strong conditions which are not…

机器学习 · 计算机科学 2021-05-07 Carles Domingo-Enrich , Samy Jelassi , Arthur Mensch , Grant Rotskoff , Joan Bruna

Mean field games (MFGs) offer a powerful framework for modeling large-scale multi-agent systems. This paper addresses MFGs formulated in continuous time with discrete state spaces, where agents' dynamics are governed by continuous-time…

计算机科学与博弈论 · 计算机科学 2026-02-27 Yannick Eich , Christian Fabian , Kai Cui , Heinz Koeppl

The goal of this paper is to study a Mean Field Game (MFG) system stemming from the harvesting of resources. Modelling the latter through a reaction-diffusion equation and the harvesters as competing rational agents, we are led to a…

偏微分方程分析 · 数学 2024-06-11 Ziad Kobeissi , Idriss Mazari-Fouquer , Domènec Ruiz-Balet

In this paper, we formulate a two-player zero-sum game under dynamic constraints defined by hybrid dynamical equations. The game consists of a min-max problem involving a cost functional that depends on the actions and resulting solutions…

最优化与控制 · 数学 2025-05-20 Santiago J. Leudo , Ricardo G. Sanfelice

We present a new combined \textit{mean field control game} (MFCG) problem which can be interpreted as a competitive game between collaborating groups and its solution as a Nash equilibrium between groups. Players coordinate their strategies…

最优化与控制 · 数学 2023-02-16 Andrea Angiuli , Nils Detering , Jean-Pierre Fouque , Mathieu Lauriere , Jimin Lin