中文
相关论文

相关论文: A Linear-quadratic Mean-Field Stochastic Stackelbe…

200 篇论文

A linear quadratic (LQ) stochastic optimization system involving large population, which is driven by forward-backward stochastic differential equation (FBSDE), is investigated in this paper. Agents cooperate with each other to minimize the…

最优化与控制 · 数学 2024-04-30 Guangchen Wang , Shujun Wang , Jie Xiong

We develop a probabilistic approach to continuous-time finite state mean field games. Based on an alternative description of continuous-time Markov chain by means of semimartingale and the weak formulation of stochastic optimal control, our…

概率论 · 数学 2018-08-24 Rene Carmona , Peiqi Wang

This paper introduces the new concept of (follower) satisfaction in Stackelberg games and compares the standard Stackelberg game with its satisfaction version. Simulation results are presented which suggest that the follower adopting…

计算机科学与博弈论 · 计算机科学 2024-08-22 Langford White , Duong Nguyen , Hung Nguyen

Model-based reinforcement learning (MBRL) has recently gained immense interest due to its potential for sample efficiency and ability to incorporate off-policy data. However, designing stable and efficient MBRL algorithms using rich…

机器学习 · 计算机科学 2021-03-12 Aravind Rajeswaran , Igor Mordatch , Vikash Kumar

We study a two-player dynamic Stackelberg game where the follower's intention is unknown to the leader. Classical formulations of the Stackelberg equilibrium (SE) assume that the follower's best response (BR) function is known to the…

系统与控制 · 电气工程与系统科学 2026-04-09 Cayetana Salinas-Rodriguez , Jonathan Rogers , Sarah H. Q. Li

This paper investigates the convergence of learning dynamics in Stackelberg games. In the class of games we consider, there is a hierarchical game being played between a leader and a follower with continuous action spaces. We establish a…

计算机科学与博弈论 · 计算机科学 2024-12-07 Tanner Fiez , Benjamin Chasnov , Lillian J. Ratliff

This paper is concerned with a Stackelberg stochastic differential game on a finite horizon in feedback information pattern. A system of parabolic partial differential equations is obtained at the level of Hamiltonian to give the…

最优化与控制 · 数学 2021-08-17 Qi Huang. Jingtao Shi

Mean-field game theory relies on approximating games that are intractable to model due to a very large to infinite population of players. While these kinds of games can be solved analytically via the associated system of partial…

机器学习 · 计算机科学 2026-04-16 Anna C. M. Thöni , Yoram Bachrach , Tal Kachman

Competitive games involving thousands or even millions of players are prevalent in real-world contexts, such as transportation, communications, and computer networks. However, learning in these large-scale multi-agent environments presents…

最优化与控制 · 数学 2025-02-04 Batuhan Yardim , Semih Cayci , Niao He

The purpose of this paper is to provide a complete probabilistic analysis of a large class of stochastic differential games for which the interaction between the players is of mean-field type. We implement the Mean-Field Games strategy…

概率论 · 数学 2012-10-23 Rene Carmona , Francois Delarue

Reinforcement learning is a powerful tool to learn the optimal policy of possibly multiple agents by interacting with the environment. As the number of agents grow to be very large, the system can be approximated by a mean-field problem.…

最优化与控制 · 数学 2020-08-18 Weichen Wang , Jiequn Han , Zhuoran Yang , Zhaoran Wang

In this paper,we mainly focus on the numerical solution of high-dimensional stochastic optimal control problem driven by fully-coupled forward-backward stochastic differential equations (FBSDEs in short) through deep learning. We first…

最优化与控制 · 数学 2024-08-21 Shaolin Ji , Shige Peng , Ying Peng , Xichuan Zhang

This paper considers linear quadratic (LQ) mean field games with a major player and analyzes an asymptotic solvability problem. It starts with a large-scale system of coupled dynamic programming equations and applies a re-scaling technique…

最优化与控制 · 数学 2019-09-04 Yan Ma , Minyi Huang

In this article, we establish precise convergence rates of a general class of $N$-Player Stackelberg games to their mean field limits, which allows the response time delay of information, empirical distribution based interactions, and the…

最优化与控制 · 数学 2025-10-06 Alain Bensoussan , Ziyu Huang , Sheng Wang , Sheung Chi Phillip Yam

Given a large number of homogeneous players that are distributed across three possible states, we consider the problem in which these players have to control their transition rates, while minimizing a cost. The optimal transition rates are…

系统与控制 · 计算机科学 2018-02-13 Leonardo Stella , Dario Bauso

In this paper, we study a linear-quadratic optimal control problem for mean-field stochastic differential equations driven by a Poisson random martingale measure and a multidimensional Brownian motion. Firstly, the existence and uniqueness…

最优化与控制 · 数学 2016-10-12 Maoning Tang , Qingxin Meng

Motivated by the omnipresence of hierarchical structures in many real-world applications, this study delves into the intricate realm of bi-level games, with a specific focus on exploring local Stackelberg equilibria as a solution concept.…

系统与控制 · 电气工程与系统科学 2024-02-23 Marko Maljkovic , Gustav Nilsson , Nikolas Geroliminis

This paper is concerned with a linear quadratic stochastic two-person zero-sum differential game with constant coefficients in an infinite time horizon. Open-loop and closed-loop saddle points are introduced. The existence of closed-loop…

最优化与控制 · 数学 2014-04-30 Jingrui Sun , Jiongmin Yong , Shuguang Zhang

Game-theoretic inverse learning is the problem of inferring a player's objectives from their actions. We formulate an inverse learning problem in a Stackelberg game between a leader and a follower, where each player's action is the…

计算机科学与博弈论 · 计算机科学 2024-10-15 William Ward , Yue Yu , Jacob Levy , Negar Mehr , David Fridovich-Keil , Ufuk Topcu

In this paper, we first prove that the mean-field stochastic linear quadratic (MFSLQ for short) control problem with random coefficients has a unique optimal control and derive a preliminary stochastic maximum principle to characterize this…

最优化与控制 · 数学 2025-05-28 Jie Xiong , Wen Xu