中文
相关论文

相关论文: Offline and Online Nonlinear Inverse Differential …

200 篇论文

This paper presents a method for solving the Inverse Stochastic Differential Game (ISDG) problem in finite-horizon linear-quadratic Gaussian (LQG) differential games. The objective is to recover cost function parameters of all players, as…

系统与控制 · 电气工程与系统科学 2026-04-29 Lucas Günther , Felix Thömmes , Karl Handwerker , Balint Varga , Sören Hohmann

In multi-agent dynamic games, the Nash equilibrium state trajectory of each agent is determined by its cost function and the information pattern of the game. However, the cost and trajectory of each agent may be unavailable to the other…

多智能体系统 · 计算机科学 2023-01-05 Jingqi Li , Chih-Yuan Chiu , Lasse Peters , Somayeh Sojoudi , Claire Tomlin , David Fridovich-Keil

In this work, we analyze the applicability of Inverse Dynamic Game (IDG) methods based on the Minimum Principle (MP). The IDG method determines unknown cost functions in a single- or multi-agent setting from observed system trajectories by…

最优化与控制 · 数学 2024-06-18 Philipp Karg , Adrian Kienzle , Jonas Kaub , Balint Varga , Sören Hohmann

In this paper, we address the inverse problem in the case of linear-quadratic discrete-time dynamic non-cooperative games. Given feedback laws of players that are known to be a Nash equilibrium pair for a discrete-time linear system, we…

最优化与控制 · 数学 2024-07-19 Emin Martirosyan , Ming Cao

$ $This paper addresses the inverse problem for Linear-Quadratic (LQ) nonzero-sum $N$-player differential games, where the goal is to learn parameters of an unknown cost function for the game, called observed, given the demonstrated…

最优化与控制 · 数学 2024-10-28 Emin Martirosyan , Ming Cao

In this paper, we address the inverse problem for linear-quadratic differential non-cooperative games with output-feedback. Given players' stabilizing feedback laws, the goal is to find cost function parameters that lead to a game for which…

最优化与控制 · 数学 2024-10-27 Emin Martirosyan , Ming Cao

We address cost identification in a finite-horizon linear quadratic Gaussian game. We characterize the set of cost parameters that generate a given Nash equilibrium policy. We propose a backpropagation algorithm to identify the time-varying…

系统与控制 · 电气工程与系统科学 2025-11-19 Kai Ren , Maryam Kamgarpour

We consider the inverse problem of dynamic games, where cost function parameters are sought which explain observed behavior of interacting players. Maximum entropy inverse reinforcement learning is extended to the N-player case in order to…

系统与控制 · 电气工程与系统科学 2020-07-27 Jairo Inga , Esther Bischoff , Florian Köpf , Sören Hohmann

This work studies the parameter identification problem of a generalized non-cooperative game, where each player's cost function is influenced by an observable signal and some unknown parameters. We consider the scenario where equilibrium of…

计算机科学与博弈论 · 计算机科学 2023-10-17 Jianguo Chen , Jinlong Lei , Hongsheng Qi , Yiguang Hong

In this letter, we study a model-based inverse problem for infinite-horizon linear-quadratic differential games with descriptor dynamics. Given an observed feedback strategy profile, we seek to identify all cost functions that rationalize…

最优化与控制 · 数学 2026-05-20 Aaditya Kumar , Puduru Viswanadha Reddy

For a non-cooperative differential game, the value functions of the various players satisfy a system of Hamilton-Jacobi equations. In the present paper, we consider a class of infinite-horizon games with nonlinear costs exponentially…

偏微分方程分析 · 数学 2014-08-07 Alberto Bressan , Fabio S. Priuli

This paper focuses on the development of an online inverse reinforcement learning (IRL) technique for a class of nonlinear systems. The developed approach utilizes observed state and input trajectories, and determines the unknown cost…

系统与控制 · 计算机科学 2018-11-27 Ryan Self , Michael Harlan , Rushikesh Kamalapurkar

We address the problem of finding conditions which guarantee the existence of open-loop Nash equilibria in discrete time dynamic games (DTDGs). The classical approach to DTDGs involves analyzing the problem using optimal control theory…

最优化与控制 · 数学 2015-09-22 Mathew P. Abraham , Ankur A. Kulkarni

In this paper, we consider a linear-quadratic-Gaussian defending assets differential game (DADG) where the attacker and the defender do not know each other's state information while they know the trajectory of a moving asset. Both players…

系统与控制 · 电气工程与系统科学 2021-03-25 Yunhan Huang , Juntao Chen , Quanyan Zhu

This paper introduces two new identification methods for linear quadratic (LQ) ordinal potential differential games (OPDGs). Potential games are notable for their benefits, such as the computability and guaranteed existence of Nash…

动力系统 · 数学 2025-03-06 Balint Varga , Da Huang , Sören Hohmann

This paper investigates online stochastic aggregative games subject to local set constraints and time-varying coupled inequality constraints, where each player possesses a time-varying expectation-valued cost function relying on not only…

最优化与控制 · 数学 2025-11-18 Kaixin Du , Min Meng

In this paper, we study Nash equilibrium payoffs for nonzero-sum stochastic differential games via the theory of backward stochastic differential equations. We obtain an existence theorem and a characterization theorem of Nash equilibrium…

概率论 · 数学 2011-11-30 Qian Lin

Decentralized online learning for seeking generalized Nash equilibrium (GNE) of noncooperative games in dynamic environments is studied in this paper. Each player aims at selfishly minimizing its own time-varying cost function subject to…

最优化与控制 · 数学 2021-05-14 Min Meng , Xiuxian Li , Yiguang Hong , Jie Chen , Long Wang

In this paper, the inverse reinforcement learning (IRL) problem is addressed to reconstruct the unknown cost function underlying an observed optimal policy in a model-free manner, whose online adaptation with completely off-policy system…

最优化与控制 · 数学 2025-11-20 Yibei Li , Yuexin Cao , Zhixin Liu , Lihua Xie

We study a zero-sum stochastic differential game (SDG) in which one controller plays an impulse control while their opponent plays a stochastic control. We consider an asymmetric setting in which the impulse player commits to, at the start…

概率论 · 数学 2019-01-31 Parsiad Azimzadeh
‹ 上一页 1 2 3 10 下一页 ›