中文
相关论文

相关论文: Reinforcement Learning for Inverse Non-Cooperative…

200 篇论文

Game theory is a very profound study on distributed decision-making behavior and has been extensively developed by many scholars. However, many existing works rely on certain strict assumptions such as knowing the opponent's private…

计算机科学与博弈论 · 计算机科学 2020-04-21 Kuo Chun Tsai , Zhu Han

We study a setting in which two players play a (possibly approximate) Nash equilibrium of a bimatrix game, while a learner observes only their actions and has no knowledge of the equilibrium or the underlying game. A natural question is…

计算机科学与博弈论 · 计算机科学 2026-05-27 Annalisa Barbara , Riccardo Poiani , Martino Bernasconi , Andrea Celli

This paper is devoted to a high-dimensional mixed leadership stochastic differential game on a finite horizon in feedback information mode, where the control variables enter into the diffusion term of state equation. A verification theorem…

最优化与控制 · 数学 2022-11-28 Qi Huang , Jingtao Shi

We address cost identification in a finite-horizon linear quadratic Gaussian game. We characterize the set of cost parameters that generate a given Nash equilibrium policy. We propose a backpropagation algorithm to identify the time-varying…

系统与控制 · 电气工程与系统科学 2025-11-19 Kai Ren , Maryam Kamgarpour

In this paper, we address the problem of a two-player linear quadratic differential game with incomplete information, a scenario commonly encountered in multi-agent control, human-robot interaction (HRI), and approximation methods for…

系统与控制 · 电气工程与系统科学 2025-04-25 Seyed Yousef Soltanian , Wenlong Zhang

Game-theoretic models are effective tools for modeling multi-agent interactions, especially when robots need to coordinate with humans. However, applying these models requires inferring their specifications from observed behaviors -- a…

机器人学 · 计算机科学 2025-02-06 Max Muchen Sun , Pete Trautman , Todd Murphey

Nearly all simulation-based games have environment parameters that affect incentives in the interaction but are not explicitly incorporated into the game model. To understand the impact of these parameters on strategic incentives, typical…

计算机科学与博弈论 · 计算机科学 2026-05-06 Madelyn Gatchel , Bryce Wiedenbeck

In inverse reinforcement learning (IRL), an agent seeks to replicate expert demonstrations through interactions with the environment. Traditionally, IRL is treated as an adversarial game, where an adversary searches over reward models, and…

机器学习 · 计算机科学 2025-04-23 Arnav Kumar Jain , Harley Wiltzer , Jesse Farebrother , Irina Rish , Glen Berseth , Sanjiban Choudhury

This paper studies a class of strongly monotone games involving non-cooperative agents that optimize their own time-varying cost functions. We assume that the agents can observe other agents' historical actions and choose actions that best…

最优化与控制 · 数学 2023-09-04 Zifan Wang , Yi Shen , Michael M. Zavlanos , Karl H. Johansson

The standard Reinforcement Learning from Human Feedback (RLHF) framework primarily focuses on optimizing the performance of large language models using pre-collected prompts. However, collecting prompts that provide comprehensive coverage…

In this work, we consider a novel inverse problem in mean-field games (MFG). We aim to recover the MFG model parameters that govern the underlying interactions among the population based on a limited set of noisy partial observations of the…

数值分析 · 数学 2022-04-12 Yat Tin Chow , Samy Wu Fung , Siting Liu , Levon Nurbekyan , Stanley Osher

In this paper, a novel approach to the output-feedback inverse reinforcement learning (IRL) problem is developed by casting the IRL problem, for linear systems with quadratic cost functions, as a state estimation problem. Two observer-based…

系统与控制 · 电气工程与系统科学 2023-07-19 Ryan Self , Kevin Coleman , He Bai , Rushikesh Kamalapurkar

We consider dynamic games with linear dynamics and quadratic objective functions. We observe that the unconstrained open-loop Nash equilibrium coincides with a linear quadratic regulator in an augmented space, thus deriving an explicit…

系统与控制 · 电气工程与系统科学 2025-07-22 Emilio Benenati , Sergio Grammatico

In this paper, we study a class of two-player deterministic finite-horizon difference games with coupled inequality constraints, where each player has two types of decision variables: one involving sequential interactions and the other…

最优化与控制 · 数学 2025-10-20 Partha Sarathi Mohapatra , Puduru Viswanadha Reddy , Georges Zaccour

This work proposes a policy learning algorithm for seeking generalised feedback Nash equilibria (GFNE) in $N_P$-player noncooperative dynamic games. We consider linear-quadratic games with stochastic dynamics and design a best-response…

最优化与控制 · 数学 2025-06-13 Otacilio B. L. Neto , Michela Mulas , Francesco Corona

This note studies the robust output feedback stabilization problem of a class of multi-input multi-output invertible nonlinear systems, for which an "ideal" state feedback based on feedback linearization can be designed under certain mild…

系统与控制 · 电气工程与系统科学 2021-01-07 Lei Wang , Christopher M. Kellett

In this paper, we study an inverse reinforcement learning problem that involves learning the reward function of a learning agent using trajectory data collected while this agent is learning its optimal policy. To address this problem, we…

机器学习 · 计算机科学 2024-10-21 Kavinayan P. Sivakumar , Yi Shen , Zachary Bell , Scott Nivison , Boyuan Chen , Michael M. Zavlanos

We address safe multi-robot interaction under uncertainty. In particular, we formulate a chance-constrained linear quadratic Gaussian game with coupling constraints and system uncertainties. We find a tractable reformulation of the game and…

机器人学 · 计算机科学 2025-08-15 Kai Ren , Giulio Salizzoni , Mustafa Emre Gürsoy , Maryam Kamgarpour

Iterative linear-quadratic (ILQ) methods are widely used in the nonlinear optimal control community. Recent work has applied similar methodology in the setting of multiplayer general-sum differential games. Here, ILQ methods are capable of…

系统与控制 · 电气工程与系统科学 2020-03-20 David Fridovich-Keil , Vicenc Rubies-Royo , Claire J. Tomlin

For the characterization of Feedback Nash Equilibria (FNE) in linear quadratic games, this paper provides a detailed analysis of the discrete-time discounted coupled best-response equations for the scalar two-player setting, together with a…

最优化与控制 · 数学 2026-03-17 Chiara Cavalagli , Alberto Bemporad , Mario Zanon