中文
相关论文

相关论文: Offline and Online Nonlinear Inverse Differential …

200 篇论文

This paper addresses the problem of online inverse reinforcement learning for nonlinear systems with modeling uncertainties while in the presence of unknown disturbances. The developed approach observes state and input trajectories for an…

系统与控制 · 电气工程与系统科学 2021-07-07 Ryan Self , Moad Abudia , Rushikesh Kamalapurkar

In this paper, we consider two-player zero-sum matrix and stochastic games and develop learning dynamics that are payoff-based, convergent, rational, and symmetric between the two players. Specifically, the learning dynamics for matrix…

机器学习 · 计算机科学 2024-09-06 Zaiwei Chen , Kaiqing Zhang , Eric Mazumdar , Asuman Ozdaglar , Adam Wierman

Inverse reinforcement learning (IRL) and dynamic discrete choice (DDC) models explain sequential decision-making by recovering reward functions that rationalize observed behavior. Flexible IRL methods typically rely on machine learning but…

机器学习 · 计算机科学 2026-01-01 Lars van der Laan , Aurelien Bibaut , Nathan Kallus

This paper considers a distributed gossip approach for finding a Nash equilibrium in networked games on graphs. In such games a player's cost function may be affected by the actions of any subset of players. An interference graph is…

计算机科学与博弈论 · 计算机科学 2020-04-10 Farzad Salehisadaghiani , Lacra Pavel

In this paper, we investigate the noncooperative games of multi-agent systems. Different from existing noncooperative games, our formulation involves the high-order nonlinear dynamics of players, and the communication topologies among…

系统与控制 · 电气工程与系统科学 2021-12-17 Zhenhua Deng , Jin Luo

In this paper, we study infinite-horizon linear-quadratic uncertain differential games with an output feedback information structure. We assume linear time-invariant nominal dynamics influenced by deterministic external disturbances, and…

最优化与控制 · 数学 2024-12-04 Aniruddha Roy , Puduru Viswanadha Reddy

Path planning plays an essential role in many areas of robotics. Various planning techniques have been presented, either focusing on learning a specific task from demonstrations or retrieving trajectories by optimizing for hand-crafted cost…

机器人学 · 计算机科学 2018-09-26 Salvatore Virga , Christian Rupprecht , Nassir Navab , Christoph Hennersperger

In this paper, we consider the problem of computing parameters of an objective function for a discrete-time optimal control problem from state and control trajectories with active control constraints. We propose a novel method of inverse…

系统与控制 · 电气工程与系统科学 2020-05-14 Timothy L. Molloy , Jason J. Ford , Tristan Perez

We consider a class of two-player dynamic stochastic nonzero-sum games where the state transition and observation equations are linear, and the primitive random variables are Gaussian. Each controller acquires possibly different dynamic…

系统与控制 · 计算机科学 2014-01-21 Abhishek Gupta , Ashutosh Nayyar , Cedric Langbort , Tamer Basar

Online decision systems routinely operate under delayed feedback and order-sensitive (noncommutative) dynamics: actions affect which observations arrive, and in what sequence. Taking a Bregman divergence $D_\Phi$ as the loss benchmark, we…

机器学习 · 计算机科学 2025-11-18 Duo Yi

We study episodic two-player zero-sum Markov games (MGs) in the offline setting, where the goal is to find an approximate Nash equilibrium (NE) policy pair based on a dataset collected a priori. When the dataset does not have uniform…

机器学习 · 计算机科学 2023-01-02 Han Zhong , Wei Xiong , Jiyuan Tan , Liwei Wang , Tong Zhang , Zhaoran Wang , Zhuoran Yang

Zero-sum games arise in a wide variety of problems, including robust optimization and adversarial learning. However, algorithms deployed for finding a local Nash equilibrium in these games often converge to non-Nash stationary points. This…

计算机科学与博弈论 · 计算机科学 2025-09-30 Kushagra Gupta , Xinjie Liu , Ross Allen , Ufuk Topcu , David Fridovich-Keil

We introduce an online learning algorithm in the bandit feedback model that, once adopted by all agents of a congestion game, results in game-dynamics that converge to an $\epsilon$-approximate Nash Equilibrium in a polynomial number of…

计算机科学与博弈论 · 计算机科学 2024-01-19 Leello Dadi , Ioannis Panageas , Stratis Skoulakis , Luca Viano , Volkan Cevher

We study two person nonzero-sum stochastic differential games with risk-sensitive discounted and ergodic cost criteria. Under certain conditions we establish a Nash equilibrium in Markov strategies for the discounted cost criterion and a…

最优化与控制 · 数学 2016-04-06 Mrinal K. Ghosh , K. Suresh Kumar , Chandan Pal

This paper considers no-regret learning for repeated continuous-kernel games with lossy bandit feedback. Since it is difficult to give the explicit model of the utility functions in dynamic environments, the players' action can only be…

机器学习 · 计算机科学 2022-05-17 Wenting Liu , Jinlong Lei , Peng Yi , Yiguang Hong

In this paper, we consider state and control path-dependent stochastic zero-sum differential games, where the dynamics and the running cost include both state and control paths of the players. Using the notion of nonanticipative strategies,…

最优化与控制 · 数学 2021-02-10 Jun Moon

This paper investigates the problem of Online Convex-Concave Optimization, which extends Online Convex Optimization to two-player time-varying convex-concave games. The goal is to minimize the dynamic duality gap (D-DGap), a critical…

机器学习 · 计算机科学 2025-09-10 Qing-xin Meng , Xia Lei , Jian-wei Liu

We propose and study several inverse problems for the mean field games (MFG) system in a bounded domain. Our focus is on simultaneously recovering the running cost and the Hamiltonian within the MFG system by the associated boundary…

最优化与控制 · 数学 2024-03-05 Hongyu Liu , Shen Zhang

Dynamic games arise when multiple agents with differing objectives choose control inputs to a dynamic system. Dynamic games model a wide variety of applications in economics, defense, and energy systems. However, compared to single-agent…

最优化与控制 · 数学 2018-09-25 Bolei Di , Andrew Lamperski

This paper focuses on a kind of linear quadratic non-zero sum differential game driven by backward stochastic differential equation with asymmetric information, which is a natural continuation of Wang and Yu [IEEE TAC (2010) 55: 1742-1747,…

最优化与控制 · 数学 2017-03-06 Guangchen Wang , Hua Xiao , Jie Xiong