中文
相关论文

相关论文: Reinforcement Learning for Inverse Non-Cooperative…

200 篇论文

In this short note, we consider an inverse problem to a mean-field games system where we are interested in reconstructing the state-independent running cost function from observed value-function data. We provide an elementary proof of a…

偏微分方程分析 · 数学 2024-08-16 Kui Ren , Nathan Soedjak , Kewei Wang , Hongyu Zhai

We investigate a novel finite-horizon linear-quadratic (LQ) feedback dynamic potential game with a priori unknown cost matrices played between two players. The cost matrices are revealed to the players sequentially, with the potential for…

最优化与控制 · 数学 2025-08-05 Yitian Chen , Timothy L. Molloy , Iman Shames

The evaluation of constitutive models, especially for high-risk and high-regret engineering applications, requires efficient and rigorous third-party calibration, validation and falsification. While there are numerous efforts to develop…

信号处理 · 电气工程与系统科学 2020-12-02 Kun Wang , WaiChing Sun , Qiang Du

Starting from a heuristic learning scheme for N-person games, we derive a new class of continuous-time learning dynamics consisting of a replicator-like drift adjusted by a penalty term that renders the boundary of the game's strategy space…

最优化与控制 · 数学 2014-04-08 Pierre Coucheney , Bruno Gaujal , Panayotis Mertikopoulos

Dynamic noncooperative game theory is a field of mathematics and economics in which a lot of research is being carried out at present featuring a great number of applications in many different areas of economics and management science like…

最优化与控制 · 数学 2014-09-19 Philipp Hungerländer

This paper proposes and studies a general form of dynamic $N$-player non-cooperative games called $\alpha$-potential games, where the change of a player's value function upon her unilateral deviation from her strategy is equal to the change…

最优化与控制 · 数学 2025-04-02 Xin Guo , Xinyu Li , Yufei Zhang

We propose a framework to compute approximate Nash equilibria in integer programming games with nonlinear payoffs, i.e., simultaneous and non-cooperative games where each player solves a parametrized mixed-integer nonlinear program. We…

最优化与控制 · 数学 2025-08-04 Aloïs Duguet , Margarida Carvalho , Gabriele Dragotto , Sandra Ulrich Ngueveu

A fundamental problem in noncooperative dynamic game theory is the computation of Nash equilibria under different information structures, which specify the information available to each agent during decision-making. Prior work has…

计算机科学与博弈论 · 计算机科学 2026-03-20 Janani S K , Kushagra Gupta , Ufuk Topcu , David Fridovich-Keil

This paper investigates the problem of computing the equilibrium of competitive games, which is often modeled as a constrained saddle-point optimization problem with probability simplex constraints. Despite recent efforts in understanding…

最优化与控制 · 数学 2023-01-23 Shicong Cen , Yuting Wei , Yuejie Chi

In this paper, we consider a Nash equilibrium seeking problem for a class of high-order multi-agent systems with unknown dynamics. Different from existing results for single integrators, we aim to steer the outputs of this class of…

系统与控制 · 电气工程与系统科学 2021-01-11 Yutao Tang , Peng Yi

In this book, we present a curated collection of existing results on inverse problems for Mean Field Games (MFGs), a cutting-edge and rapidly evolving field of research. Our aim is to provide fresh insights, novel perspectives, and a…

偏微分方程分析 · 数学 2025-03-20 Hongyu Liu , Catharine W. K. Lo , Shen Zhang

Making decisions in the presence of a strategic opponent requires one to take into account the opponent's ability to actively mask its intended objective. To describe such strategic situations, we introduce the non-cooperative inverse…

计算机科学与博弈论 · 计算机科学 2020-01-07 Xiangyuan Zhang , Kaiqing Zhang , Erik Miehling , Tamer Başar

Reinforcement-based learning has attracted considerable attention both in modeling human behavior as well as in engineering, for designing measurement- or payoff-based optimization schemes. Such learning schemes exhibit several advantages,…

机器学习 · 计算机科学 2025-11-26 Georgios C. Chasparis

Signaling game problems investigate communication scenarios where encoder(s) and decoder(s) have misaligned objectives due to the fact that they either employ different cost functions or have inconsistent priors. This problem has been…

信息论 · 计算机科学 2023-05-09 Ertan Kazıklı , Sinan Gezici , Serdar Yüksel

We consider multi-agent decision making, where each agent optimizes its cost function subject to constraints. Agents' actions belong to a compact convex Euclidean space and the agents' cost functions are coupled. We propose a distributed…

最优化与控制 · 数学 2016-12-01 Tatiana Tatarenko , Maryam Kamgarpour

This paper proposes a novel approach for locally stable convergence to Nash equilibrium in duopoly noncooperative games based on a distributed event-triggered control scheme. The proposed approach employs extremum seeking, with sinusoidal…

最优化与控制 · 数学 2024-04-12 Victor Hugo Pereira Rodrigues , Tiago Roux Oliveira , Miroslav Krstić , Tamer Başar

In this paper, we investigate a class of nonzero-sum dynamic stochastic games, where players have linear dynamics and quadratic cost functions. The players are coupled in both dynamics and cost through a linear regression (weighted average)…

最优化与控制 · 数学 2020-10-20 Jalal Arabneydi , Amir G. Aghdam , Roland P. Malhamé

This paper develops a predictive compensation framework for finite-horizon, discrete-time linear quadratic dynamic games subject to Gauss-Markov execution deviations from feedback Nash strategies. One player's control is corrupted by…

系统与控制 · 电气工程与系统科学 2025-11-18 Navid Mojahed , Mahdis Rabbani , Shima Nazari

Reinforcement-based learning dynamics may exhibit several limitations when applied in a distributed setup. In (repeatedly-played) multi-player/action strategic-form games, and when each player applies an independent copy of the learning…

计算机科学与博弈论 · 计算机科学 2025-11-25 Georgios C. Chasparis

For the unicycle system, we provide constructive methods for the design of feedback laws that have one or more of the following properties: being nonmodular and globally exponentially stabilizing, inverse optimal, robust to arbitrary…

系统与控制 · 电气工程与系统科学 2025-12-25 Kwang Hak Kim , Velimir Todorovski , Miroslav Krstić