中文
相关论文

相关论文: Offline and Online Nonlinear Inverse Differential …

200 篇论文

We consider stochastic differential games with $N$ players, linear-Gaussian dynamics in arbitrary state-space dimension, and long-time-average cost with quadratic running cost. Admissible controls are feedbacks for which the system is…

偏微分方程分析 · 数学 2014-07-10 Martino Bardi , Fabio S. Priuli

We propose a deep neural network-based algorithm to identify the Markovian Nash equilibrium of general large $N$-player stochastic differential games. Following the idea of fictitious play, we recast the $N$-player game into $N$ decoupled…

最优化与控制 · 数学 2020-06-08 Jiequn Han , Ruimeng Hu

In this paper, we consider a large class of constrained non-cooperative stochastic Markov games with countable state spaces and discounted cost criteria. In one-player case, i.e., constrained discounted Markov decision models, it is…

最优化与控制 · 数学 2021-12-16 Anna Jaśkiewicz , Andrzej S. Nowak

For the characterization of Feedback Nash Equilibria (FNE) in linear quadratic games, this paper provides a detailed analysis of the discrete-time discounted coupled best-response equations for the scalar two-player setting, together with a…

最优化与控制 · 数学 2026-03-17 Chiara Cavalagli , Alberto Bemporad , Mario Zanon

We provide a distributed algorithm to learn a Nash equilibrium in a class of non-cooperative games with strongly monotone mappings and unconstrained action sets. Each player has access to her own smooth local cost function and can…

最优化与控制 · 数学 2019-07-17 Tatiana Tatarenko , Angelia Nedich

The recently defined class of integer programming games (IPG) models situations where multiple self-interested decision makers interact, with their strategy sets represented by a finite set of linear constraints together with integer…

计算机科学与博弈论 · 计算机科学 2023-04-25 Margarida Carvalho , Andrea Lodi , João Pedro Pedroso

Multi-turn LLM evaluation is typically reported as a single win-rate scalar, conflating distinct capabilities. We introduce AIDG (Adversarial Information Deduction Game), formalizing multi-turn adversarial dialogue as a two-player partially…

计算与语言 · 计算机科学 2026-05-27 Adib Sakhawat , Fardeen Sadab , Rakin Shahriar

From large-scale organizations to decentralized political systems, hierarchical strategic decision making is commonplace. We introduce a novel class of structured hierarchical games (SHGs) that formally capture such hierarchical strategic…

计算机科学与博弈论 · 计算机科学 2022-06-28 Zun Li , Feiran Jia , Aditya Mate , Shahin Jabbari , Mithun Chakraborty , Milind Tambe , Yevgeniy Vorobeychik

Deep learning is built on the foundational guarantee that gradient descent on an objective function converges to local minima. Unfortunately, this guarantee fails in settings, such as generative adversarial nets, that exhibit multiple…

Deep reinforcement learning (DRL) techniques have become increasingly used in various fields for decision-making processes. However, a challenge that often arises is the trade-off between both the computational efficiency of the…

机器学习 · 计算机科学 2023-08-21 Anthony Kobanda , Valliappan C. A. , Joshua Romoff , Ludovic Denoyer

We examine global non-asymptotic convergence properties of policy gradient methods for multi-agent reinforcement learning (RL) problems in Markov potential games (MPG). To learn a Nash equilibrium of an MPG in which the size of state space…

机器学习 · 计算机科学 2022-08-08 Dongsheng Ding , Chen-Yu Wei , Kaiqing Zhang , Mihailo R. Jovanović

Game theory finds nowadays a broad range of applications in engineering and machine learning. However, in a derivative-free, expensive black-box context, very few algorithmic solutions are available to find game equilibria. Here, we propose…

机器学习 · 统计学 2018-02-28 Victor Picheny , Mickael Binois , Abderrahmane Habbal

Static reduction of information structures (ISs) is a method that is commonly adopted in stochastic control, team theory, and game theory. One approach entails change of measure arguments, which has been crucial for stochastic analysis and…

最优化与控制 · 数学 2023-07-13 Sina Sanjari , Tamer Başar , Serdar Yüksel

We consider a weighted Shapley network design game, where selfish players choose paths in a network to minimize their cost. The cost function of each edge in the network is affine linear with respect to the sum of weights of the players who…

计算机科学与博弈论 · 计算机科学 2023-12-19 Hangxin Gan , Xianhao Meng , Chunying Ren , Yongtang Shi

Inverse Optimal Control (IOC) seeks to recover an unknown cost from expert demonstrations, and it provides a systematic way of modeling experts' decision mechanisms while considering the prior information of the cost functions.…

最优化与控制 · 数学 2025-12-01 Ziliang Wang , Han Zhang , Axel Ringh

Congestion games are a classical type of games studied in game theory, in which n players choose a resource, and their individual cost increases with the number of other players choosing the same resource. In network congestion games…

计算机科学与博弈论 · 计算机科学 2020-09-30 Nathalie Bertrand , Nicolas Markey , Suman Sadhukhan , Ocan Sankur

In this paper, we study the problem of learning the set of pure strategy Nash equilibria and the exact structure of a continuous-action graphical game with quadratic payoffs by observing a small set of perturbed equilibria. A…

计算机科学与博弈论 · 计算机科学 2019-11-12 Adarsh Barik , Jean Honorio

We test the performance of deep deterministic policy gradient (DDPG), a deep reinforcement learning algorithm, able to handle continuous state and action spaces, to learn Nash equilibria in a setting where firms compete in prices. These…

计算机科学与博弈论 · 计算机科学 2025-09-30 Christoph Graf , Viktor Zobernig , Johannes Schmidt , Claude Klöckl

We consider non-differentiable dynamic optimization problems such as those arising in robotics and subspace tracking. Given the computational constraints and the time-varying nature of the problem, a low-complexity algorithm is desirable,…

最优化与控制 · 数学 2019-02-20 Rishabh Dixit , Amrit Singh Bedi , Ruchi Tripathi , Ketan Rajawat

In this paper, online game is studied, where at each time, a group of players aim at selfishly minimizing their own time-varying cost function simultaneously subject to time-varying coupled constraints and local feasible set constraints.…

计算机科学与博弈论 · 计算机科学 2023-06-29 Min Meng , Xiuxian Li , Yiguang Hong , Jie Chen , Long Wang