中文
相关论文

相关论文: Learning Stationary Correlated Equilibria in Const…

200 篇论文

This paper examines the convergence of no-regret learning in Cournot games with continuous actions. Cournot games are the essential model for many socio-economic systems, where players compete by strategically setting their output quantity.…

计算机科学与博弈论 · 计算机科学 2020-02-12 Yuanyuan Shi , Baosen Zhang

Iterative linear-quadratic (ILQ) methods are widely used in the nonlinear optimal control community. Recent work has applied similar methodology in the setting of multiplayer general-sum differential games. Here, ILQ methods are capable of…

系统与控制 · 电气工程与系统科学 2020-03-20 David Fridovich-Keil , Vicenc Rubies-Royo , Claire J. Tomlin

This paper considers a class of noncooperative games in which the feasible decision sets of all players are coupled together by a coupled inequality constraint. Adopting the variational inequality formulation of the game, we first introduce…

计算机科学与博弈论 · 计算机科学 2024-02-13 Huaqing Li , Liang Ran , Lifeng Zheng , Zhe Li , Jinhui Hu , Jun Li , Tingwen Huang

This paper investigates a fully distributed adaptive Nash equilibrium (NE) seeking algorithm for constrained noncooperative games with prescribed-time stability. On the one hand, prescribed-time stability for the proposed NE seeking…

最优化与控制 · 数学 2024-11-07 Sichen Qian

Reinforcement-based learning dynamics may exhibit several limitations when applied in a distributed setup. In (repeatedly-played) multi-player/action strategic-form games, and when each player applies an independent copy of the learning…

计算机科学与博弈论 · 计算机科学 2025-11-25 Georgios C. Chasparis

This work introduces a new general approach for the numerical analysis of stable equilibria to second order mean field games systems in cases where the uniqueness of solutions may fail. For the sake of simplicity, we focus on a simple…

偏微分方程分析 · 数学 2024-10-30 Jules Berry , Olivier Ley , Francisco J Silva

$ $This paper addresses the inverse problem for Linear-Quadratic (LQ) nonzero-sum $N$-player differential games, where the goal is to learn parameters of an unknown cost function for the game, called observed, given the demonstrated…

最优化与控制 · 数学 2024-10-28 Emin Martirosyan , Ming Cao

CNN-based steganalysis has recently achieved very good performance in detecting content-adaptive steganography. At the same time, recent works have shown that, by adopting an approach similar to that used to build adversarial examples, a…

多媒体 · 计算机科学 2019-06-04 Xiaoyu Shi , Benedetta Tondi , Bin Li , Mauro Barni

In this paper, we study a class of linear-quadratic (LQ) mean field games of controls with common noises and their corresponding $N$-player games. The theory of mean field game of controls considers a class of mean field games where the…

最优化与控制 · 数学 2022-06-13 Min Li , Chenchen Mou , Zhen Wu , Chao Zhou

This paper studies distributed Q-learning for Linear Quadratic Regulator (LQR) in a multi-agent network. The existing results often assume that agents can observe the global system state, which may be infeasible in large-scale systems due…

多智能体系统 · 计算机科学 2020-12-24 Hang Wang , Sen Lin , Hamid Jafarkhani , Junshan Zhang

In this paper, we investigate the robustness of stationary mean-field equilibria in the presence of model uncertainties, specifically focusing on infinite-horizon discounted cost functions. To achieve this, we initially establish…

系统与控制 · 电气工程与系统科学 2026-04-10 Uğur Aydın , Naci Saldi

The property of the communication network and the constraints on the strategic space are two factors that determine the complexity of the distributed Nash equilibrium (DNE) seeking problem. The DNE seeking problem of aggregative games has…

最优化与控制 · 数学 2025-02-26 Zhaocong Liu , Jie Huang

We consider a class of nonsmooth aggregative games over networks in stochastic regimes, where each player is characterized by a composite cost function $f_i+r_i$, $f_i$ is a smooth expectation-valued function dependent on its own strategy…

最优化与控制 · 数学 2024-06-28 Jinlong Lei , Uday V. Shanbhag , Jie Chen

Within the context of video games the notion of perfectly rational agents can be undesirable as it leads to uninteresting situations, where humans face tough adversarial decision makers. Current frameworks for stochastic games and…

人工智能 · 计算机科学 2019-01-09 Jordi Grau-Moya , Felix Leibfried , Haitham Bou-Ammar

Learning algorithm design for state-based games is investigated. A heuristic uncoupled learning algorithm, which is a two memory better reply with inertia dynamics, is proposed. Under certain reasonable conditions it is proved that for any…

最优化与控制 · 数学 2018-09-18 Changxi Li , Yu Xing , Fenghua He , Daizhan Cheng

We consider network aggregative games to model and study multi-agent populations in which each rational agent is influenced by the aggregate behavior of its neighbors, as specified by an underlying network. Specifically, we examine systems…

系统与控制 · 计算机科学 2015-06-26 Francesca Parise , Sergio Grammatico , Basilio Gentile , John Lygeros

We develop a flexible stochastic approximation framework for analyzing the long-run behavior of learning in games (both continuous and finite). The proposed analysis template incorporates a wide array of popular learning algorithms,…

计算机科学与博弈论 · 计算机科学 2023-07-04 Panayotis Mertikopoulos , Ya-Ping Hsieh , Volkan Cevher

This paper presents a game theoretic solution for joint channel allocation and power control in cognitive radio networks analyzed under the physical interference model. The objective is to find a distributed solution that maximizes the…

网络与互联网体系结构 · 计算机科学 2025-01-29 J. R. Gallego , M. Canales , J. Ortin

Adversarial training, a special case of multi-objective optimization, is an increasingly prevalent machine learning technique: some of its most notable applications include GAN-based generative modeling and self-play techniques in…

Reinforcement learning has been successful both empirically and theoretically in single-agent settings, but extending these results to multi-agent reinforcement learning in general-sum Markov games remains challenging. This paper studies…

机器学习 · 计算机科学 2026-04-07 Narim Jeong , Donghwan Lee