中文
相关论文

相关论文: Disturbance Decoupling for Gradient-based Multi-Ag…

200 篇论文

In this paper, a hierarchical one-leader-multi-followers game for a class of continuous-time nonlinear systems with disturbance is investigated by a novel policy iteration reinforcement learning technique in which, the game model consists…

系统与控制 · 电气工程与系统科学 2019-07-29 Mohammad reza Satouri , Hamed Kebriaei , Abolhassan Razminia , Mohammad javad Yazdanpanah

Learning problems commonly exhibit an interesting feedback mechanism wherein the population data reacts to competing decision makers' actions. This paper formulates a new game theoretic framework for this phenomenon, called "multi-player…

计算机科学与博弈论 · 计算机科学 2022-04-08 Adhyyan Narang , Evan Faulkner , Dmitriy Drusvyatskiy , Maryam Fazel , Lillian J. Ratliff

In this paper, we address the inverse problem for linear-quadratic differential non-cooperative games with output-feedback. Given players' stabilizing feedback laws, the goal is to find cost function parameters that lead to a game for which…

最优化与控制 · 数学 2024-10-27 Emin Martirosyan , Ming Cao

This paper presents a robust reinforcement learning algorithm called robust deterministic policy gradient (RDPG), which reformulates the H-infinity control problem as a two-player zero-sum dynamic game between a user and an adversary. The…

机器人学 · 计算机科学 2025-12-04 Taeho Lee , Donghwan Lee

Protecting quantum states from the decohering effects of the environment is of great importance for the development of quantum computation devices and quantum simulators. Here, we introduce a continuous dynamical decoupling protocol that…

量子物理 · 物理学 2019-04-10 İ. Yalçınkaya , B. Çakmak , G. Karpat , F. F. Fanchini

This study considers a federated learning setup where cost-sensitive and strategic agents train a learning model with a server. During each round, each agent samples a minibatch of training data and sends his gradient update. As an…

机器学习 · 计算机科学 2022-12-06 Abdullah Basar Akbay , Cihan Tepedelenlioglu

In this paper we introduce the novel framework of distributionally robust games. These are multi-player games where each player models the state of nature using a worst-case distribution, also called adversarial distribution. Thus each…

最优化与控制 · 数学 2017-07-25 Dario Bauso , Jian Gao , Hamidou Tembine

Multi-agent networked linear dynamic systems have attracted attention of researchers in power systems, intelligent transportation, and industrial automation. The agents might cooperatively optimize a global performance objective, resulting…

系统与控制 · 计算机科学 2017-01-12 Feier Lian , Aranya Chakrabortty , Alexandra Duel-Hallen

We study adaptive learning in a typical p-player game. The payoffs of the games are randomly generated and then held fixed. The strategies of the players evolve through time as the players learn. The trajectories in the strategy space…

经济学 · 定量金融 2018-04-09 James B. T. Sanders , J. Doyne Farmer , Tobias Galla

Current quantum computers suffer from noise that stems from interactions between the quantum system that constitutes the quantum device and its environment. These interactions can be suppressed through dynamical decoupling to reduce…

量子物理 · 物理学 2024-12-06 Arefur Rahman , Daniel J. Egger , Christian Arenz

We consider a problem where multiple agents must learn an action profile that maximises the sum of their utilities in a distributed manner. The agents are assumed to have no knowledge of either the utility functions or the actions and…

系统与控制 · 计算机科学 2016-03-31 Chithrupa Ramesh , Marius Schmitt , John Lygeros

In Federated Learning (FL), multiple clients jointly train a machine learning model by sharing gradient information, instead of raw data, with a server over multiple rounds. To address the possibility of information leakage in spite of…

机器学习 · 计算机科学 2025-08-12 Yashwant Krishna Pagoti , Arunesh Sinha , Shamik Sural

This paper addresses the challenge of limited observations in non-cooperative multi-agent systems where agents can have partial access to other agents' actions. We present the generalized individual Q-learning dynamics that combine…

计算机科学与博弈论 · 计算机科学 2024-09-05 Ahmed Said Donmez , Muhammed O. Sayin

Reinforcement-based learning has attracted considerable attention both in modeling human behavior as well as in engineering, for designing measurement- or payoff-based optimization schemes. Such learning schemes exhibit several advantages,…

机器学习 · 计算机科学 2025-11-26 Georgios C. Chasparis

In this work, we present a learning-based nonlinear $H^\infty$ control algorithm that guarantee system performance under learned dynamics and disturbance estimate. The Gaussian Process (GP) regression is utilized to update the nominal…

系统与控制 · 电气工程与系统科学 2021-07-12 Wei Sun , Theodore B. Trafalis

We train two neural networks adversarially to play static games. At each iteration, a row and column network observe a new random bimatrix game and output individual mixed strategies. The parameters of each network are independently updated…

理论经济学 · 经济学 2025-05-09 Daniele Condorelli , Massimiliano Furlan

Many problems in robotics involve multiple decision making agents. To operate efficiently in such settings, a robot must reason about the impact of its decisions on the behavior of other agents. Differential games offer an expressive…

系统与控制 · 电气工程与系统科学 2020-03-19 David Fridovich-Keil , Ellis Ratner , Lasse Peters , Anca D. Dragan , Claire J. Tomlin

This paper studies the stability and convergence properties of a class of multi-agent concurrent learning (CL) algorithms with momentum and restart. Such algorithms can be integrated as part of the estimation pipelines of data-enabled…

最优化与控制 · 数学 2024-06-24 Daniel E. Ochoa , Muhammad U. Javed , Xudong Chen , Jorge I. Poveda

We analyse the strategy equilibrium of dilemma games considering a payoff matrix affected by small and random perturbations on the off-diagonal. Notably, a recent work [1] reported that, while cooperation is sustained by perturbations…

物理与社会 · 物理学 2020-07-01 Marco A. Amaral , Marco A. Javarone

We consider control of heterogeneous players repeatedly playing an anti-coordination network game. In an anti-coordination game, each player has an incentive to differentiate its action from its neighbors. At each round of play, players…

系统与控制 · 计算机科学 2018-12-13 Ceyhun Eksin , Keith Paarporn