English
Related papers

Related papers: Disturbance Decoupling for Gradient-based Multi-Ag…

200 papers

Policy optimization has drawn increasing attention in reinforcement learning, particularly in the context of derivative-free methods for linear quadratic regulator (LQR) problems with unknown dynamics. This paper focuses on characterizing…

Optimization and Control · Mathematics 2025-06-17 Weijian Li , Panagiotis Kounatidis , Zhong-Ping Jiang , Andreas A. Malikopoulos

Multi-agent reinforcement learning in mixed-motive settings presents a fundamental challenge: agents must balance individual interests with collective goals, which are neither fully aligned nor strictly opposed. To address this, reward…

Multiagent Systems · Computer Science 2025-08-26 Woojun Kim , Katia Sycara

Self-interested individuals often fail to cooperate, posing a fundamental challenge for multi-agent learning. How can we achieve cooperation among self-interested, independent learning agents? Promising recent work has shown that in certain…

Games research and industry have developed a solid understanding of how to design engaging, playful experiences that draws players in for hours and causes them to lose their sense of time. While these designs can provide enjoyable…

Human-Computer Interaction · Computer Science 2023-03-28 Meshaiel Alsheail , Dmitry Alexandrovskz , Kathrin Gerling

In this paper, we study the problem of the distributed Nash equilibrium seeking of N-player games over jointly strongly connected switching networks. The action of each player is governed by a class of uncertain nonlinear systems. Our…

Optimization and Control · Mathematics 2024-11-05 Jie Huang

The paper considers a class of multi-agent Markov decision processes (MDPs), in which the network agents respond differently (as manifested by the instantaneous one-stage random costs) to a global controlled state and the control actions of…

Machine Learning · Statistics 2015-06-04 Soummya Kar , Jose' M. F. Moura , H. Vincent Poor

In this paper, we address the problem of a two-player linear quadratic differential game with incomplete information, a scenario commonly encountered in multi-agent control, human-robot interaction (HRI), and approximation methods for…

Systems and Control · Electrical Eng. & Systems 2025-04-25 Seyed Yousef Soltanian , Wenlong Zhang

Gradient descent (GD) and stochastic gradient descent (SGD) have been widely used in a large number of application domains. Therefore, understanding the dynamics of GD and improving its convergence speed is still of great importance. This…

Machine Learning · Computer Science 2024-09-11 Jinwei Zhao , Marco Gori , Alessandro Betti , Stefano Melacci , Hongtao Zhang , Jiedong Liu , Xinhong Hei

This paper considers a class of mean field linear-quadratic-Gaussian (LQG) games with model uncertainty. The drift term in the dynamics of the agents contains a common unknown function. We take a robust optimization approach where a…

Optimization and Control · Mathematics 2017-01-03 Jianhui Huang , Minyi Huang

In this paper, Nash equilibrium seeking among a network of players is considered. Different from many existing works on Nash equilibrium seeking in non-cooperative games, the players considered in this paper cannot directly observe the…

Optimization and Control · Mathematics 2017-03-28 Maojiao Ye , Guoqiang Hu

An alternating graph is a directed graph whose vertex set is partitioned into two classes, existential and universal. This forms the basic arena for a plethora of infinite duration two-player games where Player~$\square$ and~$\ocircle$…

Data Structures and Algorithms · Computer Science 2025-08-14 Carlo Comin , Romeo Rizzi

Conventional gradient descent methods compute the gradients for multiple variables through the partial derivative. Treating the coupled variables independently while ignoring the interaction, however, leads to an insufficient optimization…

Machine Learning · Computer Science 2021-06-22 Runqi Wang , Baochang Zhang , Li'an Zhuo , Qixiang Ye , David Doermann

We study stochastic effects on the lagging anchor dynamics, a reinforcement learning algorithm used to learn successful strategies in iterated games, which is known to converge to Nash points in the absence of noise. The dynamics is…

Adaptation and Self-Organizing Systems · Physics 2012-04-20 James B. T. Sanders , Tobias Galla , Jonathan Shapiro

We consider multi-agent decision making where each agent's cost function depends on all agents' strategies. We propose a distributed algorithm to learn a Nash equilibrium, whereby each agent uses only obtained values of her cost function at…

Multiagent Systems · Computer Science 2019-04-04 Tatiana Tatarenko , Maryam Kamgarpour

An open problem in linear quadratic (LQ) games has been characterizing the Nash equilibria. This problem has renewed relevance given the surge of work on understanding the convergence of learning algorithms in dynamic games. This paper…

Computer Science and Game Theory · Computer Science 2025-04-18 Giulio Salizzoni , Reda Ouhamma , Maryam Kamgarpour

We present the first massively distributed architecture for deep reinforcement learning. This architecture uses four main components: parallel actors that generate new behaviour; parallel learners that are trained from stored experience; a…

Learning in stochastic games is arguably the most standard and fundamental setting in multi-agent reinforcement learning (MARL). In this paper, we consider decentralized MARL in stochastic games in the non-asymptotic regime. In particular,…

Computer Science and Game Theory · Computer Science 2021-12-17 Zuguang Gao , Qianqian Ma , Tamer Başar , John R. Birge

We study how strategic interaction can arise from controlled quantum dynamics rather than being imposed as an external mathematical structure. We introduce a class of interaction-defined quantum games in which players are represented by…

Quantum Physics · Physics 2026-04-23 Rashid Ahmad

We study a differential game that governs the moderate-deviation heavy-traffic asymptotics of a multiclass single-server queueing control problem with a risk-sensitive cost. We consider a cost set on a finite but sufficiently large time…

Probability · Mathematics 2018-05-02 Rami Atar , Asaf Cohen

We formulate a new class of two-person zero-sum differential games, in a stochastic setting, where a specification on a target terminal state distribution is imposed on the players. We address such added specification by introducing…

Systems and Control · Electrical Eng. & Systems 2019-09-13 Yongxin Chen , Tryphon T. Georgiou , Michele Pavon