中文
相关论文

相关论文: Disturbance Decoupling for Gradient-based Multi-Ag…

200 篇论文

The effects of a collection of classical two-level charge fluctuators on the coherence of a dynamically-decoupled qubit are studied. Distinct dynamics are found at different qubit working positions. Exact analytical formulae are derived at…

介观与纳米尺度物理 · 物理学 2015-10-21 Guy Ramon

With increasing game size, a problem of computational complexity arises. This is especially true in real world problems such as in social systems, where there is a significant population of players involved in the game, and the complexity…

计算机科学与博弈论 · 计算机科学 2016-09-12 Tatsuya Iwase , Takahiro Shiga

Individual agents in a multi-agent system (MAS) may have decoupled open-loop dynamics, but a cooperative control objective usually results in coupled closed-loop dynamics thereby making the control design computationally expensive. The…

系统与控制 · 电气工程与系统科学 2021-03-09 Gangshan Jing , He Bai , Jemin George , Aranya Chakrabortty

A dynamical decoupling method is presented which is based on embedding a deterministic decoupling scheme into a stochastic one. This way it is possible to combine the advantages of both methods and to increase the suppression of undesired…

量子物理 · 物理学 2007-05-23 Oliver Kern , Gernot Alber

We discuss the effect of correlated noise on the robustness of quantum coherent phenomena. First we consider a simple, toy model to illustrate the effect of such correlations on the decoherence process. Then we show how decoherence rates…

量子物理 · 物理学 2007-05-23 Chiu Fan Lee , Neil F. Johnson , Ferney Rodriguez , Luis Quiroga

Constrained Markov games offer a formal mathematical framework for modeling multi-agent reinforcement learning problems where the behavior of the agents is subject to constraints. In this work, we focus on the recently introduced class of…

机器学习 · 计算机科学 2024-02-29 Philip Jordan , Anas Barakat , Niao He

In recent years, multi-player multi-armed bandits (MP-MAB) have been extensively studied due to their wide applications in cognitive radio networks and Internet of Things systems. While most existing research on MP-MAB focuses on…

机器学习 · 计算机科学 2025-10-01 Jingqi Fan , Canzhe Zhao , Shuai Li , Siwei Wang

We study binary coordination games over graphs under log-linear learning when neighbor actions are conveyed through explicit noisy communication links. Each edge is modeled as either a binary symmetric channel (BSC) or a binary erasure…

系统与控制 · 电气工程与系统科学 2026-01-29 Emrah Akyol , Marcos Vasconcelos

We address safe multi-robot interaction under uncertainty. In particular, we formulate a chance-constrained linear quadratic Gaussian game with coupling constraints and system uncertainties. We find a tractable reformulation of the game and…

机器人学 · 计算机科学 2025-08-15 Kai Ren , Giulio Salizzoni , Mustafa Emre Gürsoy , Maryam Kamgarpour

We examine the effect of multilevels on decoherence and dephasing properties of a quantum system consisting of a non-ideal two level subspace, identified as the qubit and a finite set of higher energy levels above this qubit subspace. The…

介观与纳米尺度物理 · 物理学 2007-05-23 T. Hakioglu , K. Savran

Achieving high-precision control for robotic systems is hindered by the low-fidelity dynamical model and external disturbances. Especially, the intricate coupling between internal uncertainties and external disturbances further exacerbates…

机器人学 · 计算机科学 2026-02-05 Jindou Jia , Meng Wang , Zihan Yang , Bin Yang , Yuhang Liu , Kexin Guo , Xiang Yu

This paper proposes a differentiable robust LQR layer for reinforcement learning and imitation learning under model uncertainty and stochastic dynamics. The robust LQR layer can exploit the advantages of robust optimal control and…

机器人学 · 计算机科学 2021-06-11 Ngo Anh Vien , Gerhard Neumann

Multiagent systems appear in most social, economical, and political situations. In the present work we extend the Deep Q-Learning Network architecture proposed by Google DeepMind to multiagent environments and investigate how two agents…

人工智能 · 计算机科学 2015-11-30 Ardi Tampuu , Tambet Matiisen , Dorian Kodelja , Ilya Kuzovkin , Kristjan Korjus , Juhan Aru , Jaan Aru , Raul Vicente

This work describes a technique for active rejection of multiple independent and time-correlated stochastic disturbances for a nonlinear flexible inverted pendulum with cart system with uncertain model parameters. The control law is…

系统与控制 · 电气工程与系统科学 2024-04-09 Vincent W. Hill

We consider a class of linear-quadratic-Gaussian mean-field games with a major agent and considerable heterogeneous minor agents in the presence of mean-field interactions. The individual admissible controls are constrained in closed convex…

最优化与控制 · 数学 2017-10-10 Ying Hu , Jianhui Huang , Tianyang Nie

The effect of entanglement and correlated noise in a four-player quantum Minority game is investigated. Different time correlated quantum memory channels are considered to analyze the Nash equilibrium payoff of the 1st player. It is seen…

量子物理 · 物理学 2013-05-10 M. Ramzan , M. K. Khan

This paper studies the finite-time horizon Markov games where the agents' dynamics are decoupled but the rewards can possibly be coupled across agents. The policy class is restricted to local policies where agents make decisions using their…

计算机科学与博弈论 · 计算机科学 2023-04-11 Runyu Zhang , Yuyang Zhang , Rohit Konda , Bryce Ferguson , Jason Marden , Na Li

In this paper, we present a framework for multi-agent learning in a nonstationary dynamic network environment. More specifically, we examine projected gradient play in smooth monotone repeated network games in which the agents'…

计算机科学与博弈论 · 计算机科学 2024-08-13 Feras Al Taha , Kiran Rokade , Francesca Parise

Decoupled learning is a branch of model parallelism which parallelizes the training of a network by splitting it depth-wise into multiple modules. Techniques from decoupled learning usually lead to stale gradient effect because of their…

机器学习 · 计算机科学 2020-12-08 Huiping Zhuang , Zhiping Lin , Kar-Ann Toh

Reinforcement Learning (RL) agents have great successes in solving tasks with large observation and action spaces from limited feedback. Still, training the agents is data-intensive and there are no guarantees that the learned behavior is…

人工智能 · 计算机科学 2021-10-20 Helge Spieker