中文
相关论文

相关论文: Disturbance Decoupling for Gradient-based Multi-Ag…

200 篇论文

Generally, Reinforcement Learning (RL) agent updates its policy by repetitively interacting with the environment, contingent on the received rewards to observed states and undertaken actions. However, the environmental disturbance, commonly…

人工智能 · 计算机科学 2024-11-07 Wei Geng , Baidi Xiao , Rongpeng Li , Ning Wei , Dong Wang , Zhifeng Zhao

We propose a novel independent and payoff-based learning framework for stochastic games that is model-free, game-agnostic, and gradient-free. The learning dynamics follow a best-response-type actor-critic architecture, where agents update…

机器学习 · 计算机科学 2026-02-03 Ahmed Said Donmez , Yuksel Arslantas , Muhammed O. Sayin

This paper studies attack detection for discrete-time linear systems with stochastic process noise that produce both a vulnerable (i.e., attackable) linear measurement and a secured (i.e., unattackable) quadratic measurement. The motivating…

最优化与控制 · 数学 2026-02-05 Muyan Jiang , Anil Aswani

We analyze offline designs of linear quadratic regulator (LQR) strategies with uncertain disturbances. First, we consider the scenario where the exogenous variable can be estimated in a controlled environment, and subsequently, consider a…

系统与控制 · 电气工程与系统科学 2025-09-26 Sayak Mukherjee , Ramij R. Hossain , Mahantesh Halappanavar

We have studied an evolutionary prisoner's dilemma game with players located on two types of random regular graphs with a degree of 4. The analysis is focused on the effects of payoffs and noise (temperature) on the maintenance of…

统计力学 · 物理学 2009-11-11 Jeromos Vukov , György Szabó , Attila Szolnoki

We analyze quantum game with correlated noise through generalized quantization scheme. Four different combinations on the basis of entanglement of initial quantum state and the measurement basis are analyzed. It is shown that the advantage…

量子物理 · 物理学 2009-11-13 Ahmad Nawaz , A. H. Toor

Resilience to noise and to decoherence processes is an important ingredient for the implementation of quantum information processing, and quantum technologies. To this end, techniques such as pulsed and continuous dynamical decoupling have…

量子物理 · 物理学 2016-12-02 Itsik Cohen , Nati Aharon , Alex Retzker

Quantum game theory is a rapidly evolving subject that extends beyond physics. In this research work, a schematic picture of quantum game theory has been provided with the help of the famous game Prisoners' Dilemma. It has been considered…

量子物理 · 物理学 2021-03-31 Kaushik Naskar

In multi-agent autonomous systems, deception is a fundamental concept which characterizes the exploitation of unbalanced information to mislead victims into choosing oblivious actions. This effectively alters the system's long term…

系统与控制 · 电气工程与系统科学 2025-08-27 Michael Tang , Miroslav Krstic , Jorge Poveda

In this work, we study stochastic non-cooperative games, where only noisy black-box function evaluations are available to estimate the cost function for each player. Since each player's cost function depends on both its own decision…

计算机科学与博弈论 · 计算机科学 2025-11-18 Haidong Li , Anzhi Sheng , Yijie Peng , Long Wang

Learning in multi-player games can model a large variety of practical scenarios, where each player seeks to optimize its own local objective function, which at the same time relies on the actions taken by others. Motivated by the frequent…

最优化与控制 · 数学 2023-09-08 Yuanhanqing Huang , Jianghai Hu

In repeated games, such as auctions, players rely on autonomous learning agents to choose their actions. We study settings in which players have their agents make monetary transfers to other agents during play at their own expense, in order…

计算机科学与博弈论 · 计算机科学 2026-02-12 Yoav Kolumbus , Joe Halpern , Éva Tardos

This paper introduces a reinforcement learning framework that enables controllable and diverse player behaviors without relying on human gameplay data. Existing approaches often require large-scale player trajectories, train separate models…

机器学习 · 计算机科学 2025-12-12 Atahan Cilan , Atay Özgövde

Motivated by recent works addressing adversarial attacks on deep reinforcement learning, a deception attack on linear quadratic Gaussian control is studied in this paper. In the considered attack model, the adversary can manipulate the…

系统与控制 · 电气工程与系统科学 2020-09-11 Zuxing Li , György Dán , Dong Liu

Competitive non-cooperative online decision-making agents whose actions increase congestion of scarce resources constitute a model for widespread modern large-scale applications. To ensure sustainable resource behavior, we introduce a novel…

最优化与控制 · 数学 2020-10-22 Ezra Tampubolon , Holger Boche

The framework of multi-agent learning explores the dynamics of how individual agent strategies evolve in response to the evolving strategies of other agents. Of particular interest is whether or not agent strategies converge to well known…

计算机科学与博弈论 · 计算机科学 2023-11-21 Sarah A. Toonsi , Jeff S. Shamma

Privacy concerns in distributed learning often lead clients to return intentionally altered gradient information. We consider the problem of learning convex and $L$-smooth functions under adversarial gradient perturbation, where a client's…

机器学习 · 计算机科学 2026-05-06 Nawapon Sangsiri , Yufei Tao

We study testable implications of multiple equilibria in discrete games with incomplete information. Unlike de Paula and Tang (2012), we allow the players' private signals to be correlated. In static games, we leverage independence of…

计量经济学 · 经济学 2020-12-03 Aureo de Paula , Xun Tang

In addition to the traditional two-level system, the three-level system serves as another important elemental building block for the manipulation of qubits. However, the quantum information processing in the three-level system is also…

量子物理 · 物理学 2025-09-19 P. Z. Zhao , Lei Qiao

Individuals, or organizations, cooperate with or compete against one another in a wide range of practical situations. Such strategic interactions are often modeled as games played on networks, where an individual's payoff depends not only…

计算机科学与博弈论 · 计算机科学 2020-09-22 Yan Leng , Xiaowen Dong , Junfeng Wu , Alex Pentland