中文
相关论文

相关论文: Min-Max Q-Learning for Multi-Player Pursuit-Evasio…

200 篇论文

Although learning has found wide application in multi-agent systems, its effects on the temporal evolution of a system are far from understood. This paper focuses on the dynamics of Q-learning in large-scale multi-agent systems modeled as…

多智能体系统 · 计算机科学 2022-03-04 Shuyue Hu , Chin-Wing Leung , Ho-fung Leung , Harold Soh

We consider a class of pursuit-evasion differential games in which the evader has continuous access to the pursuer's location, but not vice-versa. There is a remote sensor (e.g., a radar station) that can sense the evader's location upon a…

系统与控制 · 电气工程与系统科学 2023-07-04 Dipankar Maity

We consider surveillance-evasion differential games, where a pursuer must try to constantly maintain visibility of a moving evader. The pursuer loses as soon as the evader becomes occluded. Optimal controls for game can be formulated as a…

人工智能 · 计算机科学 2022-03-29 Louis Ly , Yen-Hsi Richard Tsai

We consider a model of two competing microswimming agents engaged in a pursue-evasion task within a low-Reynolds-number environment. Agents can only perform simple maneuvers and sense hydrodynamic disturbances, which provide ambiguous…

流体动力学 · 物理学 2022-03-04 Francesco Borra , Luca Biferale , Massimo Cencini , Antonio Celani

We study a pursuit-evasion problem which can be viewed as an extension of the keep-away game. In the game, pursuer(s) will attempt to intersect or catch the evader, while the evader can visit a fixed set of locations, which we denote as the…

机器人学 · 计算机科学 2022-06-17 Weifu Wang , Ping Li

This paper presents a novel strategy for a multi-agent pursuit-evasion game involving multiple faster pursuers with heterogenous speeds and a single slower evader. We define a geometric region, the evader's safe-reachable set, as the…

多智能体系统 · 计算机科学 2025-11-24 Kamal Mammadov , Damith C. Ranasinghe

Pursuit-evasion is a multi-agent sequential decision problem wherein a group of agents known as pursuers coordinate their traversal of a spatial domain to locate an agent trying to evade them. Pursuit evasion problems arise in a number of…

机器学习 · 计算机科学 2018-11-13 Zhen Li , Nicholas J. Meyer , Eric B. Laber , Robert Brigantic

This paper addresses a multi-pursuer single-evader pursuit-evasion game where the free-moving evader moves faster than the pursuers. Most of the existing works impose constraints on the faster evader such as limited moving area and moving…

系统与控制 · 电气工程与系统科学 2020-05-19 Xu Fang , Chen Wang , Lihua Xie , Jie Chen

In this work, we develop a reinforcement learning protocol for a multiagent coordination task in a discrete state and action space: an iterated prisoner's dilemma game extended into a team based, winner-take all tournament, which forces the…

计算机科学与博弈论 · 计算机科学 2018-06-18 Aaron Goodman

Achieving convergence of multiple learning agents in general $N$-player games is imperative for the development of safe and reliable machine learning (ML) algorithms and their application to autonomous systems. Yet it is known that, outside…

计算机科学与博弈论 · 计算机科学 2023-01-24 Aamal Abbas Hussain , Francesco Belardinelli , Georgios Piliouras

In this paper we study a linear pursuit differential game described by an infinite system of first-order differential equations in Hilbert space. The control functions of players are subject to geometric constraints. The pursuer attempts to…

最优化与控制 · 数学 2020-02-19 Gafurjan Ibragimov , Massimiliano Ferrara , Idham Arif Alias , Mehdi Salimi

Reinforcement learning agents in complex game environments often suffer from sparse rewards, training instability, and poor sample efficiency. This paper presents a hybrid training approach that combines offline imitation learning with…

机器学习 · 计算机科学 2025-09-19 Thomas Ackermann , Moritz Spang , Hamza A. A. Gardi

In this paper, a novel decentralized intelligent adaptive optimal strategy has been developed to solve the pursuit-evasion game for massive Multi-Agent Systems (MAS) under uncertain environment. Existing strategies for pursuit-evasion games…

系统与控制 · 电气工程与系统科学 2020-08-10 Zejian Zhou , Hao Xu

This paper introduces a new family of pursuit strategies for multi-pursuer single-evader games in a planar environment. They leverage conditions under which the minimum-time solution of the game becomes equivalent to that of a suitable…

系统与控制 · 电气工程与系统科学 2024-10-11 Marco Casini , Andrea Garulli

In a pursuit-evasion game, a team of pursuers attempt to capture an evader. The players alternate turns, move with equal speed, and have full information about the state of the game. We consider the most restictive capture condition: a…

度量几何 · 数学 2015-05-05 Andrew Beveridge , Yiqing Cai

We study a pursuit-evasion game between two players with car-like dynamics and sensing limitations by formalizing it as a partially observable stochastic zero-sum game. The partial observability caused by the sensing constraints is…

机器人学 · 计算机科学 2025-06-17 Burak M. Gonultas , Volkan Isler

In practical application, the pursuit-evasion game (PEG) often involves multiple complex and conflicting objectives. The single-objective reinforcement learning (RL) usually focuses on a single optimization objective, and it is difficult to…

系统与控制 · 电气工程与系统科学 2025-03-11 Penglin Hu , Chunhui Zhao , Quan Pan

Here we present a decomposition technique for a class of differential games. The technique consists in a decomposition of the target set which produces, for geometrical reasons, a decomposition in the dimensionality of the problem. Using…

最优化与控制 · 数学 2013-03-14 Adriano Festa , Richard B. Vinter

A new surveillance-evasion differential game is posed and solved in which an agile pursuer (the prying pedestrian) seeks to remain within a given surveillance range of a less agile evader that aims to escape. In contrast to previous…

系统与控制 · 电气工程与系统科学 2025-05-13 Philipp Braun , Timothy L. Molloy , Iman Shames

We employ the Deep Q-Learning algorithm with Experience Replay to train an agent capable of achieving a high-level of play in the L-Game while self-learning from low-dimensional states. We also employ variable batch size for training in…

机器学习 · 计算机科学 2018-02-20 Petros Giannakopoulos , Yannis Cotronis