中文
相关论文

相关论文: Min-Max Q-Learning for Multi-Player Pursuit-Evasio…

200 篇论文

A defender-attacker-target problem with non-moving target is considered. This problem is modeled by a pursuit-evasion zero-sum differential game with linear dynamics and quadratic cost functional. In this game the pursuer is the defender,…

最优化与控制 · 数学 2018-03-06 Vladimir Turetsky , Valery Y. Glizer

In this paper, we explore the susceptibility of the independent Q-learning algorithms (a classical and widely used multi-agent reinforcement learning method) to strategic manipulation of sophisticated opponents in normal-form games played…

计算机科学与博弈论 · 计算机科学 2024-07-17 Yuksel Arslantas , Ege Yuceel , Muhammed O. Sayin

We study pursuit-evasion differential games between a faster pursuer moving in 3D space and an evader moving in a plane. We first extend the well-known Apollonius circle to 3D space, by which we construct the isochron for the considered two…

系统与控制 · 电气工程与系统科学 2022-03-01 Shuai Li , Chen Wang , Guangming Xie

This paper studies a variant of the multi-player reach-avoid game played between intruders and defenders with applications to perimeter defense. The intruder team tries to score by sending as many intruders as possible to the target area,…

系统与控制 · 电气工程与系统科学 2020-06-05 Daigo Shishika , James Paulos , Vijay Kumar

We consider a scenario where a team of two unmanned aerial vehicles (UAVs) pursue an evader UAV within an urban environment. Each agent has a limited view of their environment where buildings can occlude their field-of-view. Additionally,…

多智能体系统 · 计算机科学 2025-11-13 Addison Kalanther , Daniel Bostwick , Chinmay Maheshwari , Shankar Sastry

In this paper, we consider a territory guarding game involving pursuers, evaders and a target in an environment that contains obstacles. The goal of the evaders is to capture the target, while that of the pursuers is to capture the evaders…

人工智能 · 计算机科学 2019-10-18 Sagar Verma , Richa Verma , P. B. Sujit

Policy gradient methods are often applied to reinforcement learning in continuous multiagent games. These methods perform local search in the joint-action space, and as we show, they are susceptable to a game-theoretic pathology known as…

人工智能 · 计算机科学 2018-04-27 Ermo Wei , Drew Wicke , David Freelan , Sean Luke

We propose a hybrid approach that combines Hamilton-Jacobi (HJ) reachability and mixed-integer optimization for solving a reach-avoid game with multiple attackers and defenders. The reach-avoid game is an important problem with potential…

系统与控制 · 电气工程与系统科学 2023-09-26 Hanyang Hu , Minh Bui , Mo Chen

We consider a task of surveillance-evading path-planning in a continuous setting. An Evader strives to escape from a 2D domain while minimizing the risk of detection (and immediate capture). The probability of detection is path-dependent…

机器学习 · 计算机科学 2023-02-24 Dongping Qi , David Bindel , Alexander Vladimirsky

In addressing the challenge of exponential scaling with the number of agents we adopt a cluster-based representation to approximately solve asymmetric games of very many players. A cluster groups together agents with a similar "strategic…

计算机科学与博弈论 · 计算机科学 2012-06-18 Sevan G. Ficici , David C. Parkes , Avi Pfeffer

Recent studies in the spatial prisoner's dilemma games with reinforcement learning have shown that static agents can learn to cooperate through a diverse sort of mechanisms, including noise injection, different types of learning algorithms…

人工智能 · 计算机科学 2025-07-08 Gustavo C. Mangold , Heitor C. M. Fernandes , Mendeli H. Vainstein

This paper studies a planar multiplayer Homicidal Chauffeur reach-avoid differential game, where each pursuer is a Dubins car and each evader has simple motion. The pursuers aim to protect a goal region cooperatively from the evaders. Due…

计算机科学与博弈论 · 计算机科学 2021-07-13 Rui Yan , Ruiliang Deng , Haowen Lai , Weixian Zhang , Zongying Shi , Yisheng Zhong

The classical setting of optimal control theory assumes full knowledge of the process dynamics and the costs associated with every control strategy. The problem becomes much harder if the controller only knows a finite set of possible…

最优化与控制 · 数学 2019-08-27 Marc Aurèle Gilles , Alexander Vladimirsky

Unmanned Aerial Vehicles (UAVs), autonomously-guided aircraft, are widely used for tasks involving surveillance and reconnaissance. A version of the pursuit-evasion problems centered around UAVs and its variants has been extensively studied…

机器人学 · 计算机科学 2019-11-06 Loren Anderson , Sahitya Senapathy

Within the context of video games the notion of perfectly rational agents can be undesirable as it leads to uninteresting situations, where humans face tough adversarial decision makers. Current frameworks for stochastic games and…

人工智能 · 计算机科学 2019-01-09 Jordi Grau-Moya , Felix Leibfried , Haitham Bou-Ammar

This paper studies a variant of multi-player reach-avoid game played between intruders and defenders. The intruder team tries to score by sending as many intruders as possible to the target area, while the defender team tries to minimize…

系统与控制 · 电气工程与系统科学 2021-05-04 Daigo Shishika , Vijay Kumar

Underlying relationships among multiagent systems (MAS) in hazardous scenarios can be represented as game-theoretic models. In adversarial environments, the adversaries can be intentional or unintentional based on their needs and…

机器人学 · 计算机科学 2022-06-03 Qin Yang , Ramviyas Parasuraman

In this paper, we study inverse game theory (resp. inverse multiagent learning) in which the goal is to find parameters of a game's payoff functions for which the expected (resp. sampled) behavior is an equilibrium. We formulate these…

计算机科学与博弈论 · 计算机科学 2025-02-21 Denizalp Goktas , Amy Greenwald , Sadie Zhao , Alec Koppel , Sumitra Ganesh

We propose a game-based formulation for learning dimensionality-reducing representations of feature vectors, when only a prior knowledge on future prediction tasks is available. In this game, the first player chooses a representation, and…

机器学习 · 计算机科学 2024-03-12 Neria Uzan , Nir Weinberger

Two-player pursuit-evasion differential game and time optimal zero control problem in $\ell^2$ are considered. Optimal control for the corresponding zero control problem is found. A strategy for the pursuer that guarantees the solution for…

最优化与控制 · 数学 2023-02-06 Marks Ruziboev , Khudoyor Mamayusupov , Gafurjan Ibragimov , Adkham Khaitmetov