中文
相关论文

相关论文: Min-Max Q-Learning for Multi-Player Pursuit-Evasio…

200 篇论文

This paper addresses the challenge of limited observations in non-cooperative multi-agent systems where agents can have partial access to other agents' actions. We present the generalized individual Q-learning dynamics that combine…

计算机科学与博弈论 · 计算机科学 2024-09-05 Ahmed Said Donmez , Muhammed O. Sayin

We focus on the important problem of emergency evacuation, which clearly could benefit from reinforcement learning that has been largely unaddressed. Emergency evacuation is a complex task which is difficult to solve with reinforcement…

人工智能 · 计算机科学 2019-05-30 Jivitesh Sharma , Per-Arne Andersen , Ole-Chrisoffer Granmo , Morten Goodwin

We study the class of reach-avoid dynamic games in which multiple agents interact noncooperatively, and each wishes to satisfy a distinct target criterion while avoiding a failure criterion. Reach-avoid games are commonly used to express…

系统与控制 · 电气工程与系统科学 2022-03-03 Dennis R. Anthony , Duy P. Nguyen , David Fridovich-Keil , Jaime F. Fisac

We study a simple motion differential game of many pursuers and one evader in the plane. We give a nonempty closed convex set in the plane, and the pursuers and evader move on this set. They cannot leave this set during the game. Control…

最优化与控制 · 数学 2015-05-04 Idham Arif Alias , Gafurjan Ibragimov , Massimiliano Ferrara , Mehdi Salimi , Mansor Monsi

In this paper, we consider the problem of controlling an underactuated system in unknown, and potentially adversarial environments. The emphasis will be on autonomous aerial vehicles, modelled by Dubins dynamics. The proposed control law is…

多智能体系统 · 计算机科学 2019-09-04 Shashwat Shivam , Aris Kanellopoulos , Kyriakos G. Vamvoudakis , Yorai Wardi

This report aims to survey multi-agent Q-Learning algorithms, analyze different game theory frameworks used, address each framework's applications, and report challenges and future directions. The target application for this study is…

多智能体系统 · 计算机科学 2021-05-07 Arvin Tashakori

In this paper we investigate a differential game in which countably many dynamical objects pursue a single one. All the players perform simple motions. The duration of the game is fixed. The controls of a group of pursuers are subject to…

最优化与控制 · 数学 2014-10-10 Mehdi Salimi , Gafurjan Ibragimov , Stefan Siegmund , Somayeh Sharifi

We develop a method based on computer algebra systems to represent the mutual pure strategy best-response dynamics of symmetric two-player, two-action repeated games played by players with a one-period memory. We apply this method to the…

动力系统 · 数学 2022-10-04 Janusz M Meylahn , Lars Janssen

Solutions to pursuit-evasion and surveillance-evasion differential games are typically computed and expressed using open-loop representations, with the synthesis of feedback strategies significantly less common. We propose a numerical…

系统与控制 · 电气工程与系统科学 2026-05-07 Philipp Braun , Timothy L. Molloy , Gal Barkai , Iman Shames

Deep Q-learning has achieved significant success in single-agent decision making tasks. However, it is challenging to extend Q-learning to large-scale multi-agent scenarios, due to the explosion of action space resulting from the complex…

多智能体系统 · 计算机科学 2019-10-14 Ming Zhou , Yong Chen , Ying Wen , Yaodong Yang , Yufeng Su , Weinan Zhang , Dell Zhang , Jun Wang

Training agents in multi-agent competitive games presents significant challenges due to their intricate nature. These challenges are exacerbated by dynamics influenced not only by the environment but also by opponents' strategies. Existing…

机器学习 · 计算机科学 2023-08-22 The Viet Bui , Tien Mai , Thanh Hong Nguyen

We extend the adversarial/non-stochastic multi-play multi-armed bandit (MPMAB) to the case where the number of arms to play is variable. The work is motivated by the fact that the resources allocated to scan different critical locations in…

机器学习 · 计算机科学 2021-10-28 Yiyang Wang , Neda Masoud

This paper considers a pursuit-evasion scenario among three agents -- an evader, a pursuer, and a defender. We design cooperative guidance laws for the evader and the defender team to safeguard the evader from an attacking pursuer. Unlike…

系统与控制 · 电气工程与系统科学 2022-01-19 Abhinav Sinha , Shashi Ranjan Kumar , Dwaipayan Mukherjee

Traditional game-theoretic research for security applications primarily focuses on the allocation of external protection resources to defend targets. This work puts forward the study of a new class of games centered around strategically…

计算机科学与博弈论 · 计算机科学 2024-10-29 Niclas Boehmer , Minbiao Han , Haifeng Xu , Milind Tambe

The behaviour of multi-agent learning in many player games has been shown to display complex dynamics outside of restrictive examples such as network zero-sum games. In addition, it has been shown that convergent behaviour is less likely to…

计算机科学与博弈论 · 计算机科学 2023-07-27 Aamal Hussain , Dan Leonte , Francesco Belardinelli , Georgios Piliouras

This paper presents a multiplayer Homicidal Chauffeur reach-avoid differential game, which involves Dubins-car pursuers and simple-motion evaders. The goal of the pursuers is to cooperatively protect a planar convex region from the evaders,…

系统与控制 · 电气工程与系统科学 2023-12-25 Rui Yan , Xiaoming Duan , Rui Zou , Xin He , Zongying Shi , Francesco Bullo

This paper investigates a pursuit-evasion problem involving three agents: a pursuer, an evader, and a defender. Cooperative guidance laws are developed for the evader-defender team that guarantee interception of the pursuer by the defender…

系统与控制 · 电气工程与系统科学 2025-09-16 Saurabh Kumar , Shashi Ranjan Kumar , Abhinav Sinha

When modeling robot interactions as Nash equilibrium problems, it is desirable to place coupled constraints which restrict these interactions to be safe and acceptable (for instance, to avoid collisions). Such games are continuous with…

计算机科学与博弈论 · 计算机科学 2025-06-03 Mel Krusniak , Forrest Laine

In this paper, we present a game-theoretic feedback terminal guidance law for an autonomous, unpowered hypersonic pursuit vehicle that seeks to intercept an evading ground target whose motion is constrained in a one-dimensional space. We…

系统与控制 · 电气工程与系统科学 2022-01-14 Yoonjae Lee , Efstathios Bakolas , Maruthi R. Akella

Video game playing is an extremely structured domain where algorithmic decision-making can be tested without adverse real-world consequences. While prevailing methods rely on image inputs to avoid the problem of hand-crafting state space…

机器学习 · 计算机科学 2024-09-24 Abhishek Jaiswal , Nisheeth Srivastava