中文
相关论文

相关论文: Iterative Best Response for Multi-Body Asset-Guard…

200 篇论文

Given a two-dimensional polygonal space, the multi-robot visibility-based pursuit-evasion problem tasks several pursuer robots with the goal of establishing visibility with an arbitrarily fast evader. The best known complete algorithm for…

机器人学 · 计算机科学 2021-04-12 Trevor Olsen , Anne M. Tumlin , Nicholas M. Stiffler , Jason M. O'Kane

We present a semi-infinite program (SIP) solver for trajectory optimizations of general articulated robots. These problems are more challenging than standard Nonlinear Program (NLP) by involving an infinite number of non-convex, collision…

机器人学 · 计算机科学 2023-11-06 Duo Zhang , Chen Liang , Xifeng Gao , Kui Wu , Zherong Pan

Conditional Value at Risk (CVaR) is widely used to account for the preferences of a risk-averse agent in the extreme loss scenarios. To study the effectiveness of randomization in interdiction games with an interdictor that is both risk and…

计算机科学与博弈论 · 计算机科学 2020-03-19 Utsav Sadana , Erick Delage

Deriving competitive, distributed solutions to multi-agent problems is crucial for many developing application domains; Game theory has emerged as a useful framework to design such algorithms. However, much of the attention within this…

系统与控制 · 电气工程与系统科学 2024-06-27 Rohit Konda , Rahul Chandan , David Grimsman , Jason R. Marden

We propose a novel framework for robust dynamic games with nonlinear dynamics corrupted by state-dependent additive noise, and nonlinear agent-specific and shared constraints. Leveraging system-level synthesis (SLS), each agent designs a…

系统与控制 · 电气工程与系统科学 2026-04-08 Shuyu Zhan , Chih-Yuan Chiu , Antoine P. Leeman , Glen Chou

We study automated intrusion response for an IT infrastructure and formulate the interaction between an attacker and a defender as a partially observed stochastic game. To solve the game we follow an approach where attack and defense…

系统与控制 · 电气工程与系统科学 2024-04-23 Kim Hammar , Rolf Stadler

Repeated games are difficult to analyze, especially when agents play mixed strategies. We study one-memory strategies in iterated prisoner's dilemma, then generalize the result to k-memory strategies in repeated games. Our result shows that…

计算机科学与博弈论 · 计算机科学 2019-02-26 Shiheng Wang , Fangzhen Lin

This paper examines the convergence behaviour of simultaneous best-response dynamics in random potential games. We provide a theoretical result showing that, for two-player games with sufficiently many actions, the dynamics converge quickly…

计算机科学与博弈论 · 计算机科学 2025-05-19 Galit Ashkenazi-Golan , Domenico Mergoni Cecchelli , Edward Plumb

We focus on adversarial patrolling games on arbitrary graphs, where the Defender can control a mobile resource, the targets are alarmed by an alarm system, and the Attacker can observe the actions of the mobile resource of the Defender and…

人工智能 · 计算机科学 2018-06-20 Giuseppe De Nittis , Nicola Gatti

Traditional game-theoretic research for security applications primarily focuses on the allocation of external protection resources to defend targets. This work puts forward the study of a new class of games centered around strategically…

计算机科学与博弈论 · 计算机科学 2024-10-29 Niclas Boehmer , Minbiao Han , Haifeng Xu , Milind Tambe

This paper proposes real-time sequential convex programming (RTSCP), a method for solving a sequence of nonlinear optimization problems depending on an online parameter. We provide a contraction estimate for the proposed method and, as a…

最优化与控制 · 数学 2015-03-19 Tran Dinh Quoc , Carlo Savorgnan , Moritz Diehl

This paper introduces an evolutionary dynamics based on imitate the better realization (IBR) rule. Under this rule, agents in a population game imitate the strategy of a randomly chosen opponent whenever the opponent`s realized payoff is…

理论经济学 · 经济学 2019-07-10 George Loginov

In this paper, we present a game-theoretic feedback terminal guidance law for an autonomous, unpowered hypersonic pursuit vehicle that seeks to intercept an evading ground target whose motion is constrained in a one-dimensional space. We…

系统与控制 · 电气工程与系统科学 2022-01-14 Yoonjae Lee , Efstathios Bakolas , Maruthi R. Akella

This paper presents a novel control strategy to herd groups of non-cooperative evaders by means of a team of robotic herders. In herding problems, the motion of the evaders is typically determined by strongly nonlinear and heterogeneous…

系统与控制 · 电气工程与系统科学 2022-06-14 Eduardo Sebastián , Eduardo Montijano , Carlos Sagüés

We study reward maximisation in a wide class of structured stochastic multi-armed bandit problems, where the mean rewards of arms satisfy some given structural constraints, e.g. linear, unimodal, sparse, etc. Our aim is to develop methods…

机器学习 · 统计学 2020-07-03 Rémy Degenne , Han Shao , Wouter M. Koolen

We describe an algorithm for computing best response strategies in a class of two-player infinite games of incomplete information, defined by payoffs piecewise linear in agents' types and actions, conditional on linear comparisons of…

计算机科学与博弈论 · 计算机科学 2012-07-19 Daniel Reeves , Michael P. Wellman

This paper aims to solve the optimal strategy against a well-known adaptive algorithm, the Hedge algorithm, in a finitely repeated $2\times 2$ zero-sum game. In the literature, related theoretical results are very rare. To this end, we make…

最优化与控制 · 数学 2023-12-18 Xinxiang Guo , Yifen Mu

We present a nonlinear non-convex model predictive control approach to solving a real-world labyrinth game. We introduce adaptive nonlinear constraints, representing the non-convex obstacles within the labyrinth. Our method splits the…

机器人学 · 计算机科学 2025-02-11 Johannes Gaber , Thomas Bi , Raffaello D'Andrea

We study a finite-horizon differential game of pursuit-evasion like, between a single player and a mass of agents. The player and the mass directly control their own evolution, which for the mass is given by a first order PDE of transport…

最优化与控制 · 数学 2025-02-28 Fabio Bagagiolo , Rossana Capuani , Luciano Marzufero

We study an independent best-response dynamics on network games in which the nodes (players) decide to revise their strategies independently with some probability. We provide several bounds on the convergence time to an equilibrium as a…

计算机科学与博弈论 · 计算机科学 2019-02-07 Paolo Penna , Laurent Viennot