中文
相关论文

相关论文: Visibility Optimization for Surveillance-Evasion G…

200 篇论文

In this paper, we formulate an evolutionarymultiple access control game with continuousvariable actions and coupled constraints. We characterize equilibria of the game and show that the pure equilibria are Pareto optimal and also resilient…

计算机科学与博弈论 · 计算机科学 2015-03-19 Quanyan Zhu , Hamidou Tembine , Tamer Basar

The design of the performance index, also referred to as cost or reward shaping, is central to both optimal control and reinforcement learning, as it directly determines the behaviors, trade-offs, and objectives that the resulting control…

系统与控制 · 电气工程与系统科学 2025-10-14 Ayush Rai , Shaoshuai Mou , Brian D. O. Anderson

In this paper, we formulate a two-player zero-sum game under dynamic constraints defined by hybrid dynamical equations. The game consists of a min-max problem involving a cost functional that depends on the actions and resulting solutions…

最优化与控制 · 数学 2025-05-20 Santiago J. Leudo , Ricardo G. Sanfelice

Partially observable stochastic games provide a rich mathematical paradigm for modeling multi-agent dynamic decision making under uncertainty and partial information. However, they generally do not admit closed-form solutions and are…

最优化与控制 · 数学 2020-04-15 Yanling Chang , Chelsea C. White

The deduction game is a variation of the game of cops and robber on graphs in which searchers must capture an invisible evader in at most one move. Searchers know each others' initial locations, but can only communicate if they are on the…

组合数学 · 数学 2024-12-24 Andrea Burgess , Danny Dyer , Mozhgan Farahani

We introduce the "inverse bandit" problem of estimating the rewards of a multi-armed bandit instance from observing the learning process of a low-regret demonstrator. Existing approaches to the related problem of inverse reinforcement…

We study a differential game that governs the moderate-deviation heavy-traffic asymptotics of a multiclass single-server queueing control problem with a risk-sensitive cost. We consider a cost set on a finite but sufficiently large time…

概率论 · 数学 2018-05-02 Rami Atar , Asaf Cohen

We consider the problem of learning Nash equilibrial policies for two-player risk-sensitive collision-avoiding interactions. Solving the Hamilton-Jacobi-Isaacs equations of such general-sum differential games in real time is an open…

机器人学 · 计算机科学 2025-03-21 Lei Zhang , Siddharth Das , Tanner Merry , Wenlong Zhang , Yi Ren

In dynamic noncooperative games, each player makes conjectures about other players' reactions before choosing a strategy. However, resulting equilibria may be multiple and do not always lead to desirable outcomes. These issues are typically…

计算机科学与博弈论 · 计算机科学 2025-11-24 Francesco Morri , Hélène Le Cadre , David Salas , Didier Aussel

We frame the meta-learning of prediction procedures as a search for an optimal strategy in a two-player game. In this game, Nature selects a prior over distributions that generate labeled data consisting of features and an associated…

机器学习 · 统计学 2020-09-29 Alex Luedtke , Incheoul Chung , Oleg Sofrygin

As the dimension of a system increases, traditional methods for control and differential games rapidly become intractable, making the design of safe autonomous agents challenging in complex or team settings. Deep-learning approaches avoid…

最优化与控制 · 数学 2025-04-29 William Sharpless , Zeyuan Feng , Somil Bansal , Sylvia Herbert

Under the assumption of no-arbitrage, the pricing of American and Bermudan options can be casted into optimal stopping problems. We propose a new adaptive simulation based algorithm for the numerical solution of optimal stopping problems in…

概率论 · 数学 2009-09-29 Daniel Egloff , Michael Kohler , Nebojsa Todorovic

Given a two-dimensional polygonal space, the multi-robot visibility-based pursuit-evasion problem tasks several pursuer robots with the goal of establishing visibility with an arbitrarily fast evader. The best known complete algorithm for…

机器人学 · 计算机科学 2021-04-12 Trevor Olsen , Anne M. Tumlin , Nicholas M. Stiffler , Jason M. O'Kane

Modifying the reward-biased maximum likelihood method originally proposed in the adaptive control literature, we propose novel learning algorithms to handle the explore-exploit trade-off in linear bandits problems as well as generalized…

机器学习 · 计算机科学 2020-10-09 Yu-Heng Hung , Ping-Chun Hsieh , Xi Liu , P. R. Kumar

We present solutions to a continuous patrolling game played on network. In this zero-sum game, an Attacker chooses a time and place to attack a network for a fixed amount of time. A Patroller patrols the network with the aim of intercepting…

计算机科学与博弈论 · 计算机科学 2023-01-31 Thuy Bui , Thomas Lidbetter

Stochastic patrol routing is known to be advantageous in adversarial settings; however, the optimal choice of stochastic routing strategy is dependent on a model of the adversary. We adopt a worst-case omniscient adversary model from the…

系统与控制 · 电气工程与系统科学 2025-04-10 Yohan John , Gilberto Diaz-Garcia , Xiaoming Duan , Jason R. Marden , Francesco Bullo

A two-player finite horizon linear-quadratic Stackelberg differential game is considered. The feature of this game is that the control cost of a follower in the cost functionals of both players is small, which means that the game under…

最优化与控制 · 数学 2025-12-11 Valery Y. Glizer , Vladimir Turetsky

We study a two-player zero-sum stochastic differential game with both players adopting impulse controls, on a finite time horizon. The Hamilton-Jacobi-Bellman-Isaacs (HJBI) partial differential equation of the game turns out to be a…

概率论 · 数学 2012-06-26 Andrea Cosso

An important feature of a dynamic game is its monitoring structure namely, what the players effectively see from the played actions. We consider games with arbitrary monitoring structures. One of the purposes of this paper is to know to…

信息论 · 计算机科学 2012-10-24 Maël Le Treust , Samson Lasaulce

This paper considers a game-theoretic formulation of the covert communications problem with finite blocklength, where the transmitter (Alice) can randomly vary her transmit power in different blocks, while the warden (Willie) can randomly…

信息论 · 计算机科学 2020-05-28 Alex S. Leong , Daniel E. Quevedo , Subhrakanti Dey