English
Related papers

Related papers: Visibility Optimization for Surveillance-Evasion G…

200 papers

In this paper, we formulate an evolutionarymultiple access control game with continuousvariable actions and coupled constraints. We characterize equilibria of the game and show that the pure equilibria are Pareto optimal and also resilient…

Computer Science and Game Theory · Computer Science 2015-03-19 Quanyan Zhu , Hamidou Tembine , Tamer Basar

The design of the performance index, also referred to as cost or reward shaping, is central to both optimal control and reinforcement learning, as it directly determines the behaviors, trade-offs, and objectives that the resulting control…

Systems and Control · Electrical Eng. & Systems 2025-10-14 Ayush Rai , Shaoshuai Mou , Brian D. O. Anderson

In this paper, we formulate a two-player zero-sum game under dynamic constraints defined by hybrid dynamical equations. The game consists of a min-max problem involving a cost functional that depends on the actions and resulting solutions…

Optimization and Control · Mathematics 2025-05-20 Santiago J. Leudo , Ricardo G. Sanfelice

Partially observable stochastic games provide a rich mathematical paradigm for modeling multi-agent dynamic decision making under uncertainty and partial information. However, they generally do not admit closed-form solutions and are…

Optimization and Control · Mathematics 2020-04-15 Yanling Chang , Chelsea C. White

The deduction game is a variation of the game of cops and robber on graphs in which searchers must capture an invisible evader in at most one move. Searchers know each others' initial locations, but can only communicate if they are on the…

Combinatorics · Mathematics 2024-12-24 Andrea Burgess , Danny Dyer , Mozhgan Farahani

We introduce the "inverse bandit" problem of estimating the rewards of a multi-armed bandit instance from observing the learning process of a low-regret demonstrator. Existing approaches to the related problem of inverse reinforcement…

Machine Learning · Statistics 2022-02-23 Wenshuo Guo , Kumar Krishna Agrawal , Aditya Grover , Vidya Muthukumar , Ashwin Pananjady

We study a differential game that governs the moderate-deviation heavy-traffic asymptotics of a multiclass single-server queueing control problem with a risk-sensitive cost. We consider a cost set on a finite but sufficiently large time…

Probability · Mathematics 2018-05-02 Rami Atar , Asaf Cohen

We consider the problem of learning Nash equilibrial policies for two-player risk-sensitive collision-avoiding interactions. Solving the Hamilton-Jacobi-Isaacs equations of such general-sum differential games in real time is an open…

Robotics · Computer Science 2025-03-21 Lei Zhang , Siddharth Das , Tanner Merry , Wenlong Zhang , Yi Ren

In dynamic noncooperative games, each player makes conjectures about other players' reactions before choosing a strategy. However, resulting equilibria may be multiple and do not always lead to desirable outcomes. These issues are typically…

Computer Science and Game Theory · Computer Science 2025-11-24 Francesco Morri , Hélène Le Cadre , David Salas , Didier Aussel

We frame the meta-learning of prediction procedures as a search for an optimal strategy in a two-player game. In this game, Nature selects a prior over distributions that generate labeled data consisting of features and an associated…

Machine Learning · Statistics 2020-09-29 Alex Luedtke , Incheoul Chung , Oleg Sofrygin

As the dimension of a system increases, traditional methods for control and differential games rapidly become intractable, making the design of safe autonomous agents challenging in complex or team settings. Deep-learning approaches avoid…

Optimization and Control · Mathematics 2025-04-29 William Sharpless , Zeyuan Feng , Somil Bansal , Sylvia Herbert

Under the assumption of no-arbitrage, the pricing of American and Bermudan options can be casted into optimal stopping problems. We propose a new adaptive simulation based algorithm for the numerical solution of optimal stopping problems in…

Probability · Mathematics 2009-09-29 Daniel Egloff , Michael Kohler , Nebojsa Todorovic

Given a two-dimensional polygonal space, the multi-robot visibility-based pursuit-evasion problem tasks several pursuer robots with the goal of establishing visibility with an arbitrarily fast evader. The best known complete algorithm for…

Robotics · Computer Science 2021-04-12 Trevor Olsen , Anne M. Tumlin , Nicholas M. Stiffler , Jason M. O'Kane

Modifying the reward-biased maximum likelihood method originally proposed in the adaptive control literature, we propose novel learning algorithms to handle the explore-exploit trade-off in linear bandits problems as well as generalized…

Machine Learning · Computer Science 2020-10-09 Yu-Heng Hung , Ping-Chun Hsieh , Xi Liu , P. R. Kumar

We present solutions to a continuous patrolling game played on network. In this zero-sum game, an Attacker chooses a time and place to attack a network for a fixed amount of time. A Patroller patrols the network with the aim of intercepting…

Computer Science and Game Theory · Computer Science 2023-01-31 Thuy Bui , Thomas Lidbetter

Stochastic patrol routing is known to be advantageous in adversarial settings; however, the optimal choice of stochastic routing strategy is dependent on a model of the adversary. We adopt a worst-case omniscient adversary model from the…

Systems and Control · Electrical Eng. & Systems 2025-04-10 Yohan John , Gilberto Diaz-Garcia , Xiaoming Duan , Jason R. Marden , Francesco Bullo

A two-player finite horizon linear-quadratic Stackelberg differential game is considered. The feature of this game is that the control cost of a follower in the cost functionals of both players is small, which means that the game under…

Optimization and Control · Mathematics 2025-12-11 Valery Y. Glizer , Vladimir Turetsky

We study a two-player zero-sum stochastic differential game with both players adopting impulse controls, on a finite time horizon. The Hamilton-Jacobi-Bellman-Isaacs (HJBI) partial differential equation of the game turns out to be a…

Probability · Mathematics 2012-06-26 Andrea Cosso

An important feature of a dynamic game is its monitoring structure namely, what the players effectively see from the played actions. We consider games with arbitrary monitoring structures. One of the purposes of this paper is to know to…

Information Theory · Computer Science 2012-10-24 Maël Le Treust , Samson Lasaulce

This paper considers a game-theoretic formulation of the covert communications problem with finite blocklength, where the transmitter (Alice) can randomly vary her transmit power in different blocks, while the warden (Willie) can randomly…

Information Theory · Computer Science 2020-05-28 Alex S. Leong , Daniel E. Quevedo , Subhrakanti Dey