English
Related papers

Related papers: Lead distance under a pickoff limit in Major Leagu…

200 papers

We introduce a new non-zero-sum game of optimal stopping with asymmetric exercise opportunities. Given a stochastic process modelling the value of an asset, one player observes and can act on the process continuously, while the other player…

Probability · Mathematics 2024-05-16 José Luis Pérez , Neofytos Rodosthenous , Kazutoshi Yamazaki

We study a model of two-player, zero-sum, stopping games with asymmetric information. We assume that the payoff depends on two continuous-time Markov chains (X, Y), where X is only observed by player 1 and Y only by player 2, implying that…

Optimization and Control · Mathematics 2017-12-06 Fabien Gensbittel , Christine Grün

Although parallelism has been extensively used in reinforcement learning (RL), the quantitative effects of parallel exploration are not well understood theoretically. We study the benefits of simple parallel exploration for reward-free RL…

Machine Learning · Computer Science 2023-03-03 Pedro Cisneros-Velarde , Boxiang Lyu , Sanmi Koyejo , Mladen Kolar

In this paper, a leader-follower stochastic differential game is studied for a linear stochastic differential equation with a quadratic cost functional. The coefficients in the state equation and the weighting matrices in the cost…

Optimization and Control · Mathematics 2021-07-13 Zixuan Li , Jingtao Shi

This paper is concerned with a three-level multi-leader-follower incentive Stackelberg game with $H_\infty$ constraint. Based on $H_2/H_\infty$ control theory, we firstly obtain the worst-case disturbance and the team-optimal strategy by…

Optimization and Control · Mathematics 2024-12-13 Na Xiang , Jingtao Shi

We consider a distributed stochastic approximation (SA) scheme for computing an equilibrium of a stochastic Nash game. Standard SA schemes employ diminishing steplength sequences that are square summable but not summable. Such requirements…

Optimization and Control · Mathematics 2013-03-20 Farzad Yousefian , Angelia Nedich , Uday V. Shanbhag

We study two-player security games which can be viewed as sequences of nonzero-sum matrix games played by an Attacker and a Defender. The evolution of the game is based on a stochastic fictitious play process. Players do not have access to…

Computer Science and Game Theory · Computer Science 2010-03-16 Kien C. Nguyen , Tansu Alpcan , Tamer Basar

Using methods from the statistical mechanics of disordered systems we analyze the properties of bimatrix games with random payoffs in the limit where the number of pure strategies of each player tends to infinity. We analytically calculate…

Disordered Systems and Neural Networks · Physics 2009-10-31 Johannes Berg

Recent results in the ML community have revealed that learning algorithms used to compute the optimal strategy for the leader to commit to in a Stackelberg game, are susceptible to manipulation by the follower. Such a learning algorithm…

Computer Science and Game Theory · Computer Science 2022-09-12 Georgios Birmpas , Jiarui Gan , Alexandros Hollender , Francisco J. Marmolejo-Cossío , Ninad Rajgopal , Alexandros A. Voudouris

In society, mutual cooperation, defection, and asymmetric exploitative relationships are common. Whereas cooperation and defection are studied extensively in the literature on game theory, asymmetric exploitative relationships between…

Optimization and Control · Mathematics 2019-11-13 Yuma Fujimoto , Kunihiko Kaneko

This paper studies open-loop and feedback solutions to leader-follower mean field linear-quadratic-Gaussian games with multiplicative noise by the direct approach. The leader-follower game involves a leader and many followers, where the…

Optimization and Control · Mathematics 2025-12-04 Bing-Chang Wang , Huanshui Zhang , Ji-Feng Zhang

Two-player mean-payoff Stackelberg games are nonzero-sum infinite duration games played on a bi-weighted graph by Leader (Player 0) and Follower (Player 1). Such games are played sequentially: first, Leader announces her strategy, second,…

Optimization and Control · Mathematics 2021-08-04 Mrudula Balachander , Shibashis Guha , Jean-François Raskin

Markov decision processes (MDP) are finite-state systems with both strategic and probabilistic choices. After fixing a strategy, an MDP produces a sequence of probability distributions over states. The sequence is eventually synchronizing…

Computer Science and Game Theory · Computer Science 2013-11-01 Laurent Doyen , Thierry Massart , Mahsa Shirmohammadi

We study a class of stochastic dynamic games that exhibit strategic complementarities between players; formally, in the games we consider, the payoff of a player has increasing differences between her own state and the empirical…

Computer Science and Game Theory · Computer Science 2010-12-13 Sachin Adlakha , Ramesh Johari

This paper studies two-player zero-sum repeated Bayesian games in which every player has a private type that is unknown to the other player, and the initial probability of the type of every player is publicly known. The types of players are…

Computer Science and Game Theory · Computer Science 2017-11-08 Lichun Li , Cedric Langbort , Jeff Shamma

Given a finite set $K$, we denote by $X=\Delta(K)$ the set of probabilities on $K$ and by $Z=\Delta_f(X)$ the set of Borel probabilities on $X$ with finite support. Studying a Markov Decision Process with partial information on $K$…

Optimization and Control · Mathematics 2012-02-29 Jérôme Renault , Xavier Venel

We study countably infinite stochastic 2-player games with reachability objectives. Our results provide a complete picture of the memory requirements of $\varepsilon$-optimal (resp. optimal) strategies. These results depend on the size of…

Computer Science and Game Theory · Computer Science 2024-07-03 Stefan Kiefer , Richard Mayr , Mahsa Shirmohammadi , Patrick Totzke

We study nondeterministic strategies in parity games with the aim of computing a most permissive winning strategy. Following earlier work, we measure permissiveness in terms of the average number/weight of transitions blocked by the…

Logic in Computer Science · Computer Science 2013-01-14 Patricia Bouyer , Nicolas Markey , Jörg Olschewski , Michael Ummels

In the game of cricket, the result of coin toss is assumed to be one of the determinants of match outcome. The decision to bat first after winning the toss is often taken to make the best use of superior pitch conditions and set a big…

Applications · Statistics 2020-07-14 Manar D. Samad , Sumen Sen

In finite games mixed Nash equilibria always exist, but pure equilibria may fail to exist. To assess the relevance of this nonexistence, we consider games where the payoffs are drawn at random. In particular, we focus on games where a large…

Computer Science and Game Theory · Computer Science 2020-06-18 Ben Amiet , Andrea Collevecchio , Marco Scarsini , Ziwen Zhong