English
Related papers

Related papers: Entropy-Regularized Stochastic Games

200 papers

This paper considers a time-varying game with $N$ players. Every time slot, players observe their own random events and then take a control action. The events and control actions affect the individual utilities earned by each player. The…

Computer Science and Game Theory · Computer Science 2014-02-04 Michael J. Neely

This paper studies a discrete-time major-minor mean field game of stopping where the major player can choose either an optimal control or stopping time. We look for the relaxed equilibrium as a randomized stopping policy, which is…

Optimization and Control · Mathematics 2025-10-13 Xiang Yu , Jiacheng Zhang , Keyu Zhang , Zhou Zhou

Strategic-form min-max game theory examines the existence, multiplicity, selection of equilibria, and the worst-case computational complexity under perfect rationality. However, in many applications, games are drawn from an ensemble, and…

Computer Science and Game Theory · Computer Science 2026-02-17 Yuma Ichikawa

In this paper we consider two-person zero-sum risk-sensitive stochastic dynamic games with Borel state and action spaces and bounded reward. The term risk-sensitive refers to the fact that instead of the usual risk neutral optimization…

Optimization and Control · Mathematics 2021-07-21 Nicole Bäuerle , Ulrich Rieder

We consider two-player stochastic games played on a finite graph for infinitely many rounds. Stochastic games generalize both Markov decision processes (MDP) by adding an adversary player, and two-player deterministic games by adding…

Computer Science and Game Theory · Computer Science 2022-02-28 Laurent Doyen

We consider an N-player hierarchical game in which the i-th player's objective comprises of an expectation-valued term, parametrized by rival decisions, and a hierarchical term. Such a framework allows for capturing a broad range of…

Optimization and Control · Mathematics 2024-01-26 Shisheng Cui , Uday V. Shanbhag , Mathias Staudigl

Stochastic games generalize Markov decision processes (MDPs) to a multiagent setting by allowing the state transitions to depend jointly on all player actions, and having rewards determined by multiplayer matrix games at each state. We…

Computer Science and Game Theory · Computer Science 2013-01-18 Michael Kearns , Yishay Mansour , Satinder Singh

This report investigates the optimal design of event-triggered estimation for first-order linear stochastic systems. The problem is posed as a two-player team problem with a partially nested information pattern. The two players are given by…

Optimization and Control · Mathematics 2012-03-23 Adam Molin , Sandra Hirche

In this work, we present a novel characterization of approximate Nash equilibria in a class of convex games over the simplex. To achieve this, we regularize the utility functions using the Shannon entropy term, connect the solutions to the…

Optimization and Control · Mathematics 2025-07-18 Tatiana Tatarenko , S. Rasoul Etesami

We consider a subclass of $n$-player stochastic games, in which players have their own internal state/action spaces while they are coupled through their payoff functions. It is assumed that players' internal chains are driven by independent…

Machine Learning · Computer Science 2023-03-23 S. Rasoul Etesami

We introduce a simple stochastic dynamics for game theory. It assumes ``local'' rationality in the sense that any player climbs the gradient of his utility function in the presence of a stochastic force which represents deviation from…

Statistical Mechanics · Physics 2008-11-23 Matteo Marsili , Yi-Cheng Zhang

Entropy regularization has been extensively adopted to improve the efficiency, the stability, and the convergence of algorithms in reinforcement learning. This paper analyzes both quantitatively and qualitatively the impact of entropy…

Optimization and Control · Mathematics 2021-12-10 Xin Guo , Renyuan Xu , Thaleia Zariphopoulou

We use techniques from the statistical mechanics of disordered systems to analyse the properties of Nash equilibria of bimatrix games with large random payoff matrices. By means of an annealed bound, we calculate their number and analyse…

Disordered Systems and Neural Networks · Physics 2009-10-31 Johannes Berg , Martin Weigt

In many game-theoretic settings, agents are challenged with taking decisions against the uncertain behavior exhibited by others. Often, this uncertainty arises from multiple sources, e.g., incomplete information, limited computation,…

Computer Science and Game Theory · Computer Science 2025-07-22 Nicolas Lanzetti , Sylvain Fricker , Saverio Bolognani , Florian Dörfler , Dario Paccagnan

Probabilistic model checking for stochastic games enables formal verification of systems that comprise competing or collaborating entities operating in a stochastic environment. Despite good progress in the area, existing approaches focus…

Logic in Computer Science · Computer Science 2019-07-09 Marta Kwiatkowska , Gethin Norman , David Parker , Gabriel Santos

We study the game modification problem, where a benevolent game designer or a malevolent adversary modifies the reward function of a zero-sum Markov game so that a target deterministic or stochastic policy profile becomes the unique Markov…

Computer Science and Game Theory · Computer Science 2024-08-27 Young Wu , Jeremy McMahan , Yiding Chen , Yudong Chen , Xiaojin Zhu , Qiaomin Xie

We discuss similarities and differences between systems of interacting players maximizing their individual payoffs and particles minimizing their interaction energy. Long-run behavior of stochastic dynamics of spatial games with multiple…

Statistical Mechanics · Physics 2009-11-10 Jacek Miekisz

We consider 2-player stochastic games with perfectly observed actions, and study the limit, as the discount factor goes to one, of the equilibrium payoffs set. In the usual setup where current states are observed by the players, we show…

Optimization and Control · Mathematics 2014-12-11 Jérôme Renault , Bruno Ziliotto

The standard risk minimization paradigm of machine learning is brittle when operating in environments whose test distributions are different from the training distribution due to spurious correlations. Training on data from many…

Machine Learning · Computer Science 2020-03-20 Kartik Ahuja , Karthikeyan Shanmugam , Kush R. Varshney , Amit Dhurandhar

Computing the Nash equilibrium (NE) for N-player non-zerosum stochastic games is a formidable challenge. Currently, algorithmic methods in stochastic game theory are unable to compute NE for stochastic games (SGs) for settings in all but…

Optimization and Control · Mathematics 2021-03-25 David Mguni