English
Related papers

Related papers: Finite-horizon Approximations and Episodic Equilib…

200 papers

Policy optimization is among the most popular and successful reinforcement learning algorithms, and there is increasing interest in understanding its theoretical guarantees. In this work, we initiate the study of policy optimization for the…

Machine Learning · Computer Science 2022-02-08 Liyu Chen , Haipeng Luo , Aviv Rosenberg

In this paper, we introduce a novel equilibrium concept, called the equilibrium cycle, which seeks to capture the outcome of oscillatory game dynamics. Unlike the (pure) Nash equilibrium, which defines a fixed point of mutual best…

Theoretical Economics · Economics 2025-10-07 Tushar Shankar Walunj , Shiksha Singhal , Veeraruna Kavitha , Jayakrishnan Nair

We formulate and study a class of two-player zero-sum stochastic dynamic games with partial and asymmetric information. Information asymmetry introduces fundamental challenges involving \emph{belief representation} and \emph{theory of mind}…

Optimization and Control · Mathematics 2026-03-20 Yuxiang Guan , Iman Shames , Tyler Summers

We consider learning approximate Nash equilibria for discrete-time mean-field games with nonlinear stochastic state dynamics subject to both average and discounted costs. To this end, we introduce a mean-field equilibrium (MFE) operator,…

Systems and Control · Electrical Eng. & Systems 2022-11-11 Berkay Anahtarcı , Can Deha Karıksız , Naci Saldi

In this paper, we investigate infinite horizon jump-diffusion forward-backward stochastic differential equations under some monotonicity conditions. We establish an existence and uniqueness theorem, two stability results and a comparison…

Probability · Mathematics 2016-08-22 Zhiyong Yu

We propose a new mean-field game model with two states to study synchronization phenomena, and we provide a comprehensive characterization of stationary and dynamic equilibria along with their stability properties. The game undergoes a…

Optimization and Control · Mathematics 2024-08-21 Felix Höfer , H. Mete Soner

Recently, a deep-learning algorithm referred to as Deep Galerkin Method (DGM), has gained a lot of attention among those trying to solve numerically Mean Field Games with finite horizon, even if the performance seems to be decreasing…

Optimization and Control · Mathematics 2024-03-01 René Carmona , Claire Zeng

Diffusion approximation provides weak approximation for stochastic gradient descent algorithms in a finite time horizon. In this paper, we introduce new tools motivated by the backward error analysis of numerical stochastic differential…

Machine Learning · Computer Science 2019-09-05 Yuanyuan Feng , Tingran Gao , Lei Li , Jian-Guo Liu , Yulong Lu

We consider a class of continuous-time dynamic games involving a large number of players. Each player selects actions from a finite set and evolves through a finite set of states. State transitions occur stochastically and depend on the…

Systems and Control · Electrical Eng. & Systems 2025-11-12 Leonardo Pedroso , Andrea Agazzi , W. P. M. H. Heemels , Mauro Salazar

In this paper, we develop a provably correct optimal control strategy for a finite deterministic transition system. By assuming that penalties with known probabilities of occurrence and dynamics can be sensed locally at the states of the…

Robotics · Computer Science 2013-03-15 Mária Svoreňová , Ivana Černá , Calin Belta

Infinitely repeated games can support cooperative outcomes that are not equilibria in the one-shot game. The idea is to make sure that any gains from deviating will be offset by retaliation in future rounds. However, this model of…

Computer Science and Game Theory · Computer Science 2024-06-04 Ratip Emin Berker , Vincent Conitzer

Power system operators and electric utility companies often impose a coincident peak demand charge on customers when the aggregate system demand reaches its maximum. This charge incentivizes customers to strategically shift their peak usage…

Systems and Control · Electrical Eng. & Systems 2025-05-16 Liudong Chen , Jay Sethuraman , Bolun Xu

Game theory has emerged as a powerful framework for modeling a large range of multi-agent scenarios. Many algorithmic solutions require discrete, finite games with payoffs that have a closed-form specification. In contrast, many real-world…

Computer Science and Game Theory · Computer Science 2018-06-13 Abdullah Al-Dujaili , Erik Hemberg , Una-May O'Reilly

Multi-agent reinforcement learning, despite its popularity and empirical success, faces significant scalability challenges in large-population dynamic games. Graphon mean field games (GMFGs) offer a principled framework for approximating…

Optimization and Control · Mathematics 2025-06-09 Philipp Plank , Yufei Zhang

This paper investigates a two-person non-homogeneous linear-quadratic stochastic differential game (LQ-SDG, for short) in an infinite horizon for a system regulated by a time-invariant Markov chain. Both non-zero-sum and zero-sum LQ-SDG…

Optimization and Control · Mathematics 2024-08-26 Fan Wu , Xun Li , Jie Xiong , Xin Zhang

The infinite horizon setting is widely adopted for problems of reinforcement learning (RL). These invariably result in stationary policies that are optimal. In many situations, finite horizon control problems are of interest and for such…

Machine Learning · Computer Science 2025-03-21 Soumyajit Guin , Shalabh Bhatnagar

We study a finite-horizon two-person zero-sum risk-sensitive stochastic game for continuous-time Markov chains and Borel state and action spaces, in which payoff rates, transition rates and terminal reward functions are allowed to be…

Optimization and Control · Mathematics 2021-03-09 Junyu Zhang , Xianping Guo , Li Xia

We consider a finite-horizon, zero-sum game in which both players control a stochastic differential equation by invoking impulses. We derive a control randomization formulation of the game and use the existence of a value for the randomized…

Optimization and Control · Mathematics 2025-05-13 Magnus Perninge

This paper proves that the episodic learning environment of every finite-horizon decision task has a unique steady state under any behavior policy, and that the marginal distribution of the agent's input indeed converges to the steady-state…

Machine Learning · Computer Science 2021-01-14 Huang Bojun

We introduce a new hypothesis testing-based learning dynamics in which players update their strategies by combining hypothesis testing with utility-driven exploration. In this dynamics, each player forms beliefs about opponents' strategies…

Computer Science and Game Theory · Computer Science 2025-08-01 Ruifan Yang , Manxi Wu
‹ Prev 1 4 5 6 7 8 10 Next ›