English
Related papers

Related papers: Distributed Asynchronous Policy Iteration for Sequ…

200 papers

In this paper we introduce an iterative Jacobi algorithm for solving distributed model predictive control (DMPC) problems, with linear coupled dynamics and convex coupled constraints. The algorithm guarantees stability and persistent…

Optimization and Control · Mathematics 2008-09-23 Dang Doan , Tamas Keviczky , Ion Necoara , Moritz Diehl

Dynamic zero-sum games are an important class of problems with applications ranging from evasion-pursuit and heads-up poker to certain adversarial versions of control problems such as multi-armed bandit and multiclass queuing problems.…

Computer Science and Game Theory · Computer Science 2015-06-12 Martin Haugh , Chun Wang

Distributed optimization and Nash equilibrium (NE) seeking problems have drawn much attention in the control community recently. This paper studies a class of non-cooperative games, known as N-cluster game, which subsumes both cooperative…

Optimization and Control · Mathematics 2023-03-01 Yipeng Pang , Guoqiang Hu

We develop the fictitious play algorithm in the context of the linear programming approach for mean field games of optimal stopping and mean field games with regular control and absorption. This algorithm allows to approximate the mean…

Optimization and Control · Mathematics 2023-01-25 Roxana Dumitrescu , Marcos Leutscher , Peter Tankov

We consider the problem of controlling a Markov decision process (MDP) with a large state space, so as to minimize average cost. Since it is intractable to compete with the optimal policy for large scale problems, we pursue the more modest…

Optimization and Control · Mathematics 2014-02-28 Yasin Abbasi-Yadkori , Peter L. Bartlett , Alan Malek

We study episodic two-player zero-sum Markov games (MGs) in the offline setting, where the goal is to find an approximate Nash equilibrium (NE) policy pair based on a dataset collected a priori. When the dataset does not have uniform…

Machine Learning · Computer Science 2023-01-02 Han Zhong , Wei Xiong , Jiyuan Tan , Liwei Wang , Tong Zhang , Zhaoran Wang , Zhuoran Yang

We describe an approximate dynamic programming approach to compute lower bounds on the optimal value function for a discrete time, continuous space, infinite horizon setting. The approach iteratively constructs a family of lower bounding…

Systems and Control · Electrical Eng. & Systems 2024-12-20 Paul N. Beuchat , Joseph Warrington , John Lygeros

Finding equilibria via gradient play in competitive multi-agent games has been attracting a growing amount of attention in recent years, with emphasis on designing efficient strategies where the agents operate in a decentralized and…

Computer Science and Game Theory · Computer Science 2022-11-17 Ruicheng Ao , Shicong Cen , Yuejie Chi

This paper develops a deep policy iteration method for high-dimensional finite-horizon mean-field games (MFG). We reformulate the game as a regenerative problem with deterministic cycles, which allows policy evaluation (PE), policy…

Numerical Analysis · Mathematics 2026-05-18 Shuixin Fang , Shupeng Wang , Zhen Wu , Hui Zhang , Tao Zhou

Box-simplex games are a family of bilinear minimax objectives which encapsulate graph-structured problems such as maximum flow [She17], optimal transport [JST19], and bipartite matching [AJJ+22]. We develop efficient near-linear time,…

Data Structures and Algorithms · Computer Science 2022-06-15 Arun Jambulapati , Yujia Jin , Aaron Sidford , Kevin Tian

High fidelity simulation of large-sized complex networks can be realized on a distributed computing platform that leverages the combined resources of multiple processors or machines. In a discrete event driven simulation, the assignment of…

Distributed, Parallel, and Cluster Computing · Computer Science 2012-10-16 Aditya Kurve , Christopher Griffin , David J. Miller , George Kesidis

This paper provides new conditions for dynamic optimality in discrete time and uses them to establish fundamental dynamic programming results for several commonly used recursive preference specifications. These include Epstein-Zin…

General Economics · Economics 2020-06-23 Guanlong Ren , John Stachurski

This paper is about a set-based computing method for solving a general class of two-player zero-sum Stackelberg differential games. We assume that the game is modeled by a set of coupled nonlinear differential equations, which can be…

Optimization and Control · Mathematics 2019-09-10 Xuhui Feng , Mario E. Villanueva , Boris Houska

We motivate and propose a new model for non-cooperative Markov game which considers the interactions of risk-aware players. This model characterizes the time-consistent dynamic "risk" from both stochastic state transitions (inherent to the…

Computer Science and Game Theory · Computer Science 2019-11-22 Wenjie Huang , Pham Viet Hai , William B. Haskell

We investigate a time-inconsistent, non-Markovian finite-player game in continuous time, where each player's objective functional depends non-linearly on the expected value of the state process. As a result, the classical Bellman optimality…

Probability · Mathematics 2025-12-10 Dylan Possamaï , Chiara Rossato

In this paper, we consider a large class of constrained non-cooperative stochastic Markov games with countable state spaces and discounted cost criteria. In one-player case, i.e., constrained discounted Markov decision models, it is…

Optimization and Control · Mathematics 2021-12-16 Anna Jaśkiewicz , Andrzej S. Nowak

Markov games provide a powerful framework for modeling strategic multi-agent interactions in dynamic environments. Traditionally, convergence properties of decentralized learning algorithms in these settings have been established only for…

Multiagent Systems · Computer Science 2025-06-13 Chinmay Maheshwari , Manxi Wu , Shankar Sastry

The Bellman operator constitutes the foundation of dynamic programming (DP). An alternative is presented by the Gauss-Seidel operator, whose evaluation, differently from that of the Bellman operator where the states are all processed at…

Optimization and Control · Mathematics 2021-10-07 Matilde Gargiani , Andrea Martinelli , Max Ruts Martinez , John Lygeros

Advanced Persistent Threats (APTs) have recently emerged as a significant security challenge for a cyber-physical system due to their stealthy, dynamic and adaptive nature. Proactive dynamic defenses provide a strategic and holistic…

Computer Science and Game Theory · Computer Science 2019-11-11 Linan Huang , Quanyan Zhu

We propose an implementable, neural network-based structure preserving probabilistic numerical approximation for a generalized obstacle problem describing the value of a zero-sum differential game of optimal stopping with asymmetric…

Numerical Analysis · Mathematics 2025-01-28 Ľubomír Baňas , Giorgio Ferrari , Tsiry Avisoa Randrianasolo
‹ Prev 1 8 9 10 Next ›