中文
相关论文

相关论文: Value Iteration Algorithm for Mean-field Games

200 篇论文

We consider both $N$-player and mean-field games of optimal portfolio liquidation in which the players are not allowed to change the direction of trading. Players with an initially short position of stocks are only allowed to buy while…

数理金融 · 定量金融 2025-07-31 Guanxing Fu , Paul P. Hager , Ulrich Horst

Value iteration is a fundamental algorithm for solving Markov Decision Processes (MDPs). It computes the maximal $n$-step payoff by iterating $n$ times a recurrence equation which is naturally associated to the MDP. At the same time, value…

形式语言与自动机理论 · 计算机科学 2019-04-30 Nikhil Balaji , Stefan Kiefer , Petr Novotný , Guillermo A. Pérez , Mahsa Shirmohammadi

Mean field game facilitates analyzing multi-armed bandit (MAB) for a large number of agents by approximating their interactions with an average effect. Existing mean field models for multi-agent MAB mostly assume a binary reward function,…

多智能体系统 · 计算机科学 2021-05-11 Xiong Wang , Riheng Jia

Equilibrium computation in markets usually considers settings where player valuation functions are known. We consider the setting where player valuations are unknown; using a PAC learning-theoretic framework, we analyze some classes of…

计算机科学与博弈论 · 计算机科学 2021-09-10 Vignesh Viswanathan , Omer Lev , Neel Patel , Yair Zick

We construct a semi-Lagrangian scheme for first-order, time-dependent, and non-local Mean Field Games. The convergence of the scheme to a weak solution of the system is analyzed by exploiting a key monotonicity property. To solve the…

数值分析 · 数学 2026-05-12 Elisabetta Carlini , Valentina Coscetti

This paper proves several Tauberian theorems for general iterations of operators, and provides two applications to zero-sum stochastic games where the total payoff is a weighted sum of the stage payoffs. The first application is to provide…

最优化与控制 · 数学 2016-09-09 Bruno Ziliotto

This paper is a continuation work of Ren et al. (2026) aiming to further devise q-learning algorithms for mean-field control (MFC) with controlled common noise. Based on the relaxed control formulation, we first establish the martingale…

最优化与控制 · 数学 2026-05-01 Zhenjie Ren , Xiaoli Wei , Xiang Yu , Xun Yu Zhou

Value iteration is a commonly used and empirically competitive method in solving many Markov decision process problems. However, it is known that value iteration has only pseudo-polynomial complexity in general. We establish a somewhat…

人工智能 · 计算机科学 2013-01-07 Omid Madani

We present a simulation-based approach for solution of mean field games (MFGs), using the framework of empirical game-theoretical analysis (EGTA). Our primary method employs a version of the double oracle, iteratively adding strategies…

多智能体系统 · 计算机科学 2023-02-14 Yongzhao Wang , Michael P. Wellman

The designs of many large-scale systems today, from traffic routing environments to smart grids, rely on game-theoretic equilibrium concepts. However, as the size of an $N$-player game typically grows exponentially with $N$, standard game…

This paper establishes that an MDP with a unique optimal policy and ergodic associated transition matrix ensures the convergence of various versions of the Value Iteration algorithm at a geometric rate that exceeds the discount factor…

机器学习 · 计算机科学 2024-06-17 Arsenii Mustafin , Alex Olshevsky , Ioannis Ch. Paschalidis

A decentralized blockchain is a distributed ledger that is often used as a platform for exchanging goods and services. This ledger is maintained by a network of nodes that obeys a set of rules, called a consensus protocol, which helps to…

最优化与控制 · 数学 2022-01-03 Lucy Klinger , Lei Zhang , Zhennan Zhou

Continuous games are multiplayer games in which strategy sets are compact and utility functions are continuous. These games typically have a highly complicated structure of Nash equilibria, and numerical methods for the equilibrium…

计算机科学与博弈论 · 计算机科学 2022-07-12 T. Kroupa , T. Votroubek

Two standard algorithms for approximately solving two-player zero-sum concurrent reachability games are value iteration and strategy iteration. We prove upper and lower bounds of 2^(m^(Theta(N))) on the worst case number of iterations…

计算机科学与博弈论 · 计算机科学 2012-03-02 Kristoffer Arnsfelt Hansen , Rasmus Ibsen-Jensen , Peter Bro Miltersen

In this paper we consider symmetric games where a large number of players can be in any one of d states. We derive a limiting mean field model and characterize its main properties. This mean field limit is a system of coupled ordinary…

最优化与控制 · 数学 2015-09-23 Diogo A. Gomes , Joana Mohr , Rafael R. Souza

Two-player quantitative zero-sum games provide a natural framework to synthesize controllers with performance guarantees for reactive systems within an uncontrollable environment. Classical settings include mean-payoff games, where the…

计算机科学中的逻辑 · 计算机科学 2015-09-25 Patricia Bouyer , Nicolas Markey , Mickael Randour , Kim G. Larsen , Simon Laursen

Mean-payoff games are important quantitative models for open reactive systems. They have been widely studied as games of full observation. In this paper we investigate the algorithmic properties of several sub-classes of mean-payoff games…

计算机科学与博弈论 · 计算机科学 2017-10-10 Paul Hunter , Arno Pauly , Guillermo A. Pérez , Jean-François Raskin

This paper studies the relation between equilibria in single-period, discrete-time and continuous-time mean field game models. First, for single-period mean field games, we establish the existence of equilibria and then prove the…

最优化与控制 · 数学 2024-11-04 Jodi Dianetti , Max Nendel , Ludovic Tangpi , Shichun Wang

This paper has two central aims: first, to provide simple conditions under which the generalized games in choice form and, consequently, the abstract economies, admit equilibrium; second, to study the solvability of several types of systems…

最优化与控制 · 数学 2016-05-17 Monica Patriche

We introduce a mean field model for optimal holding of a representative agent of her peers as a natural expected scaling limit from the corresponding $N-$agent model. The induced mean field dynamics appear naturally in a form which is not…

最优化与控制 · 数学 2022-04-05 Mao Fabrice Djete , Nizar Touzi