中文
相关论文

相关论文: Filtered Fictitious Play for Perturbed Observation…

200 篇论文

Decentralised optimisation tasks are important components of multi-agent systems. These tasks can be interpreted as n-player potential games: therefore game-theoretic learning algorithms can be used to solve decentralised optimisation…

多智能体系统 · 计算机科学 2013-01-16 Michalis Smyrnakis

In this paper, we deepen the analysis of continuous time Fictitious Play learning algorithm to the consideration of various finite state Mean Field Game settings (finite horizon, $\gamma$-discounted), allowing in particular for the…

最优化与控制 · 数学 2020-10-27 Sarah Perrin , Julien Perolat , Mathieu Laurière , Matthieu Geist , Romuald Elie , Olivier Pietquin

It is now well known that decentralised optimisation can be formulated as a potential game, and game-theoretical learning algorithms can be used to find an optimum. One of the most common learning techniques in game theory is fictitious…

机器学习 · 统计学 2011-12-13 Michalis Smyrnakis , David S. Leslie

Fictitious play is an algorithm for computing Nash equilibria of matrix games. Recently, machine learning variants of fictitious play have been successfully applied to complicated real-world games. This paper presents a simple modification…

计算机科学与博弈论 · 计算机科学 2022-12-21 Alex Cloud , Albert Wang , Wesley Kerr

We propose a deep neural network-based algorithm to identify the Markovian Nash equilibrium of general large $N$-player stochastic differential games. Following the idea of fictitious play, we recast the $N$-player game into $N$ decoupled…

最优化与控制 · 数学 2020-06-08 Jiequn Han , Ruimeng Hu

Recent extensions to dynamic games of the well-known fictitious play learning procedure in static games were proved to globally converge to stationary Nash equilibria in two important classes of dynamic games (zero-sum and…

计算机科学与博弈论 · 计算机科学 2022-07-08 Lucas Baudin , Rida Laraki

Robots deployed to the real world must be able to interact with other agents in their environment. Dynamic game theory provides a powerful mathematical framework for modeling scenarios in which agents have individual objectives and…

Stochastic differential games have been used extensively to model agents' competitions in Finance, for instance, in P2P lending platforms from the Fintech industry, the banking system for systemic risk, and insurance markets. The recently…

最优化与控制 · 数学 2021-03-23 Jiequn Han , Ruimeng Hu , Jihao Long

In this paper, we apply the idea of fictitious play to design deep neural networks (DNNs), and develop deep learning theory and algorithms for computing the Nash equilibrium of asymmetric $N$-player non-zero-sum stochastic differential…

最优化与控制 · 数学 2020-09-07 Ruimeng Hu

Considering the interaction through mutual interference of the different radio devices, the channel selection (CS) problem in decentralized parallel multiple access channels can be modeled by strategic-form games. Here, we show that the CS…

计算机科学与博弈论 · 计算机科学 2010-09-28 S. M. Perlaza , H. Tembine , S. Lasaulce , V. Quintero-Florez

Learning problems commonly exhibit an interesting feedback mechanism wherein the population data reacts to competing decision makers' actions. This paper formulates a new game theoretic framework for this phenomenon, called "multi-player…

计算机科学与博弈论 · 计算机科学 2022-04-08 Adhyyan Narang , Evan Faulkner , Dmitriy Drusvyatskiy , Maryam Fazel , Lillian J. Ratliff

Many real-world decision problems involve the interaction of multiple self-interested agents with limited sensing ability. The partially observable stochastic game (POSG) provides a mathematical framework for modeling these problems,…

计算机科学与博弈论 · 计算机科学 2024-10-30 Tyler Becker , Zachary Sunberg

Recent techniques based on Mean Field Games (MFGs) allow the scalable analysis of multi-player games with many similar, rational agents. However, standard MFGs remain limited to homogeneous players that weakly influence each other, and…

计算机科学与博弈论 · 计算机科学 2023-12-19 Kai Cui , Gökçe Dayanıklı , Mathieu Laurière , Matthieu Geist , Olivier Pietquin , Heinz Koeppl

Robots and autonomous systems must interact with one another and their environment to provide high-quality services to their users. Dynamic game theory provides an expressive theoretical framework for modeling scenarios involving multiple…

机器人学 · 计算机科学 2021-08-10 Lasse Peters , David Fridovich-Keil , Vicenç Rubies-Royo , Claire J. Tomlin , Cyrill Stachniss

While fictitious play is guaranteed to converge to Nash equilibrium in certain game classes, such as two-player zero-sum games, it is not guaranteed to converge in non-zero-sum and multiplayer games. We show that fictitious play in fact…

计算机科学与博弈论 · 计算机科学 2024-07-30 Sam Ganzfried

This paper considers mean field games with optimal stopping time (OSMFGs) where agents make optimal exit decisions, the coupled obstacle and Fokker-Planck equations in such models pose challenges versus classic MFGs. This paper proposes a…

数值分析 · 数学 2023-10-10 Chengfeng Shen , Yifan Luo , Zhennan Zhou

We study Nash equilibrium learning in partially observable Markov games (POMGs), a multi-agent reinforcement learning framework in which agents cannot fully observe the underlying state. Prior work in this setting relies on centralization…

计算机科学与博弈论 · 计算机科学 2026-05-08 Philip Jordan , Maryam Kamgarpour

We propose a reinforcement learning algorithm for stationary mean-field games, where the goal is to learn a pair of mean-field state and stationary policy that constitutes the Nash equilibrium. When viewing the mean-field state and the…

机器学习 · 计算机科学 2020-10-12 Qiaomin Xie , Zhuoran Yang , Zhaoran Wang , Andreea Minca

We develop the fictitious play algorithm in the context of the linear programming approach for mean field games of optimal stopping and mean field games with regular control and absorption. This algorithm allows to approximate the mean…

最优化与控制 · 数学 2023-01-25 Roxana Dumitrescu , Marcos Leutscher , Peter Tankov

We investigate convergence of decentralized fictitious play (DFP) in near-potential games, wherein agents preferences can almost be captured by a potential function. In DFP agents keep local estimates of other agents' empirical frequencies,…

计算机科学与博弈论 · 计算机科学 2022-01-31 Sarper Aydin , Sina Arefizadeh , Ceyhun Eksin
‹ 上一页 1 2 3 10 下一页 ›