中文
相关论文

相关论文: Learning by Fictitious Play in Large Populations

200 篇论文

Federated Learning (FL) aims to foster collaboration among a population of clients to improve the accuracy of machine learning without directly sharing local data. Although there has been rich literature on designing federated learning…

机器学习 · 计算机科学 2023-02-20 Shengyuan Hu , Dung Daniel Ngo , Shuran Zheng , Virginia Smith , Zhiwei Steven Wu

The large-population system consists of considerable small agents whose individual behavior and mass effect are interrelated via their state-average. The mean-field game provides an efficient way to get the decentralized strategies of…

最优化与控制 · 数学 2014-03-25 Jianhui Huang , Shujun Wang

We study discrete-time, finite-state mean-field games (MFGs) under model uncertainty, where agents face ambiguity about the state transition probabilities. Each agent maximizes its expected payoff against the worst-case transitions within…

最优化与控制 · 数学 2026-01-21 Zongxia Liang , Zhou Zhou , Yaqi Zhuang , Bin Zou

Predicting the outcomes of cyber-physical systems with multiple human interactions is a challenging problem. This article reviews a game theoretical approach to address this issue, where reinforcement learning is employed to predict the…

多智能体系统 · 计算机科学 2019-10-14 Mert Albaba , Yildiray Yildiz

Evolutionary game theory has been an important tool for describing economic and social behaviour for decades. Approximate mean value equations describing the time evolution of strategy concentrations can be derived from the players'…

种群与进化 · 定量生物学 2011-02-10 Mathis Antony , Degang Wu , K Y Szeto

Altruistic cooperation is costly yet socially desirable. As a result, agents struggle to learn cooperative policies through independent reinforcement learning (RL). Indirect reciprocity, where agents consider their interaction partner's…

多智能体系统 · 计算机科学 2024-08-09 Martin Smit , Fernando P. Santos

The psychology of the individual is continuously changing in nature, which has a significant influence on the evolutionary dynamics of populations. To study the influence of the continuously changing psychology of individuals on the…

社会与信息网络 · 计算机科学 2025-02-11 Minyu Feng , Bin Pi , Liang-Jian Deng , Jürgen Kurths

An important challenge in non-cooperative game theory is coordinating on a single (approximate) equilibrium from many possibilities - a challenge that becomes even more complex when players hold private information. Recommender mechanisms…

计算机科学与博弈论 · 计算机科学 2025-05-30 Bengisu Guresti , Chongjie Zhang , Yevgeniy Vorobeychik

We consider a class of Mean Field Games in which the agents may interact through the statistical distribution of their states and controls. It is supposed that the Hamiltonian behaves like a power of its arguments as they tend to infinity,…

偏微分方程分析 · 数学 2020-06-24 Z Kobeissi

While AI systems have equaled or surpassed human performance in a wide variety of games such as Chess, Go, or Dota 2, describing these systems as truly "human-like" remains far-fetched. Despite their success, they fail to replicate the…

人工智能 · 计算机科学 2025-07-09 Aloïs Rautureau , Éric Piette

We consider stochastic differential games with a large number of players, with the aim of quantifying the gap between closed-loop, open-loop and distributed equilibria. We show that, under two different semi-monotonicity conditions, the…

概率论 · 数学 2025-05-06 Marco Cirant , Joe Jackson , Davide Francesco Redaelli

In a co-evolutionary context, the survive probability of individual elements of a system depends on their relation with their neighbors. The natural selection process depends on the whole population, which is determined by local events…

生物物理 · 物理学 2009-11-13 Juan G. Diaz Ochoa

We develop predictive models of pedestrian dynamics by encoding the coupled nature of multi-pedestrian interaction using game theory, and deep learning-based visual analysis to estimate person-specific behavior parameters. Building…

计算机视觉与模式识别 · 计算机科学 2017-03-29 Wei-Chiu Ma , De-An Huang , Namhoon Lee , Kris M. Kitani

This paper considers mean field games with optimal stopping time (OSMFGs) where agents make optimal exit decisions, the coupled obstacle and Fokker-Planck equations in such models pose challenges versus classic MFGs. This paper proposes a…

数值分析 · 数学 2023-10-10 Chengfeng Shen , Yifan Luo , Zhennan Zhou

We introduce a new virtual environment for simulating a card game known as "Big 2". This is a four-player game of imperfect information with a relatively complicated action space (being allowed to play 1,2,3,4 or 5 card combinations from an…

机器学习 · 计算机科学 2018-09-03 Henry Charlesworth

Demographic noise has profound effects on evolutionary and population dynamics, as well as on chemical reaction systems and models of epidemiology. Such noise is intrinsic and due to the discreteness of the dynamics in finite populations.…

物理与社会 · 物理学 2015-05-14 Tobias Galla

We study evolutionary game dynamics in a well-mixed populations of finite size, N. A well-mixed population means that any two individuals are equally likely to interact. In particular we consider the average abundances of two strategies, A…

种群与进化 · 定量生物学 2009-02-24 Tibor Antal , Martin A. Nowak , Arne Traulsen

We propose a multi-agent distributed reinforcement learning algorithm that balances between potentially conflicting short-term reward and sparse, delayed long-term reward, and learns with partial information in a dynamic environment. We…

机器学习 · 计算机科学 2022-04-06 Jing Tan , Ramin Khalili , Holger Karl

Model-based Reinforcement Learning approaches have the promise of being sample efficient. Much of the progress in learning dynamics models in RL has been made by learning models via supervised learning. But traditional model-based…

机器学习 · 计算机科学 2019-06-12 Shagun Sodhani , Anirudh Goyal , Tristan Deleu , Yoshua Bengio , Sergey Levine , Jian Tang

We consider an n-player symmetric stochastic game with weak interaction between the players. Time is continuous and the horizon and the number of states are finite. We show that the value function of each of the players can be approximated…

偏微分方程分析 · 数学 2018-07-13 Erhan Bayraktar , Asaf Cohen
‹ 上一页 1 8 9 10 下一页 ›