中文
相关论文

相关论文: Learning by Fictitious Play in Large Populations

200 篇论文

In repeated interactions between individuals, we do not expect that exactly the same situation will occur from one time to another. Contrary to what is common in models of repeated games in the literature, most real situations may differ a…

种群与进化 · 定量生物学 2007-05-23 Anders Eriksson , Kristian Lindgren

In Mean Field Games of Controls, the dynamics of the single agent is influenced not only by the distribution of the agents, as in the classical theory, but also by the distribution of their optimal strategies. In this paper, we study…

偏微分方程分析 · 数学 2023-02-01 Fabio Camilli , Claudio Marchi

We introduce a simple stochastic dynamics for game theory. It assumes ``local'' rationality in the sense that any player climbs the gradient of his utility function in the presence of a stochastic force which represents deviation from…

统计力学 · 物理学 2008-11-23 Matteo Marsili , Yi-Cheng Zhang

Certain but important classes of strategic-form games, including zero-sum and identical-interest games, have the fictitious-play-property (FPP), i.e., beliefs formed in fictitious play dynamics always converge to a Nash equilibrium (NE) in…

计算机科学与博弈论 · 计算机科学 2022-05-24 Muhammed O. Sayin , Kaiqing Zhang , Asuman Ozdaglar

A mathematical model for behavioral changes by pair interactions (i.e. due to direct contact) of individuals is developed. Three kinds of pair interactions can be distinguished: Imitative processes, avoidance processes, and compromising…

统计力学 · 物理学 2007-05-23 Dirk Helbing

We motivate and propose a new model for non-cooperative Markov game which considers the interactions of risk-aware players. This model characterizes the time-consistent dynamic "risk" from both stochastic state transitions (inherent to the…

计算机科学与博弈论 · 计算机科学 2019-11-22 Wenjie Huang , Pham Viet Hai , William B. Haskell

Trusting in others and reciprocating that trust with trustworthy actions are crucial to successful and prosperous societies. The Trust Game has been widely used to quantitatively study trust and trustworthiness, involving a sequential…

统计力学 · 物理学 2021-01-04 Ik Soo Lim

In a zero-sum stochastic game with signals, at each stage, two adversary players take decisions and receive a stage payoff determined by these decisions and a variable called state. The state follows a Markov chain, that is controlled by…

最优化与控制 · 数学 2021-12-02 Bruno Ziliotto

In stochastic games with incomplete information, the uncertainty is evoked by the lack of knowledge about a player's own and the other players' types, i.e. the utility function and the policy space, and also the inherent stochasticity of…

机器学习 · 计算机科学 2022-03-21 Hannes Eriksson , Debabrota Basu , Mina Alibeigi , Christos Dimitrakakis

In many stochastic games stemming from financial models, the environment evolves with latent factors and there may be common noise across agents' states. Two classic examples are: (i) multi-agent trading on electronic exchanges, and (ii)…

最优化与控制 · 数学 2019-07-24 Dena Firoozi , Peter E. Caines , Sebastian Jaimungal

We study two-player zero-sum stochastic games, and propose a form of independent learning dynamics called Doubly Smoothed Best-Response dynamics, which integrates a discrete and doubly smoothed variant of the best-response dynamics into…

计算机科学与博弈论 · 计算机科学 2023-03-07 Zaiwei Chen , Kaiqing Zhang , Eric Mazumdar , Asuman Ozdaglar , Adam Wierman

The neural activity in the visual processing is influenced by both external stimuli and internal brain states. Ideally, a neural predictive model should account for both of them. Currently, there are no dynamic encoding models that…

神经元与认知 · 定量生物学 2025-11-18 Finn Schmidt , Polina Turishcheva , Suhas Shrinivasan , Fabian H. Sinz

The notion that cooperation can aid a group of agents to solve problems more efficiently than if those agents worked in isolation is prevalent, despite the little quantitative groundwork to support it. Here we consider a primordial form of…

适应与自组织系统 · 物理学 2014-10-22 José F. Fontanari

Consider a 2-player normal-form game repeated over time. We introduce an adaptive learning procedure, where the players only observe their own realized payoff at each stage. We assume that agents do not know their own payoff function, and…

计算机科学与博弈论 · 计算机科学 2013-06-13 Mario Bravo , Mathieu Faure

Social learning is a powerful mechanism through which agents learn about the world from others. However, humans don't always choose to observe others, since social learning can carry time and cognitive resource costs. How do people balance…

多智能体系统 · 计算机科学 2025-07-15 Lance Ying , Ryan Truong , Joshua B. Tenenbaum , Samuel J. Gershman

We study two-player security games which can be viewed as sequences of nonzero-sum matrix games played by an Attacker and a Defender. The evolution of the game is based on a stochastic fictitious play process, where players do not have…

计算机科学与博弈论 · 计算机科学 2016-11-17 Kien C. Nguyen , Tansu Alpcan , Tamer Başar

In many settings of interest, a policy is set by one party, the leader, in order to influence the action of another party, the follower, where the follower's response is determined by some private information. A natural question to ask is,…

计算机科学与博弈论 · 计算机科学 2025-04-23 Michael Albert , Quinlan Dawkins , Minbiao Han , Haifeng Xu

In this paper, we study the problem of learning the skill distribution of a population of agents from observations of pairwise games in a tournament. These games are played among randomly drawn agents from the population. The agents in our…

机器学习 · 统计学 2020-06-16 Ali Jadbabaie , Anuran Makur , Devavrat Shah

In repeated games, such as auctions, players rely on autonomous learning agents to choose their actions. We study settings in which players have their agents make monetary transfers to other agents during play at their own expense, in order…

计算机科学与博弈论 · 计算机科学 2026-02-12 Yoav Kolumbus , Joe Halpern , Éva Tardos

Acquiring multiple skills has commonly involved collecting a large number of expert demonstrations per task or engineering custom reward functions. Recently it has been shown that it is possible to acquire a diverse set of skills by…

机器人学 · 计算机科学 2020-06-15 Rostam Dinyari , Pierre Sermanet , Corey Lynch