中文
相关论文

相关论文: Optimism brings accurate perception in Iterated Pr…

200 篇论文

We study the interpersonal trust of a population of agents, asking whether chance may decide if a population ends up in a high trust or low trust state. We model this by a discrete time, random matching stochastic coordination game. Agents…

物理与社会 · 物理学 2024-05-20 Benedikt V. Meylahn , Arnoud V. den Boer , Michel Mandjes

If we could define the set of all bad outcomes, we could hard-code an agent which avoids them; however, in sufficiently complex environments, this is infeasible. We do not know of any general-purpose approaches in the literature to avoiding…

人工智能 · 计算机科学 2020-06-17 Michael K. Cohen , Marcus Hutter

We present tournament results and several powerful strategies for the Iterated Prisoner's Dilemma created using reinforcement learning techniques (evolutionary and particle swarm algorithms). These strategies are trained to perform well…

计算机科学与博弈论 · 计算机科学 2018-02-07 Marc Harper , Vincent Knight , Martin Jones , Georgios Koutsovoulos , Nikoleta E. Glynatsi , Owen Campbell

In human societies the probability of strategy adoption from a given person may be affected by the personal features. Now we investigate how an artificially imposed restricted ability to reproduce, overruling ones fitness, affects an…

种群与进化 · 定量生物学 2008-04-10 A. Szolnoki , M. Perc , G. Szabo

In this work, we ask for and answer what makes classical temporal-difference reinforcement learning with epsilon-greedy strategies cooperative. Cooperating in social dilemma situations is vital for animals, humans, and machines. While…

机器学习 · 计算机科学 2023-02-22 Wolfram Barfuss , Janusz Meylahn

The principle of optimism in the face of uncertainty is prevalent throughout sequential decision making problems such as multi-armed bandits and reinforcement learning (RL). To be successful, an optimistic RL algorithm must over-estimate…

机器学习 · 计算机科学 2021-12-07 Aldo Pacchiano , Philip J. Ball , Jack Parker-Holder , Krzysztof Choromanski , Stephen Roberts

Recognise that people have many, possibly conflicting, aspects to their personality. We hypothesise that each separate characteristic of a personality may be treated as an independent player in a non-zero sum many player game. This idea is…

综合数学 · 数学 2007-05-23 A. J. Roberts

Iterated games are a fundamental component of economic and evolutionary game theory. They describe situations where two players interact repeatedly and have the possibility to use conditional strategies that depend on the outcome of…

种群与进化 · 定量生物学 2015-06-12 Christian Hilbe , Martin A. Nowak , Karl Sigmund

As humans perceive and actively engage with the world, we adjust our decisions in response to shifting group dynamics and are influenced by social interactions. This study aims to identify which aspects of interaction affect…

物理与社会 · 物理学 2024-12-23 Lucila G. Alvarez-Zuzek , Laura Ferrarotti , Bruno Lepri , Riccardo Gallotti

Agents often have individual goals which depend on a group's actions. If agents trust a forecast of collective action and adapt strategically, such prediction can influence outcomes non-trivially, resulting in a form of performative…

机器学习 · 计算机科学 2025-02-18 António Góis , Mehrnaz Mofakhami , Fernando P. Santos , Gauthier Gidel , Simon Lacoste-Julien

We study a spatial two-strategy (cooperation and defection) Prisoner's Dilemma game with two types ($A$ and $B$) of players located on the sites of a square lattice. The evolution of strategy distribution is governed by iterated strategy…

物理与社会 · 物理学 2009-01-15 Gyorgy Szabo , Attila Szolnoki

We study a condition of favoring cooperation in a Prisoner's Dilemma game on complex networks. There are two kinds of players: cooperators and defectors. Cooperators pay a benefit b to their neighbors at a cost c, whereas defectors only…

物理与社会 · 物理学 2019-11-20 Tomohiko Konno

Multi-agent reinforcement learning has received significant interest in recent years notably due to the advancements made in deep reinforcement learning which have allowed for the developments of new architectures and learning algorithms.…

多智能体系统 · 计算机科学 2018-12-27 Nicolas Anastassacos , Mirco Musolesi

The problem of two companies of agents with one-step memory playing game is investigated in the context of the Iterated Prisoner's Dilemma under the partial imitation rule, where a player can imitate only those moves that he has observed in…

物理与社会 · 物理学 2011-04-01 Liangsheng Zhang , Wenjin Chen , Mathis Antony , K. Y. Szeto

A generalized model of games is proposed, in which cooperative games and non-cooperative games are special cases. Some games that are neither cooperative nor non-cooperative can be expressed and analyzed. The model is based on relationships…

计算机科学与博弈论 · 计算机科学 2016-10-10 Jiawei Li

As artificial agents become increasingly capable, what internal structure is *necessary* for an agent to act competently under uncertainty? Classical results show that optimal control can be *implemented* using belief states or world…

机器学习 · 计算机科学 2026-04-03 Aran Nayebi

Social dilemmas, where mutual cooperation can lead to high payoffs but participants face incentives to cheat, are ubiquitous in multi-agent interaction. We wish to construct agents that cooperate with pure cooperators, avoid exploitation by…

人工智能 · 计算机科学 2019-05-27 Alexander Peysakhovich , Adam Lerer

Risk sensitivity has become a central theme in reinforcement learning (RL), where convex risk measures and robust formulations provide principled ways to model preferences beyond expected return. Recent extensions to multi-agent RL (MARL)…

机器学习 · 计算机科学 2025-11-12 Runyu Zhang , Na Li , Asuman Ozdaglar , Jeff Shamma , Gioele Zardini

Our study contributes to the debate on the evolution of cooperation in the single-shot Prisoner's Dilemma (PD) played on networks. We construct a model in which individuals are connected with positive and negative ties. Some agents play…

物理与社会 · 物理学 2014-01-21 Simone Righi , Károly Takács

In a social dilemma, cooperation is collectively optimal, yet individually each group member prefers to defect. A class of successful strategies of direct reciprocity were recently found for the iterated prisoner's dilemma and for the…

种群与进化 · 定量生物学 2020-10-12 Yohsuke Murase , Seung Ki Baek