中文
相关论文

相关论文: The State-Action-Reward-State-Action Algorithm in …

200 篇论文

We study co-evolutionary Prisoner's Dilemma games where each player can imitate both the strategy and imitation rule from a randomly chosen neighbor with a probability dependent on the payoff difference when the player's income is collected…

物理与社会 · 物理学 2009-11-04 Gyorgy Szabo , Attila Szolnoki , Jeromos Vukov

The significance of network structures in promoting group cooperation within social dilemmas has been widely recognized. Prior studies attribute this facilitation to the assortment of strategies driven by spatial interactions. Although…

多智能体系统 · 计算机科学 2024-08-20 Tianyu Ren , Xiao-Jun Zeng

Reinforcement learning (RL) is a powerful machine learning technique that has been successfully applied to a wide variety of problems. However, it can be unpredictable and produce suboptimal results in complicated learning environments.…

多智能体系统 · 计算机科学 2024-11-19 Brian Mintz , Feng Fu

Recent paradigm shifts from imitation learning to reinforcement learning (RL) is shown to be productive in understanding human behaviors. In the RL paradigm, individuals search for optimal strategies through interaction with the environment…

种群与进化 · 定量生物学 2024-12-20 Guozhong Zheng , Jiqiang Zhang , Shengfeng Deng , Weiran Cai , Li Chen

Exploration of mechanisms underlying the emergence of collective cooperation remains a focal point in field of evolution of cooperation. Prevailing studies often neglect historical information, relying on the latest rewards as the primary…

物理与社会 · 物理学 2024-02-07 Changyan Di , Jianyue Guan , Qingguo Zhou , Jingqiang Wang , Xiangyang Li

The paper studies the emergence and stability of cooperative behavior in populations of agents who interact among themselves in Prisoner's Dilemma games and who are allowed to choose their partners. The population is then subject to…

无序系统与神经网络 · 物理学 2007-05-23 Pawel Sobkowicz

In the evolutionary Prisoner's Dilemma (PD) game, agents play with each other and update their strategies in every generation according to some microscopic dynamical rule. In its spatial version, agents do not play with every other but,…

种群与进化 · 定量生物学 2009-03-08 Luis G. Moyano , Angel Sánchez

Cooperation is fundamental in Multi-Agent Systems (MAS) and Multi-Agent Reinforcement Learning (MARL), often requiring agents to balance individual gains with collective rewards. In this regard, this paper aims to investigate strategies to…

计算机科学与博弈论 · 计算机科学 2024-05-06 Vaigarai Sathi , Sabahat Shaik , Jaswanth Nidamanuri

Traditional evolutionary game theory describes how certain strategy spreads throughout the system where individual player imitates the most successful strategy among its neighborhood. Accordingly, player doesn't have own authority to change…

多智能体系统 · 计算机科学 2016-04-14 Sundong Kim , Jin-Jae Lee

Interactions among individuals in natural populations often occur in a dynamically changing environment. Understanding the role of environmental variation in population dynamics has long been a central topic in theoretical ecology and…

种群与进化 · 定量生物学 2021-05-18 Feng Huang , Ming Cao , Long Wang

Matrix games like Prisoner's Dilemma have guided research on social dilemmas for decades. However, they necessarily treat the choice to cooperate or defect as an atomic action. In real-world social dilemmas these choices are temporally…

多智能体系统 · 计算机科学 2017-02-13 Joel Z. Leibo , Vinicius Zambaldi , Marc Lanctot , Janusz Marecki , Thore Graepel

While many multi-armed bandit algorithms assume that rewards for all arms are constant across rounds, this assumption does not hold in many real-world scenarios. This paper considers the setting of recovering bandits (Pike-Burke &…

机器学习 · 计算机科学 2024-03-19 Yuto Tanimoto , Kenji Fukumizu

Game theory provides a quantitative framework for analyzing the behavior of rational agents. The Iterated Prisoner's Dilemma in particular has become a standard model for studying cooperation and cheating, with cooperation often emerging as…

种群与进化 · 定量生物学 2015-06-18 Alexander J. Stewart , Joshua B. Plotkin

Holding on to one's strategy is natural and common if the later warrants success and satisfaction. This goes against widespread simulation practices of evolutionary games, where players frequently consider changing their strategy even…

种群与进化 · 定量生物学 2012-05-04 Yongkui Liu , Xiaojie Chen , Lin Zhang , Long Wang , Matjaz Perc

As autonomous agents become more prevalent, understanding their collective behaviour in strategic interactions is crucial. This study investigates the emergent cooperative tendencies of systems of Large Language Model (LLM) agents in a…

多智能体系统 · 计算机科学 2025-01-28 Richard Willis , Yali Du , Joel Z Leibo , Michael Luck

Reciprocity is an important feature of human social interaction and underpins our cooperative nature. What is more, simple forms of reciprocity have proved remarkably resilient in matrix game social dilemmas. Most famously, the tit-for-tat…

多智能体系统 · 计算机科学 2019-03-20 Tom Eccles , Edward Hughes , János Kramár , Steven Wheelwright , Joel Z. Leibo

Social dilemmas have been widely studied to explain how humans are able to cooperate in society. Considerable effort has been invested in designing artificial agents for social dilemmas that incorporate explicit agent motivations that are…

多智能体系统 · 计算机科学 2021-08-30 Nicolas Anastassacos , Stephen Hailes , Mirco Musolesi

The universe involves many independent co-learning agents as an ever-evolving part of our observed environment. Yet, in practice, Multi-Agent Reinforcement Learning (MARL) applications are typically constrained to small, homogeneous…

机器学习 · 计算机科学 2025-04-29 Yann Bouteiller , Karthik Soma , Giovanni Beltrame

Game theory is fundamental to understanding cooperation between agents. Mainly, the Prisoner's Dilemma is a well-known model that has been extensively studied in complex networks. However, although the emergence of cooperation has been…

物理与社会 · 物理学 2023-01-04 Nastaran Lotfi , Francisco A. Rodrigues

We investigate an evolutionary prisoner's dilemma game among self-driven agents, where collective motion of biological flocks is imitated through averaging directions of neighbors. Depending on the temptation to defect and the velocity at…

物理与社会 · 物理学 2015-03-13 Zhuo Chen , Jian-Xi Gao , Yun-Ze Cai , Xiao-Ming Xu
‹ 上一页 1 2 3 10 下一页 ›