中文
相关论文

相关论文: The State-Action-Reward-State-Action Algorithm in …

200 篇论文

We study the evolution of behavior under reinforcement learning in a Prisoner's Dilemma where agents interact in a regular network and can learn about whether they play one-shot or repeatedly by incurring a cost of deliberation. With…

物理与社会 · 物理学 2024-03-28 Rossana Mastrandrea , Leonardo Boncinelli , Ennio Bilancini

We study the evolution of cooperation in the spatial prisoner's dilemma game where players are allowed to establish new interactions with others. By employing a simple coevolutionary rule entailing only two crucial parameters, we find that…

物理与社会 · 物理学 2011-12-30 Chunyan Zhang , Jianlei Zhang , Guangming Xie , Long Wang , Matjaz Perc

We study environments in which agents are randomly matched to play a Prisoner's Dilemma, and each player observes a few of the partner's past actions against previous opponents. We depart from the existing related literature by allowing a…

理论经济学 · 经济学 2020-06-30 Yuval Heller , Erik Mohlin

Situations of conflict giving rise to social dilemmas are widespread in society and game theory is one major way in which they can be investigated. Starting from the observation that individuals in society interact through networks of…

物理与社会 · 物理学 2010-11-24 Enea Pestelacci , Marco Tomassini , Leslie Luthi

Humans and other animals can adapt their social behavior in response to environmental cues including the feedback obtained through experience. Nevertheless, the effects of the experience-based learning of players in evolution and…

种群与进化 · 定量生物学 2011-04-05 Naoki Masuda , Mitsuhiro Nakamura

Recent studies in the spatial prisoner's dilemma games with reinforcement learning have shown that static agents can learn to cooperate through a diverse sort of mechanisms, including noise injection, different types of learning algorithms…

人工智能 · 计算机科学 2025-07-08 Gustavo C. Mangold , Heitor C. M. Fernandes , Mendeli H. Vainstein

Cooperation is usually represented as a Prisoner's Dilemma game. Although individual self-interest may not favour cooperation, cooperation can evolve if, for example, players interact multiple times adjusting their behaviour accordingly to…

物理与社会 · 物理学 2015-04-29 Elton J. S. Júnior , Lucas Wardil , Jafferson K. L. da Silva

In this work, we study the social learning problem, in which agents of a networked system collaborate to detect the state of the nature based on their private signals. A novel distributed graphical evolutionary game theoretic learning…

计算机科学与博弈论 · 计算机科学 2017-05-24 Xuanyu Cao , K. J. Ray Liu

A generic property of biological, social and economical networks is their ability to evolve in time, creating and suppressing interactions. We approach this issue within the framework of an adaptive network of agents playing a Prisoner's…

适应与自组织系统 · 物理学 2014-10-20 Martin G. Zimmermann , Victor M. Eguiluz , Maxi San Miguel

We study an evolutionary prisoner's dilemma game with two layered graphs, where the lower layer is the physical infrastructure on which the interactions are taking place and the upper layer represents the connections for the strategy…

物理与社会 · 物理学 2009-11-13 Zhi-Xi Wu , Ying-Hai Wang

The Iterated Prisoner's Dilemma has guided research on social dilemmas for decades. However, it distinguishes between only two atomic actions: cooperate and defect. In real-world prisoner's dilemmas, these choices are temporally extended…

人工智能 · 计算机科学 2018-03-02 Weixun Wang , Jianye Hao , Yixi Wang , Matthew Taylor

Punishment is a common tactic to sustain cooperation and has been extensively studied for a long time. While most of previous game-theoretic work adopt the imitation learning where players imitate the strategies who are better off, the…

种群与进化 · 定量生物学 2024-12-20 Chenyang Zhao , Guozhong Zheng , Chun Zhang , Jiqiang Zhang , Li Chen

In the future, artificial learning agents are likely to become increasingly widespread in our society. They will interact with both other learning agents and humans in a variety of complex settings including social dilemmas. We consider the…

计算机科学与博弈论 · 计算机科学 2019-11-21 Tobias Baumann , Thore Graepel , John Shawe-Taylor

The relationship between a reinforcement learning (RL) agent and an asynchronous environment is often ignored. Frequently used models of the interaction between an agent and its environment, such as Markov Decision Processes (MDP) or…

人工智能 · 计算机科学 2018-06-29 Jaden B. Travnik , Kory W. Mathewson , Richard S. Sutton , Patrick M. Pilarski

Reinforcement Learning (RL) based methods have seen their paramount successes in solving serial decision-making and control problems in recent years. For conventional RL formulations, Markov Decision Process (MDP) and state-action-value…

机器学习 · 计算机科学 2020-06-09 Ziyao Zhang , Liang Ma , Kin K. Leung , Konstantinos Poularakis , Mudhakar Srivatsa

Exploiting others is beneficial individually but it could also be detrimental globally. The reverse is also true: a higher cooperation level may change the environment in a way that is beneficial for all competitors. To explore the possible…

物理与社会 · 物理学 2018-02-23 Attila Szolnoki , Xiaojie Chen

The evolutionary Prisoner's Dilemma Game (PDG) and the Snowdrift Game (SG) with preferential learning mechanism are studied in the Barab\'asi-Albert network. Simulation results demonstrate that the preferential learning of individuals…

物理与社会 · 物理学 2009-09-29 Jie Ren , Wen-Xu Wang , Gang Yan , Bing-Hong Wang

A simple model for cooperation between "selfish" agents, which play an extended version of the Prisoner's Dilemma(PD) game, in which they use arbitrary payoffs, is presented and studied. A continuous variable, representing the probability…

凝聚态物理 · 物理学 2009-11-10 H. Fort

As humans perceive and actively engage with the world, we adjust our decisions in response to shifting group dynamics and are influenced by social interactions. This study aims to identify which aspects of interaction affect…

物理与社会 · 物理学 2024-12-23 Lucila G. Alvarez-Zuzek , Laura Ferrarotti , Bruno Lepri , Riccardo Gallotti

Evolutionary game theory assumes that players replicate a highly scored player's strategy through genetic inheritance. However, when learning occurs culturally, it is often difficult to recognize someone's strategy just by observing the…

种群与进化 · 定量生物学 2021-07-01 Minjae Kim , Jung-Kyoo Choi , Seung Ki Baek