中文
相关论文

相关论文: The State-Action-Reward-State-Action Algorithm in …

200 篇论文

Reinforcement learning (RL) algorithms allow artificial agents to improve their selection of actions to increase rewarding experiences in their environments. Temporal Difference (TD) Learning -- a model-free RL method -- is a leading…

机器学习 · 计算机科学 2019-09-05 Jacob Rafati , David C. Noelle

The evolution of cooperation in networked systems helps to understand the dynamics in social networks, multi-agent systems, and biological species. The self-persistence of individual strategies is common in real-world decision making. The…

社会与信息网络 · 计算机科学 2025-11-25 Ziyan Zeng , Minyu Feng , Attila Szolnoki

As artificial intelligence (AI) systems are increasingly embedded in our lives, their presence leads to interactions that shape our behaviour, decision-making, and social interactions. Existing theoretical research has primarily focused on…

Understanding the emergence of prosocial behaviours (e.g., cooperation and trust) among self-interested agents is an important problem in many disciplines. Network structure and institutional incentives (e.g., punishing antisocial agents)…

物理与社会 · 物理学 2022-03-29 Ik Soo Lim , Valerio Capraro

We study the evolution of cooperation in the evolutionary spatial prisoner's dilemma game (PDG) and snowdrift game (SG), within which a fraction $\alpha$ of the payoffs of each player gained from direct game interactions is shared equally…

物理与社会 · 物理学 2014-01-13 Zhi-Xi Wu , Han-Xin Yang

Previous studies suggest that punishment is a useful way to promote cooperation in the well-mixed public goods game, whereas it still lacks specific evidence that punishment maintains cooperation in spatial prisoner's dilemma game as well.…

物理与社会 · 物理学 2011-03-08 Qing Jin , Zhen Wang , Zhen Wang , Yi-Ling Wang

Collective behavior, and swarm formation in particular, has been studied from several perspectives within a large variety of fields, ranging from biology to physics. In this work, we apply Projective Simulation to model each individual as…

种群与进化 · 定量生物学 2021-01-27 Andrea López-Incera , Katja Ried , Thomas Müller , Hans J. Briegel

We consider a dual model of decision making, in which an individual forms its opinion based on contrasting mechanisms of imitation and rational calculation. The decision making model (DMM) implements imitating behavior by means of a network…

适应与自组织系统 · 物理学 2015-06-22 Malgorzata Turalska , Bruce J. West

Commitment is a well-established mechanism for fostering cooperation in human society and multi-agent systems. However, existing research has predominantly focused on the commitment that neglects the freedom of players to abstain from an…

计算机科学与博弈论 · 计算机科学 2025-08-12 Zhao Song , The Anh Han

We seek to align agent policy with human expert behavior in a reinforcement learning (RL) setting, without any prior knowledge about dynamics, reward function, and unsafe states. There is a human expert knowing the rewards and unsafe states…

机器学习 · 计算机科学 2020-01-01 Daniel Hsu

We introduce a class of learning problems where the agent is presented with a series of tasks. Intuitively, if there is relation among those tasks, then the information gained during execution of one task has value for the execution of…

机器学习 · 计算机科学 2012-09-06 Christos Dimitrakakis

Punishment and partner switching are two well-studied mechanisms that support the evolution of cooperation. Observation of human behaviour suggests that the extent to which punishment is adopted depends on the usage of alternative…

物理与社会 · 物理学 2018-04-26 Hirofumi Takesue

The preferential treatment of in-group members is widely observed. This study examines this phenomenon in the domain of cooperation in social dilemmas using evolutionary agent-based models that consider the role of partner selection. The…

物理与社会 · 物理学 2024-03-12 Hirofumi Takesue

We use analytical techniques based on an expansion in the inverse system size to study the stochastic evolutionary dynamics of finite populations of players interacting in a repeated prisoner's dilemma game. We show that a mechanism of…

种群与进化 · 定量生物学 2012-04-20 Alex J. Bladon , Tobias Galla , Alan J. McKane

Game theory formalizes certain interactions between physical particles or between living beings in biology, sociology, and economics, and quantifies the outcomes by payoffs. The prisoner's dilemma (PD) describes situations in which it is…

物理与社会 · 物理学 2015-05-13 Dirk Helbing , Sergi Lozano

This paper explores human behavior in virtual networked communities, specifically individuals or groups' potential and expressive capacity to respond to internal and external stimuli, with assortative matching as a typical example. A…

多智能体系统 · 计算机科学 2023-09-06 Ou Deng , Qun Jin

The challenge of developing powerful and general Reinforcement Learning (RL) agents has received increasing attention in recent years. Much of this effort has focused on the single-agent setting, in which an agent maximizes a predefined…

机器学习 · 计算机科学 2020-10-21 Jiachen Yang , Ang Li , Mehrdad Farajtabar , Peter Sunehag , Edward Hughes , Hongyuan Zha

Humanity faces numerous problems of common-pool resource appropriation. This class of multi-agent social dilemma includes the problems of ensuring sustainable use of fresh water, common fisheries, grazing pastures, and irrigation systems.…

多智能体系统 · 计算机科学 2017-09-07 Julien Perolat , Joel Z. Leibo , Vinicius Zambaldi , Charles Beattie , Karl Tuyls , Thore Graepel

The theoretical description of the evolution of cooperation presented by Bergstrom based on assortative matching with partner choice allows to model the population dynamics in a game of Nonrepetitive Prisoners Dilemma. In this paper we…

无序系统与神经网络 · 物理学 2007-05-23 Pawel Sobkowicz

Reinforcement Learning (RL) agents often struggle with inefficient exploration, particularly in environments with sparse rewards. Traditional exploration strategies can lead to slow learning and suboptimal performance because agents fail to…

机器学习 · 计算机科学 2026-03-31 Gaurav Chaudhary , Laxmidhar Behera , Washim Uddin Mondal