中文
相关论文

相关论文: Online Learning in Iterated Prisoner's Dilemma to …

200 篇论文

As large language models (LLMs) become increasingly capable of autonomous decision-making, they introduce new challenges and opportunities for human-AI cooperation in mixed-motive contexts. While prior research has primarily examined AI in…

人机交互 · 计算机科学 2025-05-29 Guanxuan Jiang , Shirao Yang , Yuyang Wang , Pan Hui

In the future, artificial learning agents are likely to become increasingly widespread in our society. They will interact with both other learning agents and humans in a variety of complex settings including social dilemmas. We argue that…

人工智能 · 计算机科学 2022-02-22 Tobias Baumann

The standard iterated prisoner's dilemma is an unrealistic model of social behaviour because it forces individuals to participate in the interaction. We analyse a model in which players have the option of ending their association. If the…

最优化与控制 · 数学 2007-05-23 L. A. Khodarinova , J. N. Webb

A simple model for cooperation between "selfish" agents, which play an extended version of the Prisoner's Dilemma(PD) game, in which they use arbitrary payoffs, is presented and studied. A continuous variable, representing the probability…

凝聚态物理 · 物理学 2009-11-10 H. Fort

Prisoner's Dilemma is a game theory model used to describe altruistic behavior seen in various populations. This theoretical game is important in understanding why a seemingly selfish strategy does persist and spread throughout a population…

物理与社会 · 物理学 2020-04-16 Sharon M. Cameron , Ariel Cintrón-Arias

The Prisoner's Dilemma game has a long history stretching across the social, biological, and physical sciences. In 2012, Press and Dyson developed a method for analyzing the mapping of the 8-dimensional strategy profile onto the…

物理与社会 · 物理学 2019-11-28 Robert D. Young

New ranking algorithms are continually being developed and refined, necessitating the development of efficient methods for evaluating these rankers. Online ranker evaluation focuses on the challenge of efficiently determining, from implicit…

信息检索 · 计算机科学 2016-08-23 Brian Brost , Yevgeny Seldin , Ingemar J. Cox , Christina Lioma

We studied the effect of three strategy updating rules in coevolving prisoner's dilemma games where agents (nodes) can switch both the strategy and social partners. Under two node-based strategy updating rules, strategy updating occurs…

物理与社会 · 物理学 2018-09-25 Hirofumi Takesue

The Prisoner's Dilemma (PD) deals with the cooperation/defection conflict between two agents. The agents are represented by a cell of $L \times L$ square lattice. The agents are initially randomly distributed according to a certain…

统计力学 · 物理学 2007-05-23 Ricardo Oliveira dos Santos Soares , Alexandre Souto Martinez

In the future, artificial learning agents are likely to become increasingly widespread in our society. They will interact with both other learning agents and humans in a variety of complex settings including social dilemmas. We consider the…

计算机科学与博弈论 · 计算机科学 2019-11-21 Tobias Baumann , Thore Graepel , John Shawe-Taylor

When people play a repeated game they usually try to anticipate their opponents' moves based on past observations, and then decide what action to take next. Behavioural economics studies the mechanisms by which strategic decisions are taken…

物理与社会 · 物理学 2012-04-20 Tobias Galla

Cooperative behavior lies at the very basis of human societies, yet its evolutionary origin remains a key unsolved puzzle. Whereas reciprocity or conditional cooperation is one of the most prominent mechanisms proposed to explain the…

物理与社会 · 物理学 2016-03-22 Giulio Cimini , Angel Sánchez

Traditional evolutionary game theory describes how certain strategy spreads throughout the system where individual player imitates the most successful strategy among its neighborhood. Accordingly, player doesn't have own authority to change…

多智能体系统 · 计算机科学 2016-04-14 Sundong Kim , Jin-Jae Lee

Pricing decisions are increasingly made by AI. Thanks to their ability to train with live market data while making decisions on the fly, deep reinforcement learning algorithms are especially effective in taking such pricing decisions. In…

人工智能 · 计算机科学 2021-07-06 Michael Schlechtinger , Damaris Kosack , Heiko Paulheim , Thomas Fetzer

Social dilemmas are situations where individuals face a temptation to increase their payoffs at a cost to total welfare. Building artificially intelligent agents that achieve good outcomes in these situations is important because many real…

人工智能 · 计算机科学 2018-03-05 Adam Lerer , Alexander Peysakhovich

Obtaining knowledge and skill achievement through peer learning can lead to higher academic achievement. However, peer learning implementation is not just about putting students together and hoping for the best. At its worst-designed, peer…

计算机与社会 · 计算机科学 2019-10-29 Seyede Fatemeh Noorani , Mohammad Hossein Manshaei , Mohammad Ali Montazeri , Behnaz Omoomi

Given the increase in cybercrime, cybersecurity analysts (i.e. Defenders) are in high demand. Defenders must monitor an organization's network to evaluate threats and potential breaches into the network. Adversary simulation is commonly…

密码学与安全 · 计算机科学 2023-04-04 Baptiste Prebot , Yinuo Du , Cleotilde Gonzalez

We introduce a framework for decentralized online learning for multi-armed bandits (MAB) with multiple cooperative players. The reward obtained by the players in each round depends on the actions taken by all the players. It's a team…

机器学习 · 计算机科学 2021-09-10 William Chang , Mehdi Jafarnia-Jahromi , Rahul Jain

We propose a new framework for imitation learning -- treating imitation as a two-player ranking-based game between a policy and a reward. In this game, the reward agent learns to satisfy pairwise performance rankings between behaviors,…

机器学习 · 计算机科学 2023-01-18 Harshit Sikchi , Akanksha Saran , Wonjoon Goo , Scott Niekum

The problem of reinforcement learning is considered where the environment or the model undergoes a change. An algorithm is proposed that an agent can apply in such a problem to achieve the optimal long-time discounted reward. The algorithm…

系统与控制 · 电气工程与系统科学 2023-04-25 Wuxia Chen , Taposh Banerjee , Jemin George , Carl Busart