中文
相关论文

相关论文: Logit Dynamics with Concurrent Updates for Local-I…

200 篇论文

Reinforcement learning (RL) has recently achieved tremendous successes in many artificial intelligence applications. Many of the forefront applications of RL involve multiple agents, e.g., playing chess and Go games, autonomous driving, and…

计算机科学与博弈论 · 计算机科学 2021-11-24 Asuman Ozdaglar , Muhammed O. Sayin , Kaiqing Zhang

Evolutionary game theory assumes that players replicate a highly scored player's strategy through genetic inheritance. However, when learning occurs culturally, it is often difficult to recognize someone's strategy just by observing the…

种群与进化 · 定量生物学 2021-07-01 Minjae Kim , Jung-Kyoo Choi , Seung Ki Baek

We analyzes the logit dynamics of softmax policy gradient methods. We derive the exact formula for the L2 norm of the logit update vector: $$ \|\Delta \mathbf{z}\|_2 \propto \sqrt{1-2P_c + C(P)} $$ This equation demonstrates that update…

机器学习 · 计算机科学 2025-06-17 Yingru Li

I study dynamic network formation games in which agents meet stochastically and form links based on their valuation of the network. I show that these games can be represented in terms of the values agents assign to network sub-structures.…

理论经济学 · 经济学 2025-11-19 Jose M. Betancourt

We study the problem of stochastic stability for evolutionary dynamics under the logit choice rule. We consider general classes of coordination games, symmetric or asymmetric, with an arbitrary number of strategies, which satisfies the…

计算机科学与博弈论 · 计算机科学 2021-01-13 Sung-Ha Hwang , Luc Rey-Bellet

This technical note presents a leader-follower scheme for network aggregative games. The followers and leader are selfish cost minimizing agents. The cost function of each follower is affected by strategy of leader and aggregated strategies…

系统与控制 · 电气工程与系统科学 2019-08-08 Mohammad Shokri , Hamed Kebriaei

In many settings of interest, a policy is set by one party, the leader, in order to influence the action of another party, the follower, where the follower's response is determined by some private information. A natural question to ask is,…

计算机科学与博弈论 · 计算机科学 2025-04-23 Michael Albert , Quinlan Dawkins , Minbiao Han , Haifeng Xu

How can a social planner adaptively incentivize selfish agents who are learning in a strategic environment to induce a socially optimal outcome in the long run? We propose a two-timescale learning dynamics to answer this question in both…

计算机科学与博弈论 · 计算机科学 2022-04-13 Chinmay Maheshwari , Kshitij Kulkarni , Manxi Wu , Shankar Sastry

In this paper, we consider a learning problem among non-cooperative agents interacting in a time-varying system. Specifically, we focus on repeated linear quadratic network games, in which the network of interactions changes with time and…

计算机科学与博弈论 · 计算机科学 2023-10-23 Feras Al Taha , Kiran Rokade , Francesca Parise

Evolutionary game theory has been widely used to study the evolution of cooperation in social dilemmas where imitation-led strategy updates are typically assumed. However, results of recent behavioral experiments are not compatible with the…

物理与社会 · 物理学 2018-12-19 Ik Soo Lim , Peter Wittek

We report on new stability conditions for evolutionary dynamics in the context of population games. We adhere to the prevailing framework consisting of many agents, grouped into populations, that interact noncooperatively by selecting…

种群与进化 · 定量生物学 2021-07-08 Semih Kara , Nuno C. Martins

We study strategic games on weighted directed graphs, where the payoff of a player is defined as the sum of the weights on the edges from players who chose the same strategy augmented by a fixed non-negative bonus for picking a given…

计算机科学与博弈论 · 计算机科学 2016-04-19 Sunil Simon , Dominik Wojtczak

We studied the effect of three strategy updating rules in coevolving prisoner's dilemma games where agents (nodes) can switch both the strategy and social partners. Under two node-based strategy updating rules, strategy updating occurs…

物理与社会 · 物理学 2018-09-25 Hirofumi Takesue

Direct reciprocity is a mechanism for the evolution of cooperation based on repeated interactions. When individuals meet repeatedly, they can use conditional strategies to enforce cooperative outcomes that would not be feasible in one-shot…

种群与进化 · 定量生物学 2016-05-24 Seung Ki Baek , Hyeong-Chai Jeong , Christian Hilbe , Martin A. Nowak

The hierarchical interaction between the actor and critic in actor-critic based reinforcement learning algorithms naturally lends itself to a game-theoretic interpretation. We adopt this viewpoint and model the actor and critic interaction…

机器学习 · 计算机科学 2021-09-28 Liyuan Zheng , Tanner Fiez , Zane Alumbaugh , Benjamin Chasnov , Lillian J. Ratliff

We analyze independent policy-gradient (PG) learning in $N$-player linear-quadratic (LQ) stochastic differential games. Each player employs a distributed policy that depends only on its own state and updates the policy independently using…

最优化与控制 · 数学 2026-02-19 Philipp Plank , Yufei Zhang

We consider concurrent games played on graphs. At every round of a game, each player simultaneously and independently selects a move; the moves jointly determine the transition to a successor state. Two basic objectives are the safety…

计算机科学与博弈论 · 计算机科学 2012-07-03 Krishnendu Chatterjee , Luca de Alfaro , Thomas A. Henzinger

Most real-world social networks are inherently dynamic, composed of communities that are constantly changing in membership. To track these evolving communities, we need dynamic community detection techniques. This article evaluates the…

社会与信息网络 · 计算机科学 2016-09-13 Hamidreza Alvari , Alireza Hajibagheri , Gita Sukthankar , Kiran Lakkaraju

Interactions between people are the basis on which the structure of our society arises as a complex system and, at the same time, are the starting point of any physical description of it. In the last few years, much theoretical research has…

计算机科学与博弈论 · 计算机科学 2017-12-06 Mattia Mazzoli , Angel Sanchez

Evolutionary games are studied where the teaching activity of players can evolve in time. Initially all players following either the cooperative or defecting strategy are distributed on a square lattice. The rate of strategy adoption is…

物理与社会 · 物理学 2008-04-23 Attila Szolnoki , Matjaz Perc