中文
相关论文

相关论文: Aspiration Learning in Coordination Games

200 篇论文

The success of modern civilization is built upon widespread cooperation in human society, deciphering the mechanisms behind has being a major goal for centuries. A crucial fact is, however, largely missing in most prior studies that games…

物理与社会 · 物理学 2021-06-03 Qinqin Wang , Rizhou Liang , Jiqiang Zhang , Guozhong Zheng , Lin Ma , Li Chen

We consider average-energy games, where the goal is to minimize the long-run average of the accumulated energy. While several results have been obtained on these games recently, decidability of average-energy games with a lower-bound…

计算机科学中的逻辑 · 计算机科学 2017-01-16 Patricia Bouyer , Piotr Hofman , Nicolas Markey , Mickael Randour , Martin Zimmermann

Due to the large size of the training data, distributed learning approaches such as federated learning have gained attention recently. However, the convergence rate of distributed learning suffers from heterogeneous worker performance. In…

分布式、并行与集群计算 · 计算机科学 2019-08-09 Yunus Sarikaya , Ozgur Ercetin

This paper investigates the discrete-time asynchronous games in which noncooperative agents seek to minimize their individual cost functions. Building on the assumption of partial asynchronism, i.e., each agent updates at least once within…

最优化与控制 · 数学 2025-08-13 Zifan Wang , Xinlei Yi , Michael M. Zavlanos , Karl H. Johansson

The latest developments in AI focus on agentic systems where artificial and human agents cooperate to realize global goals. An example is collaborative learning, which aims to train a global model based on data from individual agents. A…

计算机科学与博弈论 · 计算机科学 2025-08-20 Björn Filter , Ralf Möller , Özgür Lütfü Özçep

We study the fragmentation-coagulation (or merging and splitting) evolutionary control model as introduced recently by one of the authors, where $N$ small players can form coalitions to resist to the pressure exerted by the principal. It is…

最优化与控制 · 数学 2022-05-03 Alekos Cecchin , Vassili N. Kolokoltsov

We show that learning algorithms satisfying a $\textit{low approximate regret}$ property experience fast convergence to approximate optimality in a large class of repeated games. Our property, which simply requires that each learner has…

计算机科学与博弈论 · 计算机科学 2016-12-19 Dylan J. Foster , Zhiyuan Li , Thodoris Lykouris , Karthik Sridharan , Eva Tardos

This paper focuses on "tracing player knowledge" in educational games. Specifically, given a set of concepts or skills required to master a game, the goal is to estimate the likelihood with which the current player has mastery of each of…

人工智能 · 计算机科学 2019-08-16 Pavan Kantharaju , Katelyn Alderfer , Jichen Zhu , Bruce Char , Brian Smith , Santiago Ontañón

We focus on the task of goal-oriented grasping, in which a robot is supposed to grasp a pre-assigned goal object in clutter and needs some pre-grasp actions such as pushes to enable stable grasps. However, in this task, the robot gets…

机器人学 · 计算机科学 2021-06-24 Kechun Xu , Hongxiang Yu , Qianen Lai , Yue Wang , Rong Xiong

This work proposes a novel distributed approach for computing a Nash equilibrium in convex games with merely monotone and restricted strongly monotone pseudo-gradients. By leveraging the idea of the centralized operator extrapolation method…

最优化与控制 · 数学 2025-07-18 Tatiana Tatarenko , Angelia Nedich

Previous studies suggest that punishment is a useful way to promote cooperation in the well-mixed public goods game, whereas it still lacks specific evidence that punishment maintains cooperation in spatial prisoner's dilemma game as well.…

物理与社会 · 物理学 2011-03-08 Qing Jin , Zhen Wang , Zhen Wang , Yi-Ling Wang

We propose an adaptive incentive mechanism that learns the optimal incentives in environments where players continuously update their strategies. Our mechanism updates incentives based on each player's externality, defined as the difference…

计算机科学与博弈论 · 计算机科学 2025-03-04 Chinmay Maheshwari , Kshitij Kulkarni , Manxi Wu , Shankar Sastry

The sustainable use of common-pool resources (CPRs) is a major environmental governance challenge because of their possible over-exploitation. Research in this field has overlooked the feedback between user decisions and resource dynamics.…

理论经济学 · 经济学 2021-10-04 Chengyi Tu , Paolo DOdorico , Zhe Li , Samir Suweis

Multi-agent reinforcement learning in mixed-motive settings presents a fundamental challenge: agents must balance individual interests with collective goals, which are neither fully aligned nor strictly opposed. To address this, reward…

多智能体系统 · 计算机科学 2025-08-26 Woojun Kim , Katia Sycara

Federated learning is a distributed machine learning system that uses participants' data to train an improved global model. In federated learning, participants cooperatively train a global model, and they will receive the global model and…

计算机科学与博弈论 · 计算机科学 2023-09-27 Mengda Ji , Genjiu Xu , Jianjun Ge , Mingqiang Li

This paper considers a networked aggregative game (NAG) where the players are distributed over a communication network. By only communicating with a subset of players, the goal of each player in the NAG is to minimize an individual cost…

最优化与控制 · 数学 2021-05-13 Rongping Zhu , Jiaqi Zhang , Keyou You

We study the problem of learning Markov decision processes with finite state and action spaces when the transition probability distributions and loss functions are chosen adversarially and are allowed to change with time. We introduce an…

机器学习 · 计算机科学 2013-03-14 Yasin Abbasi-Yadkori , Peter L. Bartlett , Csaba Szepesvari

Imitation dynamics for population games are studied and their asymptotic properties analyzed. In the considered class of imitation dynamics - that encompass the replicator equation as well as other models previously considered in…

系统与控制 · 计算机科学 2021-03-02 Lorenzo Zino , Giacomo Como , Fabio Fagnani

This paper studies the finite-time horizon Markov games where the agents' dynamics are decoupled but the rewards can possibly be coupled across agents. The policy class is restricted to local policies where agents make decisions using their…

计算机科学与博弈论 · 计算机科学 2023-04-11 Runyu Zhang , Yuyang Zhang , Rohit Konda , Bryce Ferguson , Jason Marden , Na Li

Often adaptive, distributed control can be viewed as an iterated game between independent players. The coupling between the players' mixed strategies, arising as the system evolves from one instant to the next, is determined by the system…

多智能体系统 · 计算机科学 2007-05-23 David H. Wolpert , Stefan Bieniawski