中文
相关论文

相关论文: Experience-replay Innovative Dynamics

200 篇论文

In evolutionary game theory, it is customary to be partial to the dynamical models possessing fixed points so that they may be understood as the attainment of evolutionary stability, and hence, Nash equilibrium. Any show of periodic or…

种群与进化 · 定量生物学 2021-02-23 Archan Mukhopadhyay , Sagar Chakraborty

If a game has a unique Nash equilibrium, then this equilibrium is arguably the solution of the game from the refinement's literature point of view. However, it might be that for almost all initial conditions, all strategies in the support…

计算机科学与博弈论 · 计算机科学 2012-11-26 Yannick Viossat

Experience replay is a core ingredient of modern deep reinforcement learning, yet its benefits in policy optimization are poorly understood beyond empirical heuristics. This paper develops a novel theoretical framework for experience replay…

机器学习 · 计算机科学 2026-02-04 Hua Zheng , Wei Xie , M. Ben Feng

Multiagent reinforcement learning (MARL) is commonly considered to suffer from non-stationary environments and exponentially increasing policy space. It would be even more challenging when rewards are sparse and delayed over long…

This paper aims to accommodate games in which the players' dynamics are subject to un-modeled and disturbance terms. The un-modeled and disturbance terms are regarded as extended states for which observers are designed to estimate them.…

最优化与控制 · 数学 2020-04-22 Maojiao Ye

This paper studies random reshuffling (RR)-based distributed Nash equilibrium seeking for noncooperative games. The game is motivated as a sample-average approximation of an underlying expected-value stochastic game, while the algorithmic…

最优化与控制 · 数学 2026-04-06 Jun Hu , Chao Sun , Chen Bo , Jianzheng Wang , Zheming Wang

Self-play is a technique for machine learning in multi-agent systems where a learning algorithm learns by interacting with copies of itself. Self-play is useful for generating large quantities of data for learning, but has the drawback that…

计算机科学与博弈论 · 计算机科学 2023-11-30 Revan MacQueen , James R. Wright

We consider learning Nash equilibria in two-player zero-sum Markov Games with nonlinear function approximation, where the action-value function is approximated by a function in a Reproducing Kernel Hilbert Space (RKHS). The key challenge is…

机器学习 · 计算机科学 2022-08-11 Chris Junchi Li , Dongruo Zhou , Quanquan Gu , Michael I. Jordan

The framework of multi-agent learning explores the dynamics of how individual agent strategies evolve in response to the evolving strategies of other agents. Of particular interest is whether or not agent strategies converge to well known…

计算机科学与博弈论 · 计算机科学 2023-11-21 Sarah A. Toonsi , Jeff S. Shamma

Game theory provides a well-established framework for the analysis of concurrent and multi-agent systems. The basic idea is that concurrent processes (agents) can be understood as corresponding to players in a game; plays represent the…

计算机科学中的逻辑 · 计算机科学 2023-06-22 Julian Gutierrez , Paul Harrenstein , Giuseppe Perelli , Michael Wooldridge

Multi-agent reinforcement learning (MARL) plays a pivotal role in tackling real-world challenges. However, the seamless transition of trained policies from simulations to real-world requires it to be robust to various environmental…

机器学习 · 计算机科学 2023-10-16 Aakriti Agrawal , Rohith Aralikatti , Yanchao Sun , Furong Huang

In this letter, we study dynamic game optimal control with imperfect state observations and introduce an iterative method to find a local Nash equilibrium. The algorithm consists of an iterative procedure combining a backward recursion…

最优化与控制 · 数学 2022-06-24 Armand Jordana , Bilal Hammoud , Justin Carpentier , Ludovic Righetti

Motivated by the scarcity of accurate payoff feedback in practical applications of game theory, we examine a class of learning dynamics where players adjust their choices based on past payoff observations that are subject to noise and…

最优化与控制 · 数学 2016-06-03 Mario Bravo , Panayotis Mertikopoulos

Agent faults pose a significant threat to the performance of multi-agent reinforcement learning (MARL) algorithms, introducing two key challenges. First, agents often struggle to extract critical information from the chaotic state space…

机器学习 · 计算机科学 2024-12-03 Yuchen Shi , Huaxin Pei , Liang Feng , Yi Zhang , Danya Yao

Many emerging agentic paradigms require agents to collaborate with one another (or people) to achieve shared goals. Unfortunately, existing approaches to learning policies for such collaborative problems produce brittle solutions that fail…

机器学习 · 计算机科学 2026-03-02 Chengrui Qu , Yizhou Zhang , Nicolas Lanzetti , Eric Mazumdar

This work studies the application of Multi-Agent Reinforcement Learning (MARL) to decentralized control of unmanned aerial vehicles to relay a critical data package to a known position. For this purpose, a family of deterministic games is…

系统与控制 · 电气工程与系统科学 2026-05-11 Mika Persson , Jonas Lidman , Jacob Ljungberg , Samuel Sandelius , Adam Andersson

We study multi-agent reinforcement learning (MARL) in a stochastic network of agents. The objective is to find localized policies that maximize the (discounted) global reward. In general, scalability is a challenge in this setting because…

机器学习 · 计算机科学 2021-11-03 Yiheng Lin , Guannan Qu , Longbo Huang , Adam Wierman

While Experience Replay - the practice of storing rollouts and reusing them multiple times during training - is a foundational technique in general RL, it remains largely unexplored in LLM post-training due to the prevailing belief that…

机器学习 · 计算机科学 2026-04-13 Charles Arnal , Vivien Cabannes , Taco Cohen , Julia Kempe , Remi Munos

Inferring reward functions from demonstrations is a key challenge in reinforcement learning (RL), particularly in multi-agent RL (MARL), where large joint state-action spaces and complex inter-agent interactions complicate the task. While…

机器学习 · 计算机科学 2025-02-03 The Viet Bui , Tien Mai , Hong Thanh Nguyen

We consider a game-theoretic model of information retrieval with strategic authors. We examine two different utility schemes: authors who aim at maximizing exposure and authors who want to maximize active selection of their content (i.e.…

计算机科学与博弈论 · 计算机科学 2019-02-21 Omer Ben-Porat , Itay Rosenberg , Moshe Tennenholtz