中文
相关论文

相关论文: Robust Accelerated Dynamics for Subnetwork Bilinea…

200 篇论文

We study reinforcement learning for two-player zero-sum Markov games with simultaneous moves in the finite-horizon setting, where the transition kernel of the underlying Markov games can be parameterized by a linear function over the…

机器学习 · 计算机科学 2022-04-21 Zixiang Chen , Dongruo Zhou , Quanquan Gu

We propose a fully asynchronous networked aggregative game (Asy-NAG) where each player minimizes a cost function that depends on its local action and the aggregate of all players' actions. In sharp contrast to the existing NAGs, each player…

最优化与控制 · 数学 2021-01-25 Rongping Zhu , Jiaqi Zhang , Keyou You , Tamer Başar

Zero-sum stochastic games are easy to solve as they can be cast as simple Markov decision processes. This is however not the case with general-sum stochastic games. A fairly general optimization problem formulation is available for…

机器学习 · 计算机科学 2015-07-02 H. L. Prasad , Shalabh Bhatnagar

A key feature of wireless communications is the spatial reuse. However, the spatial aspect is not yet well understood for the purpose of designing efficient spectrum sharing mechanisms. In this paper, we propose a framework of spatial…

网络与互联网体系结构 · 计算机科学 2014-05-19 Xu Chen , Jianwei Huang

This paper explores distributed Nash equilibrium seeking problems for games in which the players have limited knowledge on other players' actions. In particular, the involved players are considered to be high-order integrators with their…

最优化与控制 · 数学 2021-08-17 Maojiao Ye , Lei Ding , Shengyuan Xu

We study the stochastic bilinear minimax optimization problem, presenting an analysis of the same-sample Stochastic ExtraGradient (SEG) method with constant step size, and presenting variations of the method that yield favorable…

最优化与控制 · 数学 2022-04-11 Chris Junchi Li , Yaodong Yu , Nicolas Loizou , Gauthier Gidel , Yi Ma , Nicolas Le Roux , Michael I. Jordan

We consider a class of Nash games, termed as aggregative games, being played over a networked system. In an aggregative game, a player's objective is a function of the aggregate of all the players' decisions. Every player maintains an…

最优化与控制 · 数学 2016-06-10 Jayash Koshal , Angelia Nedić , Uday V. Shanbhag

This paper explores aggregative games in a network of general linear systems subject to external disturbances. To deal with external disturbances, distributed strategy-updating rules based on internal model are proposed for the case with…

最优化与控制 · 数学 2024-10-28 Xin Cai , Feng Xiao , Bo Wei , Mei Yu , Fang Fang

Towards characterizing the optimization landscape of games, this paper analyzes the stability of gradient-based dynamics near fixed points of two-player continuous games. We introduce the quadratic numerical range as a method to…

计算机科学与博弈论 · 计算机科学 2021-01-15 Benjamin J. Chasnov , Daniel Calderone , Behçet Açıkmeşe , Samuel A. Burden , Lillian J. Ratliff

We introduce DREAM, a deep reinforcement learning algorithm that finds optimal strategies in imperfect-information games with multiple agents. Formally, DREAM converges to a Nash Equilibrium in two-player zero-sum games and to an…

机器学习 · 计算机科学 2020-12-01 Eric Steinberger , Adam Lerer , Noam Brown

This paper presents algorithms for non-zero sum nonlinear constrained dynamic games with full information. Such problems emerge when multiple players with action constraints and differing objectives interact with the same dynamic system.…

系统与控制 · 电气工程与系统科学 2020-01-08 Bolei Di , Andrew Lamperski

For a linear equality constrained convex optimization problem involving two objective functions with a ``nonsmooth" + ``nonsmooth" composite structure, we study two algorithms derived from a mixed-order dynamical system which incorporates…

最优化与控制 · 数学 2026-03-25 Geng-Hua Li , Hai-Yi Zhao , Xiangkai Sun

In this paper, we investigate a prescribed-time and fully distributed Nash Equilibrium (NE) seeking problem for continuous-time noncooperative games. By exploiting pseudo-gradient play and consensus-based schemes, various distributed NE…

系统与控制 · 电气工程与系统科学 2020-09-25 Zhi Feng , Guoqiang Hu

Based on the idea of randomized coordinate descent of $\alpha$-averaged operators, a randomized primal-dual optimization algorithm is introduced, where a random subset of coordinates is updated at each iteration. The algorithm builds upon a…

最优化与控制 · 数学 2015-10-01 Pascal Bianchi , Walid Hachem , Franck Iutzeler

In this work, we establish near-linear and strong convergence for a natural first-order iterative algorithm that simulates Von Neumann's Alternating Projections method in zero-sum games. First, we provide a precise analysis of Optimistic…

最优化与控制 · 数学 2021-08-18 Ioannis Anagnostides , Paolo Penna

In this work, we investigate the distributed generalized Nash equilibrium (GNE) seeking problems for $N$-coalition games with inequality constraints. First, we study the scenario where each agent in a coalition has full information of all…

最优化与控制 · 数学 2021-09-28 Chao Sun , Guoqiang Hu

We present a new, distributed method to compute approximate Nash equilibria in bimatrix games. In contrast to previous approaches that analyze the two payoff matrices at the same time (for example, by solving a single LP that combines the…

计算机科学与博弈论 · 计算机科学 2018-10-12 Artur Czumaj , Argyrios Deligkas , Michail Fasoulakis , John Fearnley , Marcin Jurdziński , Rahul Savani

We present a hierarchical model predictive control approach for large-scale systems based on dual decomposition. The proposed scheme allows coupling in both dynamics and constraints between the subsystems and generates a primal feasible…

最优化与控制 · 数学 2011-11-10 Minh Dang Doan , Tamás Keviczky , Bart De Schutter

We study data corruption robustness in offline two-player zero-sum Markov games. Given a dataset of realized trajectories of two players, an adversary is allowed to modify an $\epsilon$-fraction of it. The learner's goal is to identify an…

计算机科学与博弈论 · 计算机科学 2024-03-14 Andi Nika , Debmalya Mandal , Adish Singla , Goran Radanović

This paper studies distributed online bandit learning of generalized Nash equilibria for online game, where cost functions of all players and coupled constraints are time-varying. The values rather than full information of cost and local…

最优化与控制 · 数学 2022-04-21 Min Meng , Xiuxian Li , Jie Chen