中文
相关论文

相关论文: Better Regularization for Sequential Decision Spac…

200 篇论文

In this paper, a new method is proposed to compute the rolling Nash equilibrium of the time-invariant nonlinear two-person zero-sum differential games. The idea is to discretize the time to transform a differential game into a sequential…

系统与控制 · 电气工程与系统科学 2020-11-13 Wei Liao , Xiaohui Wei , Jizhou Lai

This paper proposes a distributed algorithm to find the Nash equilibrium in a class of non-cooperative convex games with partial-decision information. Our method employs a distributed projected gradient play approach alongside consensus…

计算机科学与博弈论 · 计算机科学 2024-12-13 Duong Thuy Anh Nguyen , Duong Tung Nguyen , Angelia Nedić

Non-cooperative dynamic game theory provides a principled approach to modeling sequential decision-making among multiple noncommunicative agents. A key focus has been on finding Nash equilibria in two-agent zero-sum dynamic games under…

计算机科学与博弈论 · 计算机科学 2025-03-20 Kushagra Gupta , Ross Allen , David Fridovich-Keil , Ufuk Topcu

Nash equilibrium is one of the most influential solution concepts in game theory. With the development of computer science and artificial intelligence, there is an increasing demand on Nash equilibrium computation, especially for Internet…

计算机科学与博弈论 · 计算机科学 2023-12-19 Hanyu Li , Wenhan Huang , Zhijian Duan , David Henry Mguni , Kun Shao , Jun Wang , Xiaotie Deng

In this work, we study the distributed Nash equilibrium seeking problem for monotone generalized noncooperative games with set constraints and shared affine inequality constraints. A distributed regularized penalty method is proposed. The…

最优化与控制 · 数学 2021-09-28 Chao Sun , Guoqiang Hu

In this paper, we consider the problem of learning a generalized Nash equilibrium (GNE) in strongly monotone games. First, we propose a novel continuous-time solution algorithm that uses regular projections and first-order information. As…

系统与控制 · 电气工程与系统科学 2020-07-23 Suad Krilašević , Sergio Grammatico

Game theory has emerged as a powerful framework for modeling a large range of multi-agent scenarios. Many algorithmic solutions require discrete, finite games with payoffs that have a closed-form specification. In contrast, many real-world…

计算机科学与博弈论 · 计算机科学 2018-06-13 Abdullah Al-Dujaili , Erik Hemberg , Una-May O'Reilly

We study the global convergence of policy optimization for finding the Nash equilibria (NE) in zero-sum linear quadratic (LQ) games. To this end, we first investigate the landscape of LQ games, viewing it as a nonconvex-nonconcave…

机器学习 · 计算机科学 2021-02-12 Kaiqing Zhang , Zhuoran Yang , Tamer Başar

Equilibrium solution concepts of normal-form games, such as Nash equilibria, correlated equilibria, and coarse correlated equilibria, describe the joint strategy profiles from which no player has incentive to unilaterally deviate. They are…

计算机科学与博弈论 · 计算机科学 2023-04-21 Luke Marris , Ian Gemp , Georgios Piliouras

We consider zero-sum repeated games in which the players are restricted to strategies that require only a limited amount of randomness. Let $v_n$ be the max-min value of the $n$ stage game; previous works have characterized…

计算机科学与博弈论 · 计算机科学 2019-02-12 Mehrdad Valizadeh , Amin Gohari

In this paper, we propose a numerical methodology for finding the closed-loop Nash equilibrium of stochastic delay differential games through deep learning. These games are prevalent in finance and economics where multi-agent interaction…

最优化与控制 · 数学 2023-07-14 Robert Balkin , Hector D. Ceniceros , Ruimeng Hu

We introduce a novel class of Nash equilibrium seeking dynamics for non-cooperative games with a finite number of players, where the convergence to the Nash equilibrium is bounded by a KL function with a settling time that can be upper…

最优化与控制 · 数学 2020-12-25 Jorge I. Poveda , Miroslav Krstic , Tamer Basar

We consider generalized Nash equilibrium problems (GNEPs) with non-convex strategy spaces and non-convex cost functions. This general class of games includes the important case of games with mixed-integer variables for which only a few…

计算机科学与博弈论 · 计算机科学 2024-04-10 Tobias Harks , Julian Schwarz

A fundamental open problem in monotone game theory is the computation of a specific generalized Nash equilibrium (GNE) among all the available ones, e.g. the optimal equilibrium with respect to a system-level objective. The existing GNE…

系统与控制 · 电气工程与系统科学 2022-03-16 Emilio Benenati , Wicak Ananduta , Sergio Grammatico

Imperfect-recall abstraction has emerged as the leading paradigm for practical large-scale equilibrium computation in incomplete-information games. However, imperfect-recall abstractions are poorly understood, and only weak…

计算机科学与博弈论 · 计算机科学 2016-06-07 Christian Kroer , Tuomas Sandholm

In this paper, we investigate the power of {\it regularization}, a common technique in reinforcement learning and optimization, in solving extensive-form games (EFGs). We propose a series of new algorithms based on regularizing the payoff…

计算机科学与博弈论 · 计算机科学 2025-07-10 Mingyang Liu , Asuman Ozdaglar , Tiancheng Yu , Kaiqing Zhang

We analyze best response dynamics for finding a Nash equilibrium of an infinite horizon zero-sum stochastic linear quadratic dynamic game (LQDG) with partial and asymmetric information. We derive explicit expressions for each player's best…

系统与控制 · 电气工程与系统科学 2025-09-03 Yuxiang Guan , Iman Shames , Tyler H. Summers

We investigate the complexity of computing approximate Nash equilibria in anonymous games. Our main algorithmic result is the following: For any $n$-player anonymous game with a bounded number of strategies and any constant $\delta>0$, an…

计算机科学与博弈论 · 计算机科学 2016-08-29 Yu Cheng , Ilias Diakonikolas , Alistair Stewart

In this paper, Nash equilibrium seeking among a network of players is considered. Different from many existing works on Nash equilibrium seeking in non-cooperative games, the players considered in this paper cannot directly observe the…

最优化与控制 · 数学 2017-03-28 Maojiao Ye , Guoqiang Hu

The task of computing approximate Nash equilibria in large zero-sum extensive-form games has received a tremendous amount of attention due mainly to the Annual Computer Poker Competition. Immediately after its inception, two competing and…

人工智能 · 计算机科学 2014-11-19 Kevin Waugh , J. Andrew Bagnell