中文
相关论文

相关论文: Recursive Regret Matching: A General Method for So…

200 篇论文

Saddle point with a given Morse index on a potential energy surface is an important object related to energy landscape in physics and chemistry. Efficient numerical methods based on iterative minimization formulation have been proposed in…

最优化与控制 · 数学 2022-12-19 Shuting Gu , Hao Zhang , Xiang Zhou

No-regret learners seek to minimize the difference between the loss they cumulated through the actions they played, and the loss they would have cumulated in hindsight had they consistently modified their behavior according to some strategy…

计算机科学与博弈论 · 计算机科学 2023-11-09 Gabriele Farina , Charilaos Pipis

Game contingent claims (GCCs) generalize American contingent claims by allowing the writer to recall the option as long as it is not exercised, at the price of paying some penalty. In incomplete markets, an appealing approach is to analyze…

概率论 · 数学 2018-11-27 Klebert Kentia , Christoph Kühn

Wide machine learning tasks can be formulated as non-convex multi-player games, where Nash equilibrium (NE) is an acceptable solution to all players, since no one can benefit from changing its strategy unilaterally. Attributed to the…

计算机科学与博弈论 · 计算机科学 2023-01-20 Guanpu Chen , Gehui Xu , Fengxiang He , Yiguang Hong , Leszek Rutkowski , Dacheng Tao

We study the problem of no-regret learning algorithms for general monotone and smooth games and their last-iterate convergence properties. Specifically, we investigate the problem under bandit feedback and strongly uncoupled dynamics, which…

计算机科学与博弈论 · 计算机科学 2024-08-19 Jing Dong , Baoxiang Wang , Yaoliang Yu

Bargaining games, where agents attempt to agree on how to split utility, are an important class of games used to study economic behavior, which motivates a study of online learning algorithms in these games. In this work, we tackle when…

计算机科学与博弈论 · 计算机科学 2025-07-08 Serafina Kamp , Reese Liebman , Benjamin Fish

In this paper we investigate the Follow the Regularized Leader dynamics in sequential imperfect information games (IIG). We generalize existing results of Poincar\'e recurrence from normal-form games to zero-sum two-player imperfect…

In this paper, we present an online learning approach for two-player zero-sum linear quadratic games with unknown dynamics. We develop a framework combining regularized least squares model estimation, high probability confidence sets, and…

系统与控制 · 电气工程与系统科学 2026-04-06 Shanting Wang , Weihao Sun , Andreas A. Malikopoulos

This paper examines the convergence of no-regret learning in games with continuous action sets. For concreteness, we focus on learning via "dual averaging", a widely used class of no-regret learning schemes where players take small steps…

最优化与控制 · 数学 2018-01-17 Panayotis Mertikopoulos , Zhengyuan Zhou

Contemporary applications of machine learning in two-team e-sports and the superior expressivity of multi-agent generative adversarial networks raise important and overlooked theoretical questions regarding optimization in two-team games.…

计算机科学与博弈论 · 计算机科学 2023-04-18 Fivos Kalogiannis , Ioannis Panageas , Emmanouil-Vasileios Vlatakis-Gkaragkounis

We consider the problem of decentralized multi-agent reinforcement learning in Markov games. A fundamental question is whether there exist algorithms that, when adopted by all agents and run independently in a decentralized fashion, lead to…

机器学习 · 计算机科学 2023-03-23 Dylan J. Foster , Noah Golowich , Sham M. Kakade

Iterated regret minimization has been introduced recently by J.Y. Halpern and R. Pass in classical strategic games. For many games of interest, this new solution concept provides solutions that are judged more reasonable than solutions…

计算机科学与博弈论 · 计算机科学 2015-05-18 Emmanuel Filiot , Tristan Le Gall , Jean-François Raskin

A Nash equilibrium has become important solution concept for analyzing the decision making in Game theory. In this paper, we consider the problem of computing Nash equilibria of a subclass of generic finite normal form games. We define…

计算机科学与博弈论 · 计算机科学 2015-03-17 Samaresh Chatterji , Ratnik Gandhi

Nash Equilibrium (NE) is the canonical solution concept of game theory, which provides an elegant tool to understand the rationalities. Though mixed strategy NE exists in any game with finite players and actions, computing NE in two- or…

计算机科学与博弈论 · 计算机科学 2024-05-07 Xinrun Wang , Chang Yang , Shuxin Li , Pengdeng Li , Xiao Huang , Hau Chan , Bo An

This paper investigates the sublinear regret guarantees of two non-no-regret algorithms in zero-sum games: Fictitious Play, and Online Gradient Descent with constant stepsizes. In general adversarial online learning settings, both…

机器学习 · 计算机科学 2025-06-17 John Lazarsfeld , Georgios Piliouras , Ryann Sim , Andre Wibisono

In this article we consider a special class of Nash equilibrium problems that cannot be reduced to a single player control problem. Problems of this type can be solved by a semi-smooth Newton method. Applying results from the established…

最优化与控制 · 数学 2018-11-09 Veronika Karl , Frank Pörner

Many resources are provided by an ecological system that is vulnerable to tipping when exceeding a certain level of pollution, with a sudden big loss of ecosystem services. An ecological system is usually also a common-pool resource and…

综合经济学 · 经济学 2026-03-23 Yongyang Cai , Anastasios Xepapadeas , Aart de Zeeuw

In this paper we study the N-player nonzero-sum Dynkin game ($N\geq 3$) in continuous time, which is a non-cooperative game where the strategies are stopping times. We show that the game has a Nash equilibrium point for general payoff…

计算机科学与博弈论 · 计算机科学 2011-10-27 Hamadene Said , Hassani Mohammed

$ $This paper addresses the inverse problem for Linear-Quadratic (LQ) nonzero-sum $N$-player differential games, where the goal is to learn parameters of an unknown cost function for the game, called observed, given the demonstrated…

最优化与控制 · 数学 2024-10-28 Emin Martirosyan , Ming Cao

This work proposes a novel distributed approach for computing a Nash equilibrium in convex games with merely monotone and restricted strongly monotone pseudo-gradients. By leveraging the idea of the centralized operator extrapolation method…

最优化与控制 · 数学 2025-07-18 Tatiana Tatarenko , Angelia Nedich