中文
相关论文

相关论文: Actively Learning to Coordinate in Convex Games vi…

200 篇论文

The model of congestion games is widely used to analyze games related to traffic and communication. A central property of these games is that they are potential games and hence posses a pure Nash equilibrium. In reality it is often the case…

计算机科学与博弈论 · 计算机科学 2011-11-17 Sergey Kuniavsky , Rann Smorodinsky

Correlated equilibria arise naturally when agents communicate or rely on intermediaries such as recommendation systems. We study when a given Nash equilibrium can be improved within the set of correlated equilibria for general objectives.…

理论经济学 · 经济学 2026-05-01 Kirill Rudov , Fedor Sandomirskiy , Leeat Yariv

We study a novel control problem in the context of network coordination games: the individuation of the smallest set of players capable of driving the system, globally, from one Nash equilibrium to another one. Our main contribution is the…

计算机科学与博弈论 · 计算机科学 2019-12-18 Stephane Durand , Giacomo Como , Fabio Fagnani

We study reinforcement learning for two-player zero-sum Markov games with simultaneous moves in the finite-horizon setting, where the transition kernel of the underlying Markov games can be parameterized by a linear function over the…

机器学习 · 计算机科学 2022-04-21 Zixiang Chen , Dongruo Zhou , Quanquan Gu

The existence of simple, uncoupled no-regret dynamics that converge to correlated equilibria in normal-form games is a celebrated result in the theory of multi-agent systems. Specifically, it has been known for more than 20 years that when…

计算机科学与博弈论 · 计算机科学 2022-09-05 Andrea Celli , Alberto Marchesi , Gabriele Farina , Nicola Gatti

We consider the problem of decentralized multi-agent reinforcement learning in Markov games. A fundamental question is whether there exist algorithms that, when adopted by all agents and run independently in a decentralized fashion, lead to…

机器学习 · 计算机科学 2023-03-23 Dylan J. Foster , Noah Golowich , Sham M. Kakade

Behavioral diversity, expert imitation, fairness, safety goals and others give rise to preferences in sequential decision making domains that do not decompose additively across time. We introduce the class of convex Markov games that allow…

计算机科学与博弈论 · 计算机科学 2025-06-17 Ian Gemp , Andreas Haupt , Luke Marris , Siqi Liu , Georgios Piliouras

We examine the problem of regret minimization when the learner is involved in a continuous game with other optimizing agents: in this case, if all players follow a no-regret algorithm, it is possible to achieve significantly lower regret…

计算机科学与博弈论 · 计算机科学 2023-03-20 Yu-Guan Hsieh , Kimon Antonakopoulos , Volkan Cevher , Panayotis Mertikopoulos

This paper examines the convergence of no-regret learning in games with continuous action sets. For concreteness, we focus on learning via "dual averaging", a widely used class of no-regret learning schemes where players take small steps…

最优化与控制 · 数学 2018-01-17 Panayotis Mertikopoulos , Zhengyuan Zhou

In online convex optimization, the player aims to minimize regret, or the difference between her loss and that of the best fixed decision in hindsight over the entire repeated game. Algorithms that minimize (standard) regret may converge to…

机器学习 · 计算机科学 2023-02-14 Zhou Lu , Elad Hazan

In this paper, we present a method for finding approximate Nash equilibria in a broad class of reachability games. These games are often used to formulate both collision avoidance and goal satisfaction. Our method is computationally…

系统与控制 · 电气工程与系统科学 2021-03-23 David Fridovich-Keil , Claire J. Tomlin

Computational aspects of solution notions such as Nash equilibrium have been extensively studied, including settings where the ultimate goal is to find an equilibrium that possesses some additional properties. Furthermore, in order to…

计算复杂性 · 计算机科学 2023-05-09 Bruce M. Kapron , Koosha Samieefar

Bargaining games, where agents attempt to agree on how to split utility, are an important class of games used to study economic behavior, which motivates a study of online learning algorithms in these games. In this work, we tackle when…

计算机科学与博弈论 · 计算机科学 2025-07-08 Serafina Kamp , Reese Liebman , Benjamin Fish

We design a distributed algorithm to seek generalized Nash equilibria of a robust game with uncertain coupled constraints. Due to the uncertainty of parameters in set constraints, we aim to find a generalized Nash equilibrium in the worst…

最优化与控制 · 数学 2022-04-05 Gehui Xu , Guanpu Chen , Hongsheng Qi

This paper considers no-regret learning for repeated continuous-kernel games with lossy bandit feedback. Since it is difficult to give the explicit model of the utility functions in dynamic environments, the players' action can only be…

机器学习 · 计算机科学 2022-05-17 Wenting Liu , Jinlong Lei , Peng Yi , Yiguang Hong

We consider the problem of learning stable matchings with unknown preferences in a decentralized and uncoordinated manner, where "decentralized" means that players make decisions individually without the influence of a central platform, and…

计算机科学与博弈论 · 计算机科学 2024-08-16 S. Rasoul Etesami , R. Srikant

This paper designs a distributed stochastic annealing algorithm for non-convex cooperative aggregative games, whose agents' cost functions not only depend on agents' own decision variables but also rely on the sum of agents' decision…

最优化与控制 · 数学 2022-04-05 Yinghui Wang , Xiaoxue Geng , Guanpu Chen , Wenxiao Zhao

In this paper, the problem of finding a Nash equilibrium of a multi-player game is considered. The players are only aware of their own cost functions as well as the action space of all players. We develop a relatively fast algorithm within…

系统与控制 · 计算机科学 2017-05-09 Farzad Salehisadaghiani , Lacra Pavel

We consider team zero-sum network congestion games with $n$ agents playing against $k$ interceptors over a graph $G$. The agents aim to minimize their collective cost of sending traffic over paths in $G$, which is an aggregation of edge…

计算机科学与博弈论 · 计算机科学 2024-05-14 Edan Orzech , Martin Rinard

This paper examines the convergence behaviour of simultaneous best-response dynamics in random potential games. We provide a theoretical result showing that, for two-player games with sufficiently many actions, the dynamics converge quickly…

计算机科学与博弈论 · 计算机科学 2025-05-19 Galit Ashkenazi-Golan , Domenico Mergoni Cecchelli , Edward Plumb