中文
相关论文

相关论文: Game on Random Environment, Mean-field Langevin Sy…

200 篇论文

We investigate a class of reinforcement learning dynamics where players adjust their strategies based on their actions' cumulative payoffs over time - specifically, by playing mixed strategies that maximize their expected cumulative payoff…

最优化与控制 · 数学 2016-02-10 Panayotis Mertikopoulos , William H. Sandholm

We present a framework that incorporates the idea of bounded rationality into dynamic stochastic pursuit-evasion games. The solution of a stochastic game is characterized, in general, by its (Nash) equilibria in feedback form. However,…

系统与控制 · 电气工程与系统科学 2020-03-17 Yue Guan , Dipankar Maity , Christopher M. Kroninger , Panagiotis Tsiotras

Learning processes in games explain how players grapple with one another in seeking an equilibrium. We study a natural model of learning based on individual gradients in two-player continuous games. In such games, the arguably natural…

计算机科学与博弈论 · 计算机科学 2020-11-10 Benjamin J. Chasnov , Daniel Calderone , Behçet Açıkmeşe , Samuel A. Burden , Lillian J. Ratliff

We propose a new approach to mean field games with major and minor players. Our formulation involves a two player game where the optimization of the representative minor player is standard while the major player faces an optimization over…

概率论 · 数学 2014-09-26 Rene Carmona , Xiuneng Zhu

We formulate a class of mean field games on a finite state space with variational principles resembling those in continuous-state mean field games. We construct a controlled continuity equation featuring a nonlinear activation function on…

最优化与控制 · 数学 2023-10-10 Yuan Gao , Wuchen Li , Jian-Guo Liu

We develop a probabilistic approach to continuous-time finite state mean field games. Based on an alternative description of continuous-time Markov chain by means of semimartingale and the weak formulation of stochastic optimal control, our…

概率论 · 数学 2018-08-24 Rene Carmona , Peiqi Wang

In this paper, we introduce discrete-time linear mean-field games subject to an infinite-horizon discounted-cost optimality criterion. The state space of a generic agent is a compact Borel space. At every time, each agent is randomly…

系统与控制 · 电气工程与系统科学 2023-01-18 Naci Saldi

We introduce the notion of regularized Bayesian best response (RBBR) learning dynamic in heterogeneous population games. We obtain such a dynamic via perturbation by an arbitrary lower semicontinuous, strongly convex regularizer in Bayesian…

最优化与控制 · 数学 2023-03-13 Sayan Mukherjee , Souvik Roy

Zero-sum Markov Stackelberg games can be used to model myriad problems, in domains ranging from economics to human robot interaction. In this paper, we develop policy gradient methods that solve these games in continuous state and action…

计算机科学与博弈论 · 计算机科学 2024-01-24 Denizalp Goktas , Arjun Prakash , Amy Greenwald

The cornerstone underpinning deep learning is the guarantee that gradient descent on an objective converges to local minima. Unfortunately, this guarantee fails in settings, such as generative adversarial nets, where there are multiple…

机器学习 · 计算机科学 2018-06-07 David Balduzzi , Sebastien Racaniere , James Martens , Jakob Foerster , Karl Tuyls , Thore Graepel

Research in adversarial learning follows a cat and mouse game between attackers and defenders where attacks are proposed, they are mitigated by new defenses, and subsequently new attacks are proposed that break earlier defenses, and so on.…

机器学习 · 计算机科学 2020-11-13 Ambar Pal , René Vidal

We consider an N-player hierarchical game in which the i-th player's objective comprises of an expectation-valued term, parametrized by rival decisions, and a hierarchical term. Such a framework allows for capturing a broad range of…

最优化与控制 · 数学 2024-01-26 Shisheng Cui , Uday V. Shanbhag , Mathias Staudigl

We consider network aggregative games to model and study multi-agent populations in which each rational agent is influenced by the aggregate behavior of its neighbors, as specified by an underlying network. Specifically, we examine systems…

系统与控制 · 计算机科学 2015-06-26 Francesca Parise , Sergio Grammatico , Basilio Gentile , John Lygeros

Financial markets and more generally macro-economic models involve a large number of individuals interacting through variables such as prices resulting from the aggregate behavior of all the agents. Mean field games have been introduced to…

最优化与控制 · 数学 2021-07-12 René Carmona , Mathieu Laurière

Entropy games and matrix multiplication games have been recently introduced by Asarin et al. They model the situation in which one player (Despot) wishes to minimize the growth rate of a matrix product, whereas the other player (Tribune)…

计算机科学与博弈论 · 计算机科学 2019-12-30 Marianne Akian , Stéphane Gaubert , Julien Grand-Clément , Jérémie Guillaud

We develop a flexible stochastic approximation framework for analyzing the long-run behavior of learning in games (both continuous and finite). The proposed analysis template incorporates a wide array of popular learning algorithms,…

计算机科学与博弈论 · 计算机科学 2023-07-04 Panayotis Mertikopoulos , Ya-Ping Hsieh , Volkan Cevher

Starting from a heuristic learning scheme for N-person games, we derive a new class of continuous-time learning dynamics consisting of a replicator-like drift adjusted by a penalty term that renders the boundary of the game's strategy space…

最优化与控制 · 数学 2014-04-08 Pierre Coucheney , Bruno Gaujal , Panayotis Mertikopoulos

We investigate mean-field games (MFG) in which agents can actively control their speed of access to information. Specifically, the agents can dynamically decide to obtain observations with reduced delay by accepting higher observation…

最优化与控制 · 数学 2025-06-03 Dirk Becherer , Christoph Reisinger , Jonathan Tam

We provide an abstract framework for submodular mean field games and identify verifiable sufficient conditions that allow to prove existence and approximation of strong mean field equilibria in models where data may not be continuous with…

最优化与控制 · 数学 2022-01-21 Jodi Dianetti , Giorgio Ferrari , Markus Fischer , Max Nendel

We investigate the resolution of second-order, potential, and monotone mean field games with the generalized conditional gradient algorithm, an extension of the Frank-Wolfe algorithm. We show that the method is equivalent to the fictitious…

最优化与控制 · 数学 2023-08-22 Pierre Lavigne , Laurent Pfeiffer