中文
相关论文

相关论文: Discretization Drift in Two-Player Games

200 篇论文

We study convergence rates of the generalized conditional gradient (GCG) method applied to fully discretized Mean Field Games (MFG) systems. While explicit convergence rates of the GCG method have been established at the continuous PDE…

数值分析 · 数学 2026-02-13 Haruka Nakamura , Norikazu Saito

This article focuses two issues related to the first-order discounted mean field games system. The first is the time discretization problem. The time discretization approach enables us to prove the existence of solutions (u,m) of the…

偏微分方程分析 · 数学 2025-09-19 Renato Iturriaga , Cristian Mendico , Kaizhi Wang , Yuchen Xu

The two-timescale gradient descent-ascent (GDA) is a canonical gradient algorithm designed to find Nash equilibria in min-max games. We analyze the two-timescale GDA by investigating the effects of learning rate ratios on convergence…

最优化与控制 · 数学 2025-10-13 Jing An , Jianfeng Lu

Discretization of phase space usually nullifies chaos in dynamical systems. We show that if randomness is associated with discretization dynamical chaos may survive and be indistinguishable from that of the original chaotic system, when an…

混沌动力学 · 物理学 2009-11-10 M. Falcioni , A. Vulpiani , G. Mantica , S. Pigolotti

We analyze inertial coordination games: dynamic coordination games with an endogenously changing state that depends on (i) a persistent fundamental players privately learn about over time; and (ii) past play. The speed of learning…

理论经济学 · 经济学 2025-08-14 Andrew Koh , Ricky Li , Kei Uzui

The perceived risk and reward for a given situation can vary depending on resource availability, accumulated wealth, and other extrinsic factors such as individual backgrounds. Based on this general aspect of everyday life, here we use…

统计力学 · 物理学 2021-12-07 Marco A. Amaral , Marcelo M. de Oliveira

Generative Adversarial Networks (GANs) have shown remarkable performance in image generation. However, GAN training suffers from the problem of instability. One of the main approaches to address this problem is to modify the loss function,…

机器学习 · 计算机科学 2024-03-19 Iu Yahiro , Takashi Ishida , Naoto Yokoya

Deep learning systems are known to exhibit implicit regularization (alt. implicit bias), favoring simple solutions instead of merely minimizing the loss function. In some cases, we can analytically derive the implicit regularization --…

机器学习 · 统计学 2026-05-08 Joseph H. Rudoler , Kevin Tan , Giles Hooker , Konrad P. Kording

We study two-player security games which can be viewed as sequences of nonzero-sum matrix games played by an Attacker and a Defender. The evolution of the game is based on a stochastic fictitious play process. Players do not have access to…

计算机科学与博弈论 · 计算机科学 2010-03-16 Kien C. Nguyen , Tansu Alpcan , Tamer Basar

Dual gradient descent combined with early stopping represents an efficient alternative to the Tikhonov variational approach when the regularizer is strongly convex. However, for many relevant applications, it is crucial to deal with…

最优化与控制 · 数学 2023-05-12 Vassilis Apidopoulos , Cesare Molinari , Lorenzo Rosasco , Silvia Villa

In this paper we consider the problem of distributed Nash equilibrium (NE) seeking over networks, a setting in which players have limited local information. We start from a continuous-time gradient-play dynamics that converges to an NE…

最优化与控制 · 数学 2024-10-30 Dian Gadjov , Lacra Pavel

Consider a two-player game repeated N times. Player 1 can choose between two styles (for interpretability, offensive and defensive), whereas Player 2 uses a single fixed style. Let X N\,:= \#wins -\#losses for Player 1 after N games, and…

计算机科学与博弈论 · 计算机科学 2026-04-20 Jonatha ANSELMI , Bruno Gaujal

The rapid progress in machine learning in recent years has been based on a highly productive connection to gradient-based optimization. Further progress hinges in part on a shift in focus from pattern recognition to decision-making and…

机器学习 · 计算机科学 2024-02-27 Neha S. Wadia , Yatin Dandi , Michael I. Jordan

The policy iteration method is a classical algorithm for solving optimal control problems. In this paper, we introduce a policy iteration method for Mean Field Games systems, and we study the convergence of this procedure to a solution of…

偏微分方程分析 · 数学 2021-07-12 Simone Cacace , Fabio Camilli , Alessandro Goffi

Generating competitive strategies and performing continuous motion planning simultaneously in an adversarial setting is a challenging problem. In addition, understanding the intent of other agents is crucial to deploying autonomous systems…

机器人学 · 计算机科学 2023-10-12 Hongrui Zheng , Zhijun Zhuang , Johannes Betz , Rahul Mangharam

This work explores three-player game training dynamics, under what conditions three-player games converge and the equilibria the converge on. In contrast to prior work, we examine a three-player game architecture in which all players…

机器学习 · 计算机科学 2022-08-16 Kenneth Christofferson , Fernando J. Yanez

Adversarial training often suffers from a robustness-accuracy trade-off, where achieving high robustness comes at the cost of accuracy. One approach to mitigate this trade-off is leveraging invariance regularization, which encourages model…

机器学习 · 计算机科学 2025-08-29 Futa Waseda , Ching-Chun Chang , Isao Echizen

We present a fast numerical algorithm for large scale zero-sum stochastic games with perfect information, which combines policy iteration and algebraic multigrid methods. This algorithm can be applied either to a true finite state space…

最优化与控制 · 数学 2015-03-19 Marianne Akian , Sylvie Detournay

We consider deterministic mean field games where the dynamics of a typical agent is non-linear with respect to the state variable and affine with respect to the control variable. Particular instances of the problem considered here are mean…

最优化与控制 · 数学 2022-12-21 Justina Gianatti , Francisco J. Silva

The dominant line of work in domain adaptation has focused on learning invariant representations using domain-adversarial training. In this paper, we interpret this approach from a game theoretical perspective. Defining optimal solutions in…

机器学习 · 计算机科学 2022-02-14 David Acuna , Marc T Law , Guojun Zhang , Sanja Fidler