中文
相关论文

相关论文: Discretization Drift in Two-Player Games

200 篇论文

In this article, we introduce a method to approximate solutions of some variational mean field game problems with congestion, by finite sets of player trajectories. These trajectories are obtained by solving a minimization problem similar…

最优化与控制 · 数学 2022-01-14 Clément Sarrazin

Under mild regularity conditions, gradient-based methods converge globally to a critical point in the single-loss setting. This is known to break down for vanilla gradient descent when moving to multi-loss optimization, but can we hope to…

最优化与控制 · 数学 2021-01-19 Alistair Letcher

We formulate a new class of two-person zero-sum differential games, in a stochastic setting, where a specification on a target terminal state distribution is imposed on the players. We address such added specification by introducing…

系统与控制 · 电气工程与系统科学 2019-09-13 Yongxin Chen , Tryphon T. Georgiou , Michele Pavon

Learning processes in games explain how players grapple with one another in seeking an equilibrium. We study a natural model of learning based on individual gradients in two-player continuous games. In such games, the arguably natural…

计算机科学与博弈论 · 计算机科学 2020-11-10 Benjamin J. Chasnov , Daniel Calderone , Behçet Açıkmeşe , Samuel A. Burden , Lillian J. Ratliff

Physical systems can fail. For this reason the problem of identifying and reacting to faults has received a large attention in the control and computer science communities. In this paper we study the fault diagnosis problem for hybrid…

形式语言与自动机理论 · 计算机科学 2011-06-08 Davide Bresolin , Marta Capiluppi

In this paper we study two-player bilinear zero-sum games with constrained strategy spaces. An instance of natural occurrences of such constraints is when mixed strategies are used, which correspond to a probability simplex constraint. We…

计算机科学与博弈论 · 计算机科学 2022-06-10 Andre Wibisono , Molei Tao , Georgios Piliouras

Motivated by the pursuit of a systematic computational and algorithmic understanding of Generative Adversarial Networks (GANs), we present a simple yet unified non-asymptotic local convergence theory for smooth two-player games, which…

机器学习 · 统计学 2020-07-27 Tengyuan Liang , James Stokes

Dynamic difficulty adjustment ($DDA$) is a process of automatically changing a game difficulty for the optimization of user experience. It is a vital part of almost any modern game. Most existing DDA approaches concentrate on the experience…

机器学习 · 计算机科学 2021-06-08 Dvir Ben Or , Michael Kolomenkin , Gil Shabat

This article introduces differential hybrid games, which combine differential games with hybrid games. In both kinds of games, two players interact with continuous dynamics. The difference is that hybrid games also provide all the features…

计算机科学中的逻辑 · 计算机科学 2017-08-17 André Platzer

Logit dynamics are dynamical systems describing transitions and equilibria of actions of interacting players under uncertainty. An uncertainty is embodied in logit dynamic as a softmax type function often called a logit function originating…

最优化与控制 · 数学 2024-09-26 Hidekazu Yoshioka

In this work, we establish a frequency-domain framework for analyzing gradient-based algorithms in linear minimax optimization problems; specifically, our approach is based on the Z-transform, a powerful tool applied in Control Theory and…

最优化与控制 · 数学 2020-10-08 Ioannis Anagnostides , Paolo Penna

Min-max formulations have attracted great attention in the ML community due to the rise of deep generative models and adversarial methods, while understanding the dynamics of gradient algorithms for solving such formulations has remained a…

机器学习 · 计算机科学 2020-03-05 Guojun Zhang , Yaoliang Yu

Gradient regularization (GR) is a method that penalizes the gradient norm of the training loss during training. While some studies have reported that GR can improve generalization performance, little attention has been paid to it from the…

机器学习 · 计算机科学 2023-02-06 Ryo Karakida , Tomoumi Takase , Tomohiro Hayase , Kazuki Osawa

The remarkable success of the Adam in training neural networks has naturally led to the widespread use of its descent-ascent counterpart, Adam-DA, for solving zero-sum games. Despite its popularity in practice, a rigorous theoretical…

机器学习 · 计算机科学 2026-05-20 Yi Feng , Weiming Ou , Xiao Wang

Gradient regularization (GR), which aims to penalize the gradient norm atop the loss function, has shown promising results in training modern over-parameterized deep neural networks. However, can we trust this powerful technique? This paper…

机器学习 · 计算机科学 2024-06-17 Yang Zhao , Hao Zhang , Xiuyuan Hu

In decision-dependent games, multiple players optimize their decisions under a data distribution that shifts with their joint actions, creating complex dynamics in applications like market pricing. A practical consequence of these dynamics…

计算机科学与博弈论 · 计算机科学 2025-09-04 Guangzheng Zhong , Yang Liu , Jiming Liu

Unlike Poker where the action space $\mathcal{A}$ is discrete, differential games in the physical world often have continuous action spaces not amenable to discrete abstraction, rendering no-regret algorithms with…

计算机科学与博弈论 · 计算机科学 2025-02-17 Mukesh Ghimire , Zhe Xu , Yi Ren

The approximation of mixed Nash equilibria (MNE) for zero-sum games with mean-field interacting players has recently raised much interest in machine learning. In this paper we propose a mean-field gradient descent dynamics for finding the…

最优化与控制 · 数学 2025-05-13 Yulong Lu , Pierre Monmarché

We study the problem of finding the Nash equilibrium in a two-player zero-sum Markov game. Due to its formulation as a minimax optimization program, a natural approach to solve the problem is to perform gradient descent/ascent with respect…

最优化与控制 · 数学 2022-10-13 Sihan Zeng , Thinh T. Doan , Justin Romberg

This paper studies the convergence of mean field games with finite state space to mean field games with a continuous state space. We examine a space discretization of a diffusive dynamics, which is reminiscent of the Markov chain…

最优化与控制 · 数学 2024-01-18 Charles Bertucci , Alekos Cecchin