中文
相关论文

相关论文: Bessel Function Analysis of Nesterov's ODE in $N$-…

200 篇论文

Nesterov's Accelerated Gradient (NAG) for optimization has better performance than its continuous time limit (noiseless kinetic Langevin) when a finite step-size is employed \citep{shi2021understanding}. This work explores the sampling…

机器学习 · 计算机科学 2022-06-22 Ruilin Li , Hongyuan Zha , Molei Tao

In this work, we study the sample complexity of obtaining a Nash equilibrium (NE) estimate in two-player zero-sum matrix games with noisy feedback. Specifically, we propose a novel algorithm that repeatedly solves linear programs (LPs) to…

最优化与控制 · 数学 2026-02-16 Jiashuo Jiang , Mengxiao Zhang

Accelerated gradient methods like Nesterov's Accelerated Gradient (NAG) achieve faster convergence on well-conditioned problems but often diverge on ill-conditioned or non-convex landscapes due to aggressive momentum accumulation. We…

机器学习 · 计算机科学 2025-12-12 Sarwan Ali

Nesterov's accelerated gradient (AG) method for minimizing a smooth strongly convex function $f$ is known to reduce $f({\bf x}_k)-f({\bf x}^*)$ by a factor of $\epsilon\in(0,1)$ after $k=O(\sqrt{L/\ell}\log(1/\epsilon))$ iterations, where…

最优化与控制 · 数学 2019-01-11 Sahar Karimi , Stephen Vavasis

We develop a theoretical foundation for the application of Nesterov's accelerated gradient descent method (AGD) to the approximation of solutions of a wide class of partial differential equations (PDEs). This is achieved by proving the…

数值分析 · 数学 2021-02-03 Jea-Hyun Park , Abner J. Salgado , Steven M. Wise

We study strong stability of Nash equilibria in load balancing games of m (m >= 2) identical servers, in which every job chooses one of the m servers and each job wishes to minimize its cost, given by the workload of the server it chooses.…

计算机科学与博弈论 · 计算机科学 2015-06-17 Bo Chen , Song-Song Li , Yu-Zhong Zhang

A wide array of modern machine learning applications - from adversarial models to multi-agent reinforcement learning - can be formulated as non-cooperative games whose Nash equilibria represent the system's desired operational states.…

计算机科学与博弈论 · 计算机科学 2023-12-29 Iosif Sakos , Emmanouil-Vasileios Vlatakis-Gkaragkounis , Panayotis Mertikopoulos , Georgios Piliouras

In order to find Nash-equilibria for two-player zero-sum games where each player plays combinatorial objects like spanning trees, matchings etc, we consider two online learning algorithms: the online mirror descent (OMD) algorithm and the…

机器学习 · 计算机科学 2016-03-03 Swati Gupta , Michel Goemans , Patrick Jaillet

We show that, for any sufficiently small fixed $\epsilon > 0$, when both players in a general-sum two-player (bimatrix) game employ optimistic mirror descent (OMD) with smooth regularization, learning rate $\eta = O(\epsilon^2)$ and $T =…

计算机科学与博弈论 · 计算机科学 2022-10-10 Ioannis Anagnostides , Gabriele Farina , Ioannis Panageas , Tuomas Sandholm

We derive a second-order ordinary differential equation (ODE) which is the limit of Nesterov's accelerated gradient method. This ODE exhibits approximate equivalence to Nesterov's scheme and thus can serve as a tool for analysis. We show…

机器学习 · 统计学 2015-10-29 Weijie Su , Stephen Boyd , Emmanuel J. Candes

The study of learning in games typically assumes that each player always has access to all of their actions. However, in many practical scenarios, players' available actions might be restricted due to exogenous stochasticity. To model this…

计算机科学与博弈论 · 计算机科学 2026-05-12 Thomas Schwarz , Ryann Sim , Chun Kai Ling

We study infinite-horizon discounted two-player zero-sum Markov games, and develop a decentralized algorithm that provably converges to the set of Nash equilibria under self-play. Our algorithm is based on running an Optimistic Gradient…

机器学习 · 计算机科学 2021-07-08 Chen-Yu Wei , Chung-Wei Lee , Mengxiao Zhang , Haipeng Luo

We show by counterexample that policy-gradient algorithms have no guarantees of even local convergence to Nash equilibria in continuous action and state space multi-agent settings. To do so, we analyze gradient-play in N-player general-sum…

机器学习 · 计算机科学 2019-12-18 Eric Mazumdar , Lillian J. Ratliff , Michael I. Jordan , S. Shankar Sastry

In practical applications, decision-makers with heterogeneous dynamics may be engaged in the same decision-making process. This motivates us to study distributed Nash equilibrium seeking for games in which players are mixed-order (first-…

最优化与控制 · 数学 2022-09-05 Maojiao Ye , Lei Ding , Jizhao Yin

Considering a class of gradient-based multi-agent learning algorithms in non-cooperative settings, we provide local convergence guarantees to a neighborhood of a stable local Nash equilibrium. In particular, we consider continuous games…

最优化与控制 · 数学 2024-09-23 Benjamin Chasnov , Lillian J. Ratliff , Eric Mazumdar , Samuel A. Burden

In evolutionary game theory, it is customary to be partial to the dynamical models possessing fixed points so that they may be understood as the attainment of evolutionary stability, and hence, Nash equilibrium. Any show of periodic or…

种群与进化 · 定量生物学 2021-02-23 Archan Mukhopadhyay , Sagar Chakraborty

We prove that every finite two-person shortest path game, where the local cost of every move is positive for each player, has a Nash equilibrium (NE) in pure stationary strategies, which can be computed in polynomial time. We also extend…

离散数学 · 计算机科学 2025-05-22 Endre Boros , Khaled Elbassioni , Vladimir Gurvich , Mikhail Vyalyi

We show that the value function of an optimal stopping game driven by a one-dimensional diffusion can be characterised using a modification of the Legendre transformation if and only if the optimal stopping game exhibits a Nash equilibrium…

最优化与控制 · 数学 2014-01-10 Jenny Sexton

We analyse the computational complexity of finding Nash equilibria in turn-based stochastic multiplayer games with omega-regular objectives. We show that restricting the search space to equilibria whose payoffs fall into a certain interval…

计算机科学与博弈论 · 计算机科学 2015-07-01 Michael Ummels , Dominik Wojtczak

In this paper, the problem of distributively seeking the equilibria of aggregative games with bilevel structures is studied. Different from the traditional aggregative games, here the aggregation is determined by the minimizer of a virtual…

系统与控制 · 电气工程与系统科学 2025-12-02 Kaihong Lu , Huanshui Zhang , Long Wang
‹ 上一页 1 8 9 10 下一页 ›