English
Related papers

Related papers: Continuous-Time Analysis of Accelerated Gradient M…

200 papers

Classical machine learning models such as deep neural networks are usually trained by using Stochastic Gradient Descent-based (SGD) algorithms. The classical SGD can be interpreted as a discretization of the stochastic gradient flow. In…

Optimization and Control · Mathematics 2023-10-03 Valentin Leplat , Daniil Merkulov , Aleksandr Katrutsa , Daniel Bershatsky , Olga Tsymboi , Ivan Oseledets

We develop a theory of accelerated first-order optimization from the viewpoint of differential equations and Lyapunov functions. Building upon the previous work of many researchers, we consider differential equations which model the…

Optimization and Control · Mathematics 2021-04-02 Jonathan W. Siegel

Understanding the geometric properties of gradient descent dynamics is a key ingredient in deciphering the recent success of very large machine learning models. A striking observation is that trained over-parameterized models retain some…

Machine Learning · Computer Science 2024-07-11 Sibylle Marcotte , Rémi Gribonval , Gabriel Peyré

Model-free reinforcement learning attempts to find an optimal control action for an unknown dynamical system by directly searching over the parameter space of controllers. The convergence behavior and statistical properties of these…

Optimization and Control · Mathematics 2021-03-17 Hesameddin Mohammadi , Armin Zare , Mahdi Soltanolkotabi , Mihailo R. Jovanović

Gradient algorithms are classical in adaptive control and parameter estimation. For instantaneous quadratic cost functions they lead to a linear time-varying dynamic system that converges exponentially under persistence of excitation…

Optimization and Control · Mathematics 2020-10-06 Juan G. Rueda-Escobedo , Jaime A. Moreno

We consider the parallel-in-time solution of scalar nonlinear conservation laws in one spatial dimension. The equations are discretized in space with a conservative finite-volume method using weighted essentially non-oscillatory (WENO)…

Numerical Analysis · Mathematics 2025-11-04 O. A. Krzysik , H. De Sterck , R. D. Falgout , J. B. Schroder

There are much recent interests in solving noncovnex min-max optimization problems due to its broad applications in many areas including machine learning, networked resource allocations, and distributed optimization. Perhaps, the most…

Optimization and Control · Mathematics 2021-12-20 Thinh T. Doan

We propose a family of optimization methods that achieve linear convergence using first-order gradient information and constant step sizes on a class of convex functions much larger than the smooth and strongly convex ones. This larger…

Optimization and Control · Mathematics 2018-09-14 Chris J. Maddison , Daniel Paulin , Yee Whye Teh , Brendan O'Donoghue , Arnaud Doucet

In this paper, we consider gradient methods for minimizing smooth convex functions, which employ the information obtained at the previous iterations in order to accelerate the convergence towards the optimal solution. This information is…

Optimization and Control · Mathematics 2021-06-02 Yurii Nesterov , Mihai I. Florea

This paper is part of a program to combine a staggered time and staggered spatial discretization of continuum wave equations so that important properties of the continuum that are proved using vector calculus can be proven in an analogous…

Numerical Analysis · Mathematics 2020-10-13 Stanly Steinberg

Continuum computational kinetic plasma models evolve the distribution function of a plasma species $f_s$ on a phase-space grid over time. In many problems of interest the distribution function has limited extent in velocity space; hence,…

Plasma Physics · Physics 2026-01-14 Manaure Francisquez , Petr Cagas , Akash Shukla , James Juno , Gregory W. Hammett

We propose a systematic method for learning stable and physically interpretable dynamical models using sampled trajectory data from physical processes based on a generalized Onsager principle. The learned dynamics are autonomous ordinary…

Dynamical Systems · Mathematics 2021-11-25 Haijun Yu , Xinyuan Tian , Weinan E , Qianxiao Li

This paper studies the online convex optimization problem by using an Online Continuous-Time Nesterov Accelerated Gradient method (OCT-NAG). We show that the continuous-time dynamics generated by the online version of the Bregman Lagrangian…

Optimization and Control · Mathematics 2020-09-29 Chao Sun , Guoqiang Hu

Extremal graphical models encode the conditional independence structure of multivariate extremes and provide a powerful tool for quantifying the risk of rare events. Prior work on learning these graphs from data has focused on the setting…

Methodology · Statistics 2025-04-15 Sebastian Engelke , Armeen Taeb

Recent research has indicated a substantial rise in interest in understanding Nesterov's accelerated gradient methods via their continuous-time models. However, most existing studies focus on specific classes of Nesterov's methods, which…

Optimization and Control · Mathematics 2026-03-23 Chanwoong Park , Youngchae Cho , Insoon Yang

In convex optimization, first-order optimization methods efficiently minimizing function values have been a central subject study since Nesterov's seminal work of 1983. Recently, however, Kim and Fessler's OGM-G and Lee et al.'s FISTA-G…

Optimization and Control · Mathematics 2023-11-02 Jaeyeon Kim , Asuman Ozdaglar , Chanwoo Park , Ernest K. Ryu

This paper generalizes the optimized gradient method (OGM) that achieves the optimal worst-case cost function bound of first-order methods for smooth convex minimization. Specifically, this paper studies a generalized formulation of OGM and…

Optimization and Control · Mathematics 2019-06-14 Donghwan Kim , Jeffrey A. Fessler

We show a novel systematic way to construct conservative finite difference schemes for quasilinear first-order system of ordinary differential equations with conserved quantities. In particular, this includes both autonomous and…

Numerical Analysis · Mathematics 2018-05-23 Andy T. S. Wan , Alexander Bihlo , Jean-Christophe Nave

Sampling from a target distribution is a fundamental problem. Traditional Markov chain Monte Carlo (MCMC) algorithms, such as the unadjusted Langevin algorithm (ULA), derived from the overdamped Langevin dynamics, have been extensively…

Optimization and Control · Mathematics 2024-10-29 Xinzhe Zuo , Stanley Osher , Wuchen Li

Following the seminal work of Nesterov, accelerated optimization methods have been used to powerfully boost the performance of first-order, gradient-based parameter estimation in scenarios where second-order optimization strategies are…

Numerical Analysis · Computer Science 2017-11-28 Anthony Yezzi , Ganesh Sundaramoorthi