English
Related papers

Related papers: Direct Runge-Kutta Discretization Achieves Acceler…

200 papers

We study gradient-based optimization methods obtained by direct Runge-Kutta discretization of the ordinary differential equation (ODE) describing the movement of a heavy-ball under constant friction coefficient. When the function is high…

Optimization and Control · Mathematics 2019-05-30 Jingzhao Zhang , Suvrit Sra , Ali Jadbabaie

We study the connections between ordinary differential equations and optimization algorithms in a non-Euclidean setting. We propose a novel accelerated algorithm for minimising convex functions over a convex constrained set. This algorithm…

Optimization and Control · Mathematics 2026-03-30 Paul Dobson , Jesus María Sanz-Serna , Konstantinos C. Zygalakis

We derive a second-order ordinary differential equation (ODE) which is the limit of Nesterov's accelerated gradient method. This ODE exhibits approximate equivalence to Nesterov's scheme and thus can serve as a tool for analysis. We show…

Machine Learning · Statistics 2015-10-29 Weijie Su , Stephen Boyd , Emmanuel J. Candes

The aim of this paper is to construct and analyze explicit exponential Runge-Kutta methods for the temporal discretization of linear and semilinear integro-differential equations. By expanding the errors of the numerical method in terms of…

Numerical Analysis · Mathematics 2023-01-24 Alexander Ostermann , Fardin Saedpanah , Nasrin Vaisi

We study first-order optimization methods obtained by discretizing ordinary differential equations (ODEs) corresponding to Nesterov's accelerated gradient methods (NAGs) and Polyak's heavy-ball method. We consider three discretization…

Optimization and Control · Mathematics 2019-11-05 Bin Shi , Simon S. Du , Weijie J. Su , Michael I. Jordan

This paper introduces the Runge-Kutta Chebyshev descent method (RKCD) for strongly convex optimisation problems. This new algorithm is based on explicit stabilised integrators for stiff differential equations, a powerful class of numerical…

Optimization and Control · Mathematics 2020-06-30 Armin Eftekhari , Bart Vandereycken , Gilles Vilmart , Konstantinos C. Zygalakis

We present a coupled system of ODEs which, when discretized with a constant time step/learning rate, recovers Nesterov's accelerated gradient descent algorithm. The same ODEs, when discretized with a decreasing learning rate, leads to novel…

Optimization and Control · Mathematics 2020-09-02 Maxime Laborde , Adam M. Oberman

The derivation of second-order ordinary differential equations (ODEs) as continuous-time limits of optimization algorithms has been shown to be an effective tool for the analysis of these algorithms. Additionally, discretizing…

Optimization and Control · Mathematics 2019-08-29 Rachel Walker , Emily Zhang

In a previous paper, a technique was suggested to avoid order reduction with any explicit exponential Runge-Kutta method when integrating initial boundary value nonlinear problems with time-dependent boundary conditions. In this paper, we…

Numerical Analysis · Mathematics 2023-07-18 Begoña Cano , María Jesús Moreta

Gradient-based minimax optimal algorithms have greatly promoted the development of continuous optimization and machine learning. One seminal work due to Yurii Nesterov [Nes83a] established $\tilde{\mathcal{O}}(\sqrt{L/\mu})$ gradient…

Machine Learning · Computer Science 2023-12-07 Yuanshi Liu , Hanzhen Zhao , Yang Xu , Pengyun Yue , Cong Fang

A new approach for the construction of high order A-stable explicit integrators for ordinary differential equations (ODEs) is theoretically studied. Basically, the integrators are obtained by splitting, at each time step, the solution of…

Numerical Analysis · Mathematics 2012-08-24 H. de la Cruz , R. J. Biscay , J. C. Jimenez , F. Carbonell

We develop a theoretical foundation for the application of Nesterov's accelerated gradient descent method (AGD) to the approximation of solutions of a wide class of partial differential equations (PDEs). This is achieved by proving the…

Numerical Analysis · Mathematics 2021-02-03 Jea-Hyun Park , Abner J. Salgado , Steven M. Wise

We introduce a family of stochastic optimization methods based on the Runge-Kutta-Chebyshev (RKC) schemes. The RKC methods are explicit methods originally designed for solving stiff ordinary differential equations by ensuring that their…

Optimization and Control · Mathematics 2022-02-01 Tony Stillfjord , Måns Williamson

When applied to stiff, linear differential equations with time-dependent forcing, Runge-Kutta methods can exhibit convergence rates lower than predicted by the classical order condition theory. Commonly, this order reduction phenomenon is…

Numerical Analysis · Mathematics 2022-02-15 Steven Roberts , Adrian Sandu

Classical convergence theory of Runge-Kutta methods assumes that the time step is small relative to the Lipschitz constant of the ordinary differential equation (ODE). For stiff problems, that assumption is often violated, and a problematic…

Numerical Analysis · Mathematics 2026-05-05 Steven B. Roberts , David Shirokoff , Abhijit Biswas , Benjamin Seibold

In this paper, we study the behavior of solutions of the ODE associated to Nesterov acceleration. It is well-known since the pioneering work of Nesterov that the rate of convergence $O(1/t^2)$ is optimal for the class of convex functions…

Optimization and Control · Mathematics 2019-07-09 Jean François Aujol , Charles Dossal , Aude Rondepierre

In this paper, we summarize the results about the strong convergence rate of the Ninomiya-Victoir scheme and the stable convergence in law of its normalized error that we obtained in previous papers. We then recall the properties of the…

Probability · Mathematics 2016-12-22 Anis Al Gerbi , Benjamin Jourdain , Emmanuelle Clément

Modern deep learning algorithms use variations of gradient descent as their main learning methods. Gradient descent can be understood as the simplest Ordinary Differential Equation (ODE) solver; namely, the Euler method applied to the…

Machine Learning · Computer Science 2025-05-20 Benoit Dherin , Michael Munn , Hanna Mazzawi , Michael Wunder , Sourabh Medapati , Javier Gonzalvo

Discrete gradient methods are geometric integration techniques that can preserve the dissipative structure of gradient flows. Due to the monotonic decay of the function values, they are well suited for general convex and nonconvex…

Optimization and Control · Mathematics 2024-07-17 Matthias J. Ehrhardt , Erlend S. Riis , Torbjørn Ringholm , Carola-Bibiane Schönlieb

Recently Grimmer [1] showed for smooth convex optimization by utilizing longer steps periodically, gradient descent's textbook $LD^2/2T$ convergence guarantees can be improved by constant factors, conjecturing an accelerated rate strictly…

Optimization and Control · Mathematics 2023-09-28 Benjamin Grimmer , Kevin Shu , Alex L. Wang
‹ Prev 1 2 3 10 Next ›