English
Related papers

Related papers: The (1+1)-ES Reliably Overcomes Saddle Points

200 papers

Deep learning researchers commonly suggest that converged models are stuck in local minima. More recently, some researchers observed that under reasonable assumptions, the vast majority of critical points are saddle points, not true minima.…

Machine Learning · Computer Science 2016-02-25 Zachary C. Lipton

We consider the problem of finding a saddle point for the convex-concave objective $\min_x \max_y f(x) + \langle Ax, y\rangle - g^*(y)$, where $f$ is a convex function with locally Lipschitz gradient and $g$ is convex and possibly…

Optimization and Control · Mathematics 2021-10-29 Maria-Luiza Vladarean , Yura Malitsky , Volkan Cevher

Extremum seeking control (ESC) is a classical adaptive control method for steady-state optimization, purely based on output feedback. It is well known that the extremum seeking control loop, under certain mild conditions on the controller,…

Optimization and Control · Mathematics 2018-02-26 Olle Trollberg , Elling W. Jacobsen

This article provides a rigorous analysis of convergence and stability of Episodic Upside-Down Reinforcement Learning, Goal-Conditioned Supervised Learning and Online Decision Transformers. These algorithms performed competitively across…

An evolution strategy (ES) variant based on a simplification of a natural evolution strategy recently attracted attention because it performs surprisingly well in challenging deep reinforcement learning domains. It searches for neural…

Neural and Evolutionary Computing · Computer Science 2018-05-03 Joel Lehman , Jay Chen , Jeff Clune , Kenneth O. Stanley

The one-fifth success rule is one of the best-known and most widely accepted techniques to control the parameters of evolutionary algorithms. While it is often applied in the literal sense, a common interpretation sees the one-fifth success…

Neural and Evolutionary Computing · Computer Science 2021-12-30 Benjamin Doerr , Carola Doerr , Johannes Lengler

We consider a stochastic individual-based model of adaptive dynamics on a finite trait graph $G=(V,E)$. The evolution is driven by a linear birth rate, a density dependent logistic death rate an the possibility of mutations along the…

Probability · Mathematics 2024-04-08 Manuel Esser , Anna Kraut

We consider non-convex stochastic optimization using first-order algorithms for which the gradient estimates may have heavy tails. We show that a combination of gradient clipping, momentum, and normalized gradient descent yields convergence…

Machine Learning · Computer Science 2021-11-10 Ashok Cutkosky , Harsh Mehta

This work examines the conditions for asymptotic and exponential convergence of saddle flow dynamics of convex-concave functions. First, we propose an observability-based certificate for asymptotic convergence, directly bridging the gap…

Optimization and Control · Mathematics 2024-09-10 Pengcheng You , Yingzhu Liu , Enrique Mallada

Under constant selection, each trait has a fixed fitness, and small mutation rates allow populations to efficiently exploit the optimal trait. Therefore it is reasonable to expect mutation rates will evolve downwards. However, we find this…

Populations and Evolution · Quantitative Biology 2022-08-23 Brian Mintz , Feng Fu

This paper analyzes the trajectories of stochastic gradient descent (SGD) to help understand the algorithm's convergence properties in non-convex problems. We first show that the sequence of iterates generated by SGD remains bounded and…

Optimization and Control · Mathematics 2020-06-22 Panayotis Mertikopoulos , Nadav Hallak , Ali Kavis , Volkan Cevher

Evolutionary algorithms (EAs) are general-purpose optimisers that come with several parameters like the sizes of parent and offspring populations or the mutation rate. It is well known that the performance of EAs may depend drastically on…

Neural and Evolutionary Computing · Computer Science 2022-10-13 Mario Alejandro Hevia Fajardo , Dirk Sudholt

This paper study the parameter selection of predefined-time sliding mode and try to design a general nonsingular predefined-time terminal sliding mode. 1). On parameter selection: Some existing predefined-time sliding modes are designed to…

Optimization and Control · Mathematics 2020-10-07 Wen Yan

We show that, for any c>0, the (1+1) evolutionary algorithm using an arbitrary mutation rate p_n = c/n finds the optimum of a linear objective function over bit strings of length n in expected time Theta(n log n). Previously, this was only…

Data Structures and Algorithms · Computer Science 2012-04-20 Benjamin Doerr , Leslie Ann Goldberg

We study the replica free energy surface for a spin glass model near the glassy temperature. In this model the simplicity of the equilibrium solution hides non trivial metastable saddle points. By means of the stability analysis performed…

Condensed Matter · Physics 2009-10-22 Marco E. Ferrero , Miguel A. Virasoro

We study reinforcement learning in stochastic path (SP) problems. The goal in these problems is to maximize the expected sum of rewards until the agent reaches a terminal state. We provide the first regret guarantees in this general problem…

Machine Learning · Computer Science 2022-10-18 Christoph Dann , Chen-Yu Wei , Julian Zimmert

We study the iteration complexity of the optimistic gradient descent-ascent (OGDA) method and the extra-gradient (EG) method for finding a saddle point of a convex-concave unconstrained min-max problem. To do so, we first show that both…

Optimization and Control · Mathematics 2020-09-30 Aryan Mokhtari , Asuman Ozdaglar , Sarath Pattathil

We show that on the manifold of fixed-rank and symmetric positive semi-definite matrices, the Riemannian gradient descent algorithm almost surely escapes some spurious critical points on the boundary of the manifold. Our result is the first…

Optimization and Control · Mathematics 2022-06-13 Thomas Y. Hou , Zhenzhen Li , Ziyun Zhang

In this paper we fully describe the trajectory of gradient flow over diagonal linear networks in the limit of vanishing initialisation. We show that the limiting flow successively jumps from a saddle of the training loss to another until…

Machine Learning · Computer Science 2023-10-26 Scott Pesme , Nicolas Flammarion

Steady states are invaluable in the study of dynamical systems. High-dimensional dynamical systems, due to a separation of time-scales, often evolve towards a lower dimensional manifold $M$. We introduce an approach to locate saddle points…

Dynamical Systems · Mathematics 2023-10-02 A. Georgiou , H. Vandecasteele , J. M. Bello-Rivas , I. Kevrekidis