English
Related papers

Related papers: Global Convergence of Second-order Dynamics in Two…

200 papers

We present a unified convergence analysis for first order convex optimization methods using the concept of strong Lyapunov conditions. Combining this with suitable time scaling factors, we are able to handle both convex and strong convex…

Optimization and Control · Mathematics 2021-08-03 Long Chen , Hao Luo

We introduce new multilevel methods for solving large-scale unconstrained optimization problems. Specifically, the philosophy of multilevel methods is applied to Newton-type methods that regularize the Newton sub-problem using second order…

Optimization and Control · Mathematics 2024-07-16 Nick Tsipinakis , Panos Parpas

In the context of over-parameterization, there is a line of work demonstrating that randomly initialized (stochastic) gradient descent (GD) converges to a globally optimal solution at a linear convergence rate for the quadratic loss…

Machine Learning · Computer Science 2025-06-16 Xianliang Xu , Ting Du , Wang Kong , Bin Shan , Ye Li , Zhongyi Huang

In this paper, we focus on providing convergence guarantees for stochastic subgradient methods in minimizing nonsmooth nonconvex functions. We first investigate the global stability of a general framework for stochastic subgradient methods,…

Optimization and Control · Mathematics 2024-10-15 Nachuan Xiao , Xiaoyin Hu , Kim-Chuan Toh

We study the overparametrization bounds required for the global convergence of stochastic gradient descent algorithm for a class of one hidden layer feed-forward neural networks, considering most of the activation functions used in…

Machine Learning · Computer Science 2022-11-17 Bartłomiej Polaczyk , Jacek Cyranka

We give a simple proof for the global convergence of gradient descent in training deep ReLU networks with the standard square loss, and show some of its improvements over the state-of-the-art. In particular, while prior works require all…

Machine Learning · Computer Science 2021-06-14 Quynh Nguyen

Classical global convergence results for first-order methods rely on uniform smoothness and the \L{}ojasiewicz inequality. Motivated by properties of objective functions that arise in machine learning, we propose a non-uniform refinement of…

Machine Learning · Computer Science 2022-06-03 Jincheng Mei , Yue Gao , Bo Dai , Csaba Szepesvari , Dale Schuurmans

We prove that the Gini coefficient of economic inequality is a Lyapunov functional for a class of nonlinear, nonlocal integro-differential equations arising at the intersection of mathematics, economics, and statistical physics. Next, a…

Analysis of PDEs · Mathematics 2026-02-23 David W. Cohen

We consider a one dimensional transport model with nonlocal velocity given by the Hilbert transform and develop a global well-posedness theory of probability measure solutions. Both the viscous and non-viscous cases are analyzed. Both in…

Analysis of PDEs · Mathematics 2011-11-01 J. A. Carrillo , L. C. F. Ferreira , J. C. Precioso

We consider a one-dimensional kinetic model of granular media in the case where the interaction potential is quadratic. Taking advan- tage of a simple first integral, we can use a reformulation (equivalent to the initial kinetic model for…

Analysis of PDEs · Mathematics 2015-06-19 Martial Agueh , Guillaume Carlier

We show that the Nernst-Planck-Euler system, which models ionic electrodiffusion in fluids, has global strong solutions for arbitrarily large data in the two dimensional bounded domains. The assumption on species is either there are two…

Analysis of PDEs · Mathematics 2022-12-27 Dapeng Du , Jingyu Li , Yansheng Ma , Ruyi Pang

We propose an unconstrained optimization method based on the well-known primal-dual hybrid gradient (PDHG) algorithm. We first formulate the optimality condition of the unconstrained optimization problem as a saddle point problem. We then…

Optimization and Control · Mathematics 2024-08-29 X. Zuo , S. Osher , W. Li

We present a novel notion of $\lambda$-monotonicity for an $n$-species system of partial differential equations governed by mass-preserving flow dynamics, extending monotonicity in Banach spaces to the Wasserstein-2 metric space. We show…

Analysis of PDEs · Mathematics 2025-09-29 Lauren Conger , Franca Hoffmann , Eric Mazumdar , Lillian J. Ratliff

Deep learning models are often successfully trained using gradient descent, despite the worst case hardness of the underlying non-convex optimization problem. The key question is then under what conditions can one prove that optimization…

Machine Learning · Computer Science 2017-02-28 Alon Brutzkus , Amir Globerson

There are much recent interests in solving noncovnex min-max optimization problems due to its broad applications in many areas including machine learning, networked resource allocations, and distributed optimization. Perhaps, the most…

Optimization and Control · Mathematics 2021-12-20 Thinh T. Doan

Over the past fifteen years, the theory of Wasserstein gradient flows of convex (or, more generally, semiconvex) energies has led to advances in several areas of partial differential equations and analysis. In this work, we extend the…

Analysis of PDEs · Mathematics 2017-05-04 Katy Craig

The symmetric low-rank matrix factorization serves as a building block in many learning tasks, including matrix recovery and training of neural networks. However, despite a flurry of recent research, the dynamics of its training via…

Machine Learning · Computer Science 2024-11-26 Hesameddin Mohammadi , Mohammad Tinati , Stephen Tu , Mahdi Soltanolkotabi , Mihailo R. Jovanović

We analyze recurrent neural networks with diagonal hidden-to-hidden weight matrices, trained with gradient descent in the supervised learning setting, and prove that gradient descent can achieve optimality \emph{without} massive…

Machine Learning · Computer Science 2024-10-11 Semih Cayci , Atilla Eryilmaz

Minimizing loss functions is central to machine-learning training. Although first-order methods dominate practical applications, higher-order techniques such as Newton's method can deliver greater accuracy and faster convergence, yet are…

Machine Learning · Computer Science 2025-11-25 Giuseppe Carrino , Elena Loli Piccolomini , Elisa Riccietti , Theo Mary

We introduce a framework for Newton's flows in probability space with information metrics, named information Newton's flows. Here two information metrics are considered, including both the Fisher-Rao metric and the Wasserstein-2 metric. A…

Optimization and Control · Mathematics 2020-08-06 Yifei Wang , Wuchen Li
‹ Prev 1 8 9 10 Next ›