中文
相关论文

相关论文: Global Convergence of Second-order Dynamics in Two…

200 篇论文

We present a unified convergence analysis for first order convex optimization methods using the concept of strong Lyapunov conditions. Combining this with suitable time scaling factors, we are able to handle both convex and strong convex…

最优化与控制 · 数学 2021-08-03 Long Chen , Hao Luo

We introduce new multilevel methods for solving large-scale unconstrained optimization problems. Specifically, the philosophy of multilevel methods is applied to Newton-type methods that regularize the Newton sub-problem using second order…

最优化与控制 · 数学 2024-07-16 Nick Tsipinakis , Panos Parpas

In the context of over-parameterization, there is a line of work demonstrating that randomly initialized (stochastic) gradient descent (GD) converges to a globally optimal solution at a linear convergence rate for the quadratic loss…

机器学习 · 计算机科学 2025-06-16 Xianliang Xu , Ting Du , Wang Kong , Bin Shan , Ye Li , Zhongyi Huang

In this paper, we focus on providing convergence guarantees for stochastic subgradient methods in minimizing nonsmooth nonconvex functions. We first investigate the global stability of a general framework for stochastic subgradient methods,…

最优化与控制 · 数学 2024-10-15 Nachuan Xiao , Xiaoyin Hu , Kim-Chuan Toh

We study the overparametrization bounds required for the global convergence of stochastic gradient descent algorithm for a class of one hidden layer feed-forward neural networks, considering most of the activation functions used in…

机器学习 · 计算机科学 2022-11-17 Bartłomiej Polaczyk , Jacek Cyranka

We give a simple proof for the global convergence of gradient descent in training deep ReLU networks with the standard square loss, and show some of its improvements over the state-of-the-art. In particular, while prior works require all…

机器学习 · 计算机科学 2021-06-14 Quynh Nguyen

Classical global convergence results for first-order methods rely on uniform smoothness and the \L{}ojasiewicz inequality. Motivated by properties of objective functions that arise in machine learning, we propose a non-uniform refinement of…

机器学习 · 计算机科学 2022-06-03 Jincheng Mei , Yue Gao , Bo Dai , Csaba Szepesvari , Dale Schuurmans

We prove that the Gini coefficient of economic inequality is a Lyapunov functional for a class of nonlinear, nonlocal integro-differential equations arising at the intersection of mathematics, economics, and statistical physics. Next, a…

偏微分方程分析 · 数学 2026-02-23 David W. Cohen

We consider a one dimensional transport model with nonlocal velocity given by the Hilbert transform and develop a global well-posedness theory of probability measure solutions. Both the viscous and non-viscous cases are analyzed. Both in…

偏微分方程分析 · 数学 2011-11-01 J. A. Carrillo , L. C. F. Ferreira , J. C. Precioso

We consider a one-dimensional kinetic model of granular media in the case where the interaction potential is quadratic. Taking advan- tage of a simple first integral, we can use a reformulation (equivalent to the initial kinetic model for…

偏微分方程分析 · 数学 2015-06-19 Martial Agueh , Guillaume Carlier

We show that the Nernst-Planck-Euler system, which models ionic electrodiffusion in fluids, has global strong solutions for arbitrarily large data in the two dimensional bounded domains. The assumption on species is either there are two…

偏微分方程分析 · 数学 2022-12-27 Dapeng Du , Jingyu Li , Yansheng Ma , Ruyi Pang

We propose an unconstrained optimization method based on the well-known primal-dual hybrid gradient (PDHG) algorithm. We first formulate the optimality condition of the unconstrained optimization problem as a saddle point problem. We then…

最优化与控制 · 数学 2024-08-29 X. Zuo , S. Osher , W. Li

We present a novel notion of $\lambda$-monotonicity for an $n$-species system of partial differential equations governed by mass-preserving flow dynamics, extending monotonicity in Banach spaces to the Wasserstein-2 metric space. We show…

偏微分方程分析 · 数学 2025-09-29 Lauren Conger , Franca Hoffmann , Eric Mazumdar , Lillian J. Ratliff

Deep learning models are often successfully trained using gradient descent, despite the worst case hardness of the underlying non-convex optimization problem. The key question is then under what conditions can one prove that optimization…

机器学习 · 计算机科学 2017-02-28 Alon Brutzkus , Amir Globerson

There are much recent interests in solving noncovnex min-max optimization problems due to its broad applications in many areas including machine learning, networked resource allocations, and distributed optimization. Perhaps, the most…

最优化与控制 · 数学 2021-12-20 Thinh T. Doan

Over the past fifteen years, the theory of Wasserstein gradient flows of convex (or, more generally, semiconvex) energies has led to advances in several areas of partial differential equations and analysis. In this work, we extend the…

偏微分方程分析 · 数学 2017-05-04 Katy Craig

The symmetric low-rank matrix factorization serves as a building block in many learning tasks, including matrix recovery and training of neural networks. However, despite a flurry of recent research, the dynamics of its training via…

We analyze recurrent neural networks with diagonal hidden-to-hidden weight matrices, trained with gradient descent in the supervised learning setting, and prove that gradient descent can achieve optimality \emph{without} massive…

机器学习 · 计算机科学 2024-10-11 Semih Cayci , Atilla Eryilmaz

Minimizing loss functions is central to machine-learning training. Although first-order methods dominate practical applications, higher-order techniques such as Newton's method can deliver greater accuracy and faster convergence, yet are…

机器学习 · 计算机科学 2025-11-25 Giuseppe Carrino , Elena Loli Piccolomini , Elisa Riccietti , Theo Mary

We introduce a framework for Newton's flows in probability space with information metrics, named information Newton's flows. Here two information metrics are considered, including both the Fisher-Rao metric and the Wasserstein-2 metric. A…

最优化与控制 · 数学 2020-08-06 Yifei Wang , Wuchen Li
‹ 上一页 1 8 9 10 下一页 ›