中文
相关论文

相关论文: Regularized Step Directions in Nonlinear Conjugate…

200 篇论文

In this paper, we develop a new accelerated stochastic gradient method for efficiently solving the convex regularized empirical risk minimization problem in mini-batch settings. The use of mini-batches is becoming a golden standard in the…

最优化与控制 · 数学 2017-09-20 Tomoya Murata , Taiji Suzuki

Many supervised learning tasks have intrinsic symmetries, such as translational and rotational symmetry in image classifications. These symmetries can be exploited to enhance performance. We formulate the symmetry constraints into a concise…

量子物理 · 物理学 2024-08-14 Kaiming Bian , Shitao Zhang , Fei Meng , Wen Zhang , Oscar Dahlsten

This paper presents a regularized Newton method (RNM) with generalized regularization terms for unconstrained convex optimization problems. The generalized regularization includes quadratic, cubic, and elastic net regularizations as special…

最优化与控制 · 数学 2024-07-11 Yuya Yamakawa , Nobuo Yamashita

In the paper [Muhammad Aslam Noor, Khalida Inayat Noor, Three-step iterative methods for nonlinear equations, Applied Mathematics and Computation, 183 (2006), pp. 322-327 ], Authors presented an algorithm (\textbf{Algorithm 2.3}) and stated…

数值分析 · 数学 2015-03-13 Laila M Assas , Fayyaz Ahmad , Malik Zaka Ullah

We suggest a conjugate subgradient type method without any line-search for minimization of convex non differentiable functions. Unlike the custom methods of this class, it does not require monotone decrease of the goal function and reduces…

最优化与控制 · 数学 2019-04-22 Igor Konnov

In this paper, we consider minimizing a sum of local convex objective functions in a distributed setting, where the cost of communication and/or computation can be expensive. We extend and generalize the analysis for a class of nested…

最优化与控制 · 数学 2021-09-01 Albert S. Berahas , Raghu Bollapragada , Ermin Wei

We consider a distributed multi-agent optimization problem over a time-invariant undirected graph, where each agent possesses a local objective function and all agents collaboratively minimize the average of all objective functions through…

最优化与控制 · 数学 2019-07-19 Juan Gao , Xinwei Liu , Yu-Hong Dai , Yakui Huang , Peng Yang

In the setting of federated optimization, where a global model is aggregated periodically, step asynchronism occurs when participants conduct model training by efficiently utilizing their computational resources. It is well acknowledged…

分布式、并行与集群计算 · 计算机科学 2023-02-28 Feijie Wu , Song Guo , Haozhao Wang , Zhihao Qu , Haobo Zhang , Jie Zhang , Ziming Liu

We propose a fully-corrective generalized conditional gradient method (FC-GCG) for the minimization of the sum of a smooth, convex loss function and a convex one-homogeneous regularizer over a Banach space. The algorithm relies on the…

最优化与控制 · 数学 2023-07-17 Kristian Bredies , Marcello Carioni , Silvio Fanzon , Daniel Walter

We propose a stochastic variance-reduced cubic regularized Newton method for non-convex optimization. At the core of our algorithm is a novel semi-stochastic gradient along with a semi-stochastic Hessian, which are specifically designed for…

机器学习 · 计算机科学 2018-02-14 Dongruo Zhou , Pan Xu , Quanquan Gu

Mathematical reasoning has been challenging for large language models (LLMs), and the introduction of step-by-step Chain-of-Thought (CoT) inference has significantly advanced the mathematical capabilities of LLMs. However, current…

人工智能 · 计算机科学 2025-09-23 Lang Cao , Yingtian Zou , Chao Peng , Renhong Chen , Wu Ning , Yitong Li

Finite-difference methods are widely used for zeroth-order optimization in settings where gradient information is unavailable or expensive to compute. These procedures mimic first-order strategies by approximating gradients through function…

最优化与控制 · 数学 2025-05-27 Marco Rando , Cesare Molinari , Lorenzo Rosasco , Silvia Villa

The renewed interest in Steepest Descent (SD) methods following the work of Barzilai and Borwein [IMA Journal of Numerical Analysis, 8 (1988)] has driven us to consider a globalization strategy based on SD, which is applicable to any…

最优化与控制 · 数学 2020-06-24 Daniela di Serafino , Gerardo Toraldo , Marco Viola

Subgradient methods are the natural extension to the non-smooth case of the classical gradient descent for regular convex optimization problems. However, in general, they are characterized by slow convergence rates, and they require…

最优化与控制 · 数学 2023-11-20 Alessandro Scagliotti , Piero Colli Franzone

In this paper, two new subspace minimization conjugate gradient methods based on $p - $regularization models are proposed, where a special scaled norm in $p - $regularization model is analyzed. Different choices for special scaled norm lead…

最优化与控制 · 数学 2020-04-06 Ting Zhao , Hongwei Liu , Zexian Liu

In this paper, we initiate a study of functional minimization in Federated Learning. First, in the semi-heterogeneous setting, when the marginal distributions of the feature vectors on client machines are identical, we develop the federated…

机器学习 · 计算机科学 2021-03-15 Zebang Shen , Hamed Hassani , Satyen Kale , Amin Karbasi

Distributed training of massive machine learning models, in particular deep neural networks, via Stochastic Gradient Descent (SGD) is becoming commonplace. Several families of communication-reduction methods, such as quantization,…

A set of accelerated first order algorithms with memory are proposed for minimising strongly convex functions. The algorithms are differentiated by their use of the iterate history for the gradient step. The increased convergence rate of…

最优化与控制 · 数学 2018-08-31 Ross Drummond , Stephen Duncan

Approximating complex curves with simple parametric curves is widely used in CAGD, CG, and CNC. This paper presents an algorithm to compute a certified approximation to a given parametric space curve with cubic B-spline curves. By…

计算几何 · 计算机科学 2012-03-05 Liyong Shen , Chunming Yuan , Xiao-Shan Gao

Nonlinear acceleration algorithms improve the performance of iterative methods, such as gradient descent, using the information contained in past iterates. However, their efficiency is still not entirely understood even in the quadratic…

最优化与控制 · 数学 2019-03-22 Damien Scieur