English
Related papers

Related papers: Modified limited memory BFGS with displacement agg…

200 papers

Fine-tuning large foundation models presents significant memory challenges due to stateful optimizers like AdamW, often requiring several times more GPU memory than inference. While memory-efficient methods like parameter-efficient…

Machine Learning · Computer Science 2025-07-14 Pouria Mahdavinia , Mehrdad Mahdavi

The augmented Lagrangiam method (ALM), widely used in quantum chemistry constrained optimization problems, is applied in the context of the nuclear Density Functional Theory (DFT) in the self-consistent constrained Skyrme…

Nuclear Theory · Physics 2014-11-21 A. Staszczak , M. Stoitsov , A. Baran , W. Nazarewicz

This work addresses weight optimization problem for fully-connected feed-forward neural networks. Unlike existing approaches that are based on back-propagation (BP) and chain rule gradient-based optimization (which implies iterative…

Machine Learning · Computer Science 2024-06-18 Slavisa Tomic , João Pedro Matos-Carvalho , Marko Beko

In this article, a novel barrier function is introduced to convert the box-constrained convex optimization problem to an unconstrained problem. For each double-sided bounded variable, a single monomial function is added as a barrier…

Optimization and Control · Mathematics 2024-01-31 Hatem Fayed

We present a stochastic setting for optimization problems with nonsmooth convex separable objective functions over linear equality constraints. To solve such problems, we propose a stochastic Alternating Direction Method of Multipliers…

Machine Learning · Computer Science 2013-01-23 Hua Ouyang , Niao He , Alexander Gray

Recently, several works have shown that natural modifications of the classical conditional gradient method (aka Frank-Wolfe algorithm) for constrained convex optimization, provably converge with a linear rate when: i) the feasible set is a…

Optimization and Control · Mathematics 2016-05-23 Dan Garber , Ofer Meshi

We study unconstrained optimization problems with nonsmooth and convex objective function in the form of a mathematical expectation. The proposed method approximates the expected objective function with a sample average function using…

Optimization and Control · Mathematics 2022-11-03 Natasa Krejic , Natasa Krklec Jerinkic , Tijana Ostojic

Shape-constrained convex regression problem deals with fitting a convex function to the observed data, where additional constraints are imposed, such as component-wise monotonicity and uniform Lipschitz continuity. This paper provides a…

Optimization and Control · Mathematics 2020-02-27 Meixia Lin , Defeng Sun , Kim-Chuan Toh

Recent research indicates that the performance of machine learning models can be improved by aligning the geometry of the latent space with the underlying data structure. Rather than relying solely on Euclidean space, researchers have…

Machine Learning · Computer Science 2023-10-30 Haitz Saez de Ocariz Borde , Alvaro Arroyo , Ismael Morales , Ingmar Posner , Xiaowen Dong

Asynchronous federated learning (FL) with heterogeneous clients faces two key issues: curvature-induced loss barriers encountered by standard linear parameter interpolation techniques (e.g. FedAvg) and interference from stale updates…

Machine Learning · Computer Science 2025-10-13 Archie Licudi , Anshul Thakur , Soheila Molaei , Danielle Belgrave , David Clifton

In this paper, we consider a class of nonsmooth nonconvex optimization problems whose objective is the sum of a block relative smooth function and a proper and lower semicontinuous block separable function. Although the analysis of block…

Optimization and Control · Mathematics 2022-04-27 Le Thi Khanh Hien , Duy Nhat Phan , Nicolas Gillis , Masoud Ahookhosh , Panagiotis Patrinos

Many inverse problems are phrased as optimization problems in which the objective function is the sum of a data-fidelity term and a regularization. Often, the Hessian of the fidelity term is computationally unavailable while the Hessian of…

Optimization and Control · Mathematics 2024-03-12 Florian Mannel , Hari Om Aggrawal , Jan Modersitzki

In this paper, we explore the non-asymptotic global convergence rates of the Broyden-Fletcher-Goldfarb-Shanno (BFGS) method implemented with exact line search. Notably, due to Dixon's equivalence result, our findings are also applicable to…

Optimization and Control · Mathematics 2025-07-16 Qiujiang Jin , Ruichen Jiang , Aryan Mokhtari

Recently, minimax optimization received renewed focus due to modern applications in machine learning, robust optimization, and reinforcement learning. The scale of these applications naturally leads to the use of first-order methods.…

Optimization and Control · Mathematics 2023-03-07 Saeed Hajizadeh , Haihao Lu , Benjamin Grimmer

Motivated by the training of Generative Adversarial Networks (GANs), we study methods for solving minimax problems with additional nonsmooth regularizers. We do so by employing \emph{monotone operator} theory, in particular the…

Optimization and Control · Mathematics 2020-06-17 Axel Böhm , Michael Sedlmayer , Ernö Robert Csetnek , Radu Ioan Boţ

Reconstructing high-quality images with sharp edges requires the use of edge-preserving constraints in the regularized form of the inverse problem. The use of the $\ell_q$-norm on the gradient of the image is a common such constraint. For…

Numerical Analysis · Mathematics 2023-09-28 Mirjeta Pasha , Eric de Sturler , Misha E. Kilmer

We integrate the diagonal quasi-Newton update approach with the enhanced BFGS formula proposed by Wei, Z., Yu, G., Yuan, G., Lian, Z. \cite{b1}, incorporating extrapolation techniques and inertia acceleration technology. This method,…

Optimization and Control · Mathematics 2025-07-08 Zhenhua Luo , Gonglin Yuan , Hongtruong Pham

For general large-scale optimization problems compact representations exist in which recursive quasi-Newton update formulas are represented as compact matrix factorizations. For problems in which the objective function contains additional…

Optimization and Control · Mathematics 2022-08-02 Johannes J. Brust , Zichao , Di , Sven Leyffer , Cosmin G. Petra

This paper considers consensus optimization problems where each node of a network has access to a different summand of an aggregate cost function. Nodes try to minimize the aggregate cost function, while they exchange information only with…

Optimization and Control · Mathematics 2016-03-24 Mark Eisen , Aryan Mokhtari , Alejandro Ribeiro

In practice, many machine learning (ML) problems come with constraints, and their applied domains involve distributed sensitive data that cannot be shared with others, e.g., in healthcare. Collaborative learning in such practical scenarios…

Machine Learning · Computer Science 2024-05-02 Chuan He , Le Peng , Ju Sun