中文
相关论文

相关论文: Trust-Region Method for Optimization of Set-Valued…

200 篇论文

Most methods in reinforcement learning use a Policy Gradient (PG) approach to learn a parametric stochastic policy that maps states to actions. The standard approach is to implement such a mapping via a neural network (NN) whose parameters…

机器学习 · 计算机科学 2024-05-29 Sergio Rozada , Antonio G. Marques

We introduce a constrained optimization method for policy gradient reinforcement learning, which uses a virtual trust region to regulate each policy update. In addition to using the proximity of one single old policy as the normal trust…

机器学习 · 计算机科学 2022-09-19 Hung Le , Thommen Karimpanal George , Majid Abdolshah , Dung Nguyen , Kien Do , Sunil Gupta , Svetha Venkatesh

Manifold optimization has recently gained significant attention due to its wide range of applications in various areas. This paper introduces the first Riemannian trust region method for minimizing an SC$^1$ function, which is a…

最优化与控制 · 数学 2024-06-03 Chenyu Zhang , Rufeng Xiao , Wen Huang , Rujun Jiang

Topology optimization problems often support multiple local minima due to a lack of convexity. Typically, gradient-based techniques combined with continuation in model parameters are used to promote convergence to more optimal solutions;…

数值分析 · 数学 2021-01-13 Ioannis P. A. Papadopoulos , Patrick E. Farrell , Thomas M. Surowiec

Physics-informed machine learning and inverse modeling require the solution of ill-conditioned non-convex optimization problems. First-order methods, such as SGD and ADAM, and quasi-Newton methods, such as BFGS and L-BFGS, have been applied…

数值分析 · 数学 2021-05-18 Kailai Xu , Eric Darve

Convergence guarantees for optimization over bounded-rank matrices are delicate to obtain because the feasible set is a non-smooth and non-convex algebraic variety. Existing techniques include projected gradient descent, fixed-rank…

最优化与控制 · 数学 2024-06-21 Quentin Rebjock , Nicolas Boumal

We present a novel derivative-free interpolation based optimization algorithm. A trust-region method is used where a surrogate model is realized via an interpolation framework. The framework for interpolation is provided by Universal…

最优化与控制 · 数学 2018-05-31 Tom Lefebvre , Frederik De Belie , Guillaume Crevecoeur

In this paper, we solve the l2-l1 sparse recovery problem by transforming the objective function of this problem into an unconstrained differentiable function and apply a limited-memory trust-region method. Unlike gradient projection-type…

数值分析 · 数学 2016-03-01 Lasith Adhikari , Jennifer B. Erway , Shelby Lockhart , Roummel F. Marcia

Acquisition of training data for the standard semantic segmentation is expensive if requiring that each pixel is labeled. Yet, current methods significantly deteriorate in weakly supervised settings, e.g. where a fraction of pixels is…

计算机视觉与模式识别 · 计算机科学 2021-10-13 Dmitrii Marin , Yuri Boykov

Coordinate update/descent algorithms are widely used in large-scale optimization due to their low per-iteration cost and scalability, but their behavior on infeasible or misspecified problems has not been much studied compared to the…

最优化与控制 · 数学 2024-10-14 Jinhee Paeng , Jisun Park , Ernest K. Ryu

We consider the problem of provably finding a stationary point of a smooth function to be minimized on the variety of bounded-rank matrices. This turns out to be unexpectedly delicate. We trace the difficulty back to a geometric obstacle:…

最优化与控制 · 数学 2022-07-11 Eitan Levin , Joe Kileel , Nicolas Boumal

Real-world optimization problems often involve complex objective functions with costly evaluations. While Bayesian optimization (BO) with Gaussian processes is effective for these challenges, it suffers in high-dimensional spaces due to…

机器学习 · 计算机科学 2024-12-17 Nobuo Namura , Sho Takemori

The trust region subproblem (TRS) is to minimize a possibly nonconvex quadratic function over a Euclidean ball. There are typically two cases for (TRS), the so-called ``easy case'' and ``hard case''. Even in the ``easy case'', the sequence…

最优化与控制 · 数学 2022-07-13 Mengmeng Song , Yong Xia , Jinyang Zheng

Bayesian optimization is a powerful tool for solving real-world optimization tasks under tight evaluation budgets, making it well-suited for applications involving costly simulations or experiments. However, many of these tasks are also…

机器学习 · 计算机科学 2025-06-18 Paolo Ascia , Elena Raponi , Thomas Bäck , Fabian Duddeck

In this paper, we study a method for finding robust solutions to multiobjective optimization problems under uncertainty. We follow the set-based minmax approach for handling the uncertainties which leads to a certain set optimization…

最优化与控制 · 数学 2022-12-29 Gabriele Eichfelder , Ernest Quintana

In this paper, a sequential search method for finding the global minimum of an objective function is presented, The descent gradient search is repeated until the global minimum is obtained. The global minimum is located by a process of…

最优化与控制 · 数学 2024-02-06 Mohamed Tifroute , Anouar Lahmdani , Hassane Bouzahir

This short note considers an efficient variant of the trust-region algorithm with dynamic accuracy proposed Carter (1993) and Conn, Gould and Toint (2000) as a tool for very high-performance computing, an area where it is critical to allow…

数值分析 · 计算机科学 2019-04-16 S. Gratton , Ph. L. Toint

To facilitate widespread adoption of automated engineering design techniques, existing methods must become more efficient and generalizable. In the field of topology optimization, this requires the coupling of modern optimization methods…

计算工程、金融与科学 · 计算机科学 2024-02-23 Connor N. Mallon , Aaron W. Thornton , Matthew R. Hill , Santiago Badia

In this work we classify the at-point regularities of set-valued mappings into two categories and then we analyze their relationship through several implications and examples. After this theoretical tour, we use the subregularity properties…

最优化与控制 · 数学 2012-02-07 Marius Apetrii , Marius Durea , Radu Strugariu

We develop a worst-case evaluation complexity bound for trust-region methods in the presence of unbounded Hessian approximations. We use the algorithm of arXiv:2103.15993v3 as a model, which is designed for nonsmooth regularized problems,…

最优化与控制 · 数学 2025-10-14 Geoffroy Leconte , Dominique Orban