中文
相关论文

相关论文: Natural Gradient Descent for Control

200 篇论文

In this work, we introduce a novel gradient descent-based approach for optimizing control systems, leveraging a new representation of stable closed-loop dynamics as a function of two matrices i.e. the step size or direction matrix and value…

最优化与控制 · 数学 2024-09-18 Ramin Esmzad , Hamidreza Modares

Natural gradient descent is an optimization method traditionally motivated from the perspective of information geometry, and works well for many applications as an alternative to stochastic gradient descent. In this paper we critically…

机器学习 · 计算机科学 2020-09-22 James Martens

Natural Gradient Descent, a second-degree optimization method motivated by the information geometry, makes use of the Fisher Information Matrix instead of the Hessian which is typically used. However, in many cases, the Fisher Information…

机器学习 · 计算机科学 2023-03-10 Rajesh Shrestha

This paper considers the problem of regulating a linear dynamical system to the solution of a convex optimization problem with an unknown or partially-known cost. We design a data-driven feedback controller - based on gradient flow dynamics…

最优化与控制 · 数学 2022-04-05 Liliaokeawawa Cothren , Gianluca Bianchin , Emiliano Dall'Anese

We present a gradient-based optimal-control technique for open quantum systems that utilizes quantum trajectories to simulate the quantum dynamics during optimization. Using trajectories allows for optimizing open systems with less…

量子物理 · 物理学 2019-06-03 Mohamed Abdelhafez , David I. Schuster , Jens Koch

In this paper, online convex optimization is applied to the problem of controlling linear dynamical systems. An algorithm similar to online gradient descent, which can handle time-varying and unknown cost functions, is proposed. Then,…

最优化与控制 · 数学 2021-11-03 Marko Nonhoff , Matthias A. Müller

This paper presents a data-driven method to find a closed-loop optimal controller, which minimizes a specified infinite-horizon cost function for systems with unknown dynamics. Suppose the closed-loop optimal controller can be parameterized…

机器学习 · 计算机科学 2025-11-20 Wenjian Hao , Paulo C. Heredia , Shaoshuai Mou

This paper presents a discrete-time passivity-based analysis of the gradient descent method for a class of functions with sector-bounded gradients. Using a loop transformation, it is shown that the gradient descent method can be interpreted…

最优化与控制 · 数学 2024-11-26 Sepehr Moalemi , James Richard Forbes

Bearing-based distributed formation control is attractive because it can be implemented using vision-based measurements to achieve a desired formation. Gradient-descent-based controllers using bearing measurements have been shown to have…

系统与控制 · 电气工程与系统科学 2022-03-25 Zili Wang , Sean B. Andersson , Roberto Tron

In this paper we We propose GoPRONTO, a first-order, feedback-based approach to solve nonlinear discrete-time optimal control problems. This method is a generalized first-order framework based on incorporating the original dynamics into a…

最优化与控制 · 数学 2023-08-22 Lorenzo Sforni , Sara Spedicato , Ivano Notarnicola , Giuseppe Notarstefano

Motivated by recent advances of reinforcement learning and direct data-driven control, we propose policy gradient adaptive control (PGAC) for the linear quadratic regulator (LQR), which uses online closed-loop data to improve the control…

最优化与控制 · 数学 2025-06-16 Feiran Zhao , Alessandro Chiuso , Florian Dörfler

In this paper, we present a novel control scheme for feedback optimization. That is, we propose a discrete-time controller that can steer the steady state of a physical plant to the solution of a constrained optimization problem without…

系统与控制 · 电气工程与系统科学 2020-07-09 Verena Häberle , Adrian Hauswirth , Lukas Ortmann , Saverio Bolognani , Florian Dörfler

In this paper, we consider continuous-time stochastic optimal control problems where the cost is evaluated through a coherent risk measure. We provide an explicit gradient descent-ascent algorithm which applies to problems subject to…

最优化与控制 · 数学 2023-06-23 Gabriel Velho , Jean Auriol , Riccardo Bonalli

The design of the performance index, also referred to as cost or reward shaping, is central to both optimal control and reinforcement learning, as it directly determines the behaviors, trade-offs, and objectives that the resulting control…

系统与控制 · 电气工程与系统科学 2025-10-14 Ayush Rai , Shaoshuai Mou , Brian D. O. Anderson

Feedback optimization is an increasingly popular control paradigm to optimize dynamical systems, accounting for control objectives that concern the system operation at steady-state. Existing feedback optimization techniques heavily rely on…

最优化与控制 · 数学 2025-04-08 Amir Mehrnoosh , Gianluca Bianchin

This paper investigates a novel finite-time gradient descent-based adaptive neural network finite-time control strategy for the attitude tracking of a 3-DOF lab helicopter platform subject to composite disturbances. First, the radial basis…

系统与控制 · 电气工程与系统科学 2021-07-28 Xidong Wang

Feedback optimization enables autonomous optimality seeking of a dynamical system through its closed-loop interconnection with iterative optimization algorithms. Among various iteration structures, model-based approaches require the…

最优化与控制 · 数学 2026-05-26 Zhiyu He , Saverio Bolognani , Michael Muehlebach , Florian Dörfler

Natural gradient descent (NGD) provided deep insights and powerful tools to deep neural networks. However the computation of Fisher information matrix becomes more and more difficult as the network structure turns large and complex. This…

机器学习 · 计算机科学 2021-09-22 Weihua Liu , Xiabi Liu

The problem of designing adaptive stepsize sequences for the gradient descent method applied to convex and locally smooth functions is studied. We take an adaptive control perspective and design update rules for the stepsize that make use…

最优化与控制 · 数学 2025-08-27 Andrea Iannelli

One approach for feedback control using high dimensional and rich sensor measurements is to classify the measurement into one out of a finite set of situations, each situation corresponding to a (known) control action. This approach…

最优化与控制 · 数学 2019-03-12 Hasan A. Poonawala , Niklas Lauffer , Ufuk Topcu
‹ 上一页 1 2 3 10 下一页 ›