中文
相关论文

相关论文: Lie Algebraic Cost Function Design for Control on …

200 篇论文

We address the problem of designing an LQR controller in a distributed setting, where M similar but not identical systems share their locally computed policy gradient (PG) estimates with a server that aggregates the estimates and computes a…

最优化与控制 · 数学 2024-04-16 Leonardo F. Toso , Han Wang , James Anderson

This paper addresses the inverse optimal control problem of finding the state weighting function that leads to a quadratic value function when the cost on the input is fixed to be quadratic. The paper focuses on a class of infinite horizon…

最优化与控制 · 数学 2022-11-21 Luis Rodrigues

This work presents a suboptimality study of a particular model predictive control with a stage cost shaping based on the ideas of reinforcement learning. The focus of the suboptimality study is to derive quantities relating the…

最优化与控制 · 数学 2020-04-28 Lukas Beckenbach , Pavel Osinenko , Stefan Streif

We study the convergence of model-based policy gradient for the deterministic, scalar, discounted linear-quadratic regulator when the controller is an overparameterized one-hidden-layer ReLU network without biases. Although the optimal LQR…

最优化与控制 · 数学 2026-04-27 Jhojan A. Rodriguez-Gil , César A. Uribe

This papers deals with the constrained discounted control of piecewise deterministic Markov process (PDMPs) in general Borel spaces. The control variable acts on the jump rate and transition measure, and the goal is to minimize the total…

最优化与控制 · 数学 2014-02-26 Oswaldo Costa , François Dufour

The aim of this work is to study, from an intrinsic and geometric point of view, second-order constrained variational problems on Lie algebroids, that is, optimization problems defined by a cost functional which depends on higher-order…

数学物理 · 物理学 2017-01-18 Leonardo Colombo

The two main contributions of this paper are a proof of concept of the recent novel idea in the area of long-time average cost control, and a new method of overcoming the well-known difficulty of non-convexity of simultaneous optimization…

最优化与控制 · 数学 2015-04-02 Deqing Huang , Sergei Chernyshenko

With a growing interest in data-driven control techniques, Model Predictive Control (MPC) provides an opportunity to exploit the surplus of data reliably, particularly while taking safety and stability into account. In many real-world and…

人工智能 · 计算机科学 2021-06-04 Mayank Mittal , Marco Gallieri , Alessio Quaglino , Seyed Sina Mirrazavi Salehian , Jan Koutník

A Learning Model Predictive Controller (LMPC) for iterative tasks is presented. The controller is reference-free and is able to improve its performance by learning from previous iterations. A safe set and a terminal cost function are used…

系统与控制 · 计算机科学 2017-12-15 Ugo Rosolia , Francesco Borrelli

While differentiable control has emerged as a powerful paradigm combining model-free flexibility with model-based efficiency, the iterative Linear Quadratic Regulator (iLQR) remains underexplored as a differentiable component. The…

机器人学 · 计算机科学 2025-06-24 Shuyuan Wang , Philip D. Loewen , Michael Forbes , Bhushan Gopaluni , Wei Pan

Optimization problems with convex quadratic cost and polyhedral constraints are ubiquitous in signal processing, automatic control and decision-making. We consider here an enlarged problem class that allows to encode logical conditions and…

最优化与控制 · 数学 2026-04-09 Alberto De Marchi

We consider the problem of stochastic optimal control, where the state-feedback control policies take the form of a probability distribution and where a penalty on the entropy is added. By viewing the cost function as a Kullback- Leibler…

最优化与控制 · 数学 2024-12-12 Marc Lambert , Francis Bach , Silvère Bonnabel

We propose new methods for learning control policies and neural network Lyapunov functions for nonlinear control problems, with provable guarantee of stability. The framework consists of a learner that attempts to find the control and…

机器学习 · 计算机科学 2022-09-26 Ya-Chien Chang , Nima Roohi , Sicun Gao

The controller of an input-affine system is determined through minimizing a time-varying objective function, where stabilization is ensured via a Lyapunov function decay condition as constraint. This constraint is incorporated into the…

系统与控制 · 电气工程与系统科学 2021-10-12 Patrick Schmidt , Thomas Göhrt , Stefan Streif

In this paper, we propose and analyze a new method for online linear quadratic regulator (LQR) control with a priori unknown time-varying cost matrices. The cost matrices are revealed sequentially with the potential for future values to be…

最优化与控制 · 数学 2023-02-22 Yitian Chen , Timothy L. Molloy , Tyler Summers , Iman Shames

Recent advances in the reinforcement learning (RL) literature have enabled roboticists to automatically train complex policies in simulated environments. However, due to the poor sample complexity of these methods, solving RL problems using…

机器人学 · 计算机科学 2022-11-21 Tyler Westenbroek , Fernando Castaneda , Ayush Agrawal , Shankar Sastry , Koushil Sreenath

Computation of derivatives (gradient and Hessian) of a fidelity function is one of the most crucial steps in many optimization algorithms. Having access to accurate methods to calculate these derivatives is even more desired where the…

计算物理 · 物理学 2020-03-05 Mohammadali Foroozandeh , Pranav Singh

We extend the Euler Poincare formalism from Lie groups to Lie groupoids for optimal control problems. While Lie algebroids provide the standard infinitesimal framework, the groupoid formulation enables global trajectory reconstruction and…

动力系统 · 数学 2026-05-26 Ghorbanali Haghighatdoost

We propose a new risk-constrained reformulation of the standard Linear Quadratic Regulator (LQR) problem. Our framework is motivated by the fact that the classical (risk-neutral) LQR controller, although optimal in expectation, might be…

系统与控制 · 电气工程与系统科学 2020-10-30 Anastasios Tsiamis , Dionysios S. Kalogerias , Luiz F. O. Chamon , Alejandro Ribeiro , George J. Pappas

Polyhedral control Lyapunov functions (PCLFs) are exploited in finite-horizon linear model predictive control formulations in order to guarantee the maximal domain of attraction (DoA), in contrast to traditional formulations based on…

系统与控制 · 计算机科学 2012-08-14 Sergio Grammatico , Gabriele Pannocchia