中文
相关论文

相关论文: Model-Free $\delta$-Policy Iteration Based on Damp…

200 篇论文

We present an Imitation Learning approach for the control of dynamical systems with a known model. Our policy search method is guided by solutions from MPC. Typical policy search methods of this kind minimize a distance metric between the…

机器人学 · 计算机科学 2020-02-18 Jan Carius , Farbod Farshidian , Marco Hutter

This article studies the control ideas of the optimal backstepping technique, proposing an event-triggered optimal tracking control scheme for a class of strict-feedback nonlinear systems with non-affine and nonlinear faults. A simplified…

最优化与控制 · 数学 2024-06-13 Ling Wang , Xin Wang , Ziming Wang

In this work, we present a learning-based nonlinear $H^\infty$ control algorithm that guarantee system performance under learned dynamics and disturbance estimate. The Gaussian Process (GP) regression is utilized to update the nominal…

系统与控制 · 电气工程与系统科学 2021-07-12 Wei Sun , Theodore B. Trafalis

Port-Hamiltonian systems (PHS) and interconnection and damping assignment passivity-based control (IDA-PBC) have achieved broad success in modelling and stabilisation of physical systems. However, the absence of a dedicated scalar potential…

系统与控制 · 电气工程与系统科学 2026-05-14 Jinjun Jia , Yuchen Liao , Kang An , Xun Yan , Tiedong Zhang , Dapeng Jiang

This paper addresses distributional offline continuous-time reinforcement learning (DOCTR-L) with stochastic policies for high-dimensional optimal control. A soft distributional version of the classical Hamilton-Jacobi-Bellman (HJB)…

机器学习 · 计算机科学 2021-04-05 Igor Halperin

This paper proposes a spectral-based tuning method for proportional-integral (PI) controllers in integrating-plus-dead-time (IPDT) systems. The design objective is to achieve unified exponential decay for both reference tracking and…

系统与控制 · 电气工程与系统科学 2025-07-03 Dhamdhawach Horsuwan

This paper presents a generalizable methodology for data-driven identification of nonlinear dynamics that bounds the model error in terms of the prediction horizon and the magnitude of the derivatives of the system states. Using…

机器学习 · 统计学 2021-05-03 Giorgos Mamakoukas , Maria L. Castano , Xiaobo Tan , Todd D. Murphey

In this paper, we present a scalable deep learning approach to solve opinion dynamics stochastic optimal control problems with mean field term coupling in the dynamics and cost function. Our approach relies on the probabilistic…

多智能体系统 · 计算机科学 2022-04-19 Tianrong Chen , Ziyi Wang , Evangelos A. Theodorou

The iterative problem of solving nonlinear equations is studied. A new Newton like iterative method with adjustable parameters is designed based on the dynamic system theory. In order to avoid the derivative function in the iterative…

数值分析 · 数学 2022-11-09 Yonglong Liao , Limin Cui

In this paper infinite horizon optimal control problems for nonlinear high-dimensional dynamical systems are studied. Nonlinear feedback laws can be computed via the value function characterized as the unique viscosity solution to the…

最优化与控制 · 数学 2016-02-22 Alessandro Alla , Maurizio Falcone , Stefan Volkwein

Off-policy Reinforcement Learning (RL) holds the promise of better data efficiency as it allows sample reuse and potentially enables safe interaction with the environment. Current off-policy policy gradient methods either suffer from high…

机器学习 · 计算机科学 2021-06-09 Samuele Tosatto , João Carvalho , Jan Peters

Stabilizing an unknown control system is one of the most fundamental problems in control systems engineering. In this paper, we provide a simple, model-free algorithm for stabilizing fully observed dynamical systems. While model-free…

系统与控制 · 电气工程与系统科学 2021-10-14 Juan C. Perdomo , Jack Umenberger , Max Simchowitz

In this paper we introduce an iterative Jacobi algorithm for solving distributed model predictive control (DMPC) problems, with linear coupled dynamics and convex coupled constraints. The algorithm guarantees stability and persistent…

最优化与控制 · 数学 2008-09-23 Dang Doan , Tamas Keviczky , Ion Necoara , Moritz Diehl

This paper introduces a novel methodology that leverages the Hamilton-Jacobi solution to enhance non-linear model predictive control (MPC) in scenarios affected by navigational uncertainty. Using Hamilton-Jacobi-Theoretic approach, a…

最优化与控制 · 数学 2025-04-01 Amit Jain , Roshan T. Eapen , Puneet Singla

Model-reference adaptive systems refer to a consortium of techniques that guide plants to track desired reference trajectories. Approaches based on theories like Lyapunov, sliding surfaces, and backstepping are typically employed to advise…

系统与控制 · 电气工程与系统科学 2023-03-20 Mohammed Abouheaf , Wail Gueaieb , Davide Spinello , Salah Al-Sharhan

This paper proposes two cooperative optimal output tracking (COOT) algorithms based on policy iteration (PI) for discrete-time multi-agent systems with unknown model parameters. First, we establish a stabilizing PI framework that can start…

系统与控制 · 电气工程与系统科学 2026-01-27 Dongdong Li , Jiuxiang Dong

Nonlinear optimal control is vital for numerous applications but remains challenging for unknown systems due to the difficulties in accurately modelling dynamics and handling computational demands, particularly in high-dimensional settings.…

系统与控制 · 电气工程与系统科学 2024-12-03 Zhexuan Zeng , Ruikun Zhou , Yiming Meng , Jun Liu

The ergodic control problem for a non-degenerate controlled diffusion controlled through its drift is considered under a uniform stability condition that ensures the well-posedness of the associated Hamilton-Jacobi-Bellman (HJB) equation. A…

最优化与控制 · 数学 2019-03-20 Ari Arapostathis , Vivek S. Borkar

Reinforcement learning (RL) algorithms still suffer from high sample complexity despite outstanding recent successes. The need for intensive interactions with the environment is especially observed in many widely popular policy gradient…

机器学习 · 计算机科学 2020-08-04 Samuele Tosatto , Joao Carvalho , Hany Abdulsamad , Jan Peters

We present a semi-real-time algorithm for minimal-time optimal path planning based on optimal control theory, dynamic programming, and Hamilton-Jacobi (HJ) equations. Partial differential equation (PDE) based optimal path planning methods…

最优化与控制 · 数学 2023-09-06 Christian Parkinson , Kyle Polage