中文
相关论文

相关论文: Passivity and Immersion based-modified gradient es…

200 篇论文

This paper presents a tutorial overview of path integral (PI) control approaches for stochastic optimal control and trajectory optimization. We concisely summarize the theoretical development of path integral control to compute a solution…

机器人学 · 计算机科学 2023-12-05 Muhammad Kazim , JunGee Hong , Min-Gyeom Kim , Kwang-Ki K. Kim

Recovery type a posteriori error estimators are popular, particularly in the engineering community, for their computationally inexpensive, easy to implement, and generally asymptotically exactness. Unlike the residual type error estimators,…

数值分析 · 数学 2025-03-26 Ying Liu , Jingjing Xiao , Nianyu Yi , Huihui Cao

This paper presents extensions of finite-time stability results to some prototypical adaptive control and estimation frameworks. First, we present a novel scheme of online parameter estimation that guarantees convergence of the estimation…

最优化与控制 · 数学 2020-10-20 Kunal Garg , Parag Bobade , Dimitra Panagou

Differential-algebraic equations (DAEs) with state-dependent events arise in systems whose continuous dynamics are constrained by algebraic equations and interrupted by mode changes, switching logic, impacts, or state reinitializations.…

机器学习 · 计算机科学 2026-05-08 Ion Matei , Maksym Zhenirovskyy , Anthony Wong

Policy gradient (PG) methods are the backbone of many reinforcement learning algorithms due to their good performance in policy optimization problems. As a gradient-based approach, PG methods typically rely on knowledge of the system…

系统与控制 · 电气工程与系统科学 2026-04-02 Bowen Song , Andrea Iannelli

This letter proposes a new method for joint state and parameter estimation in uncertain dynamical systems. We exploit the partial errors-in-variables (PEIV) principle and formulate a regression problem in the sense of weighted total least…

信号处理 · 电气工程与系统科学 2024-07-03 Peng Liu , Kailai Li , Gustaf Hendeby , Fredrik Gustafsson

In many scientific fields, the generation and evolution of data are governed by partial differential equations (PDEs) which are typically informed by established physical laws at the macroscopic level to describe general and predictable…

统计方法学 · 统计学 2025-07-01 Ziyuan Chen , Shunxing Yan , Fang Yao

State-dependent parameter identification, where unknown model parameters depend on one or more state variables in partial differential equations (PDEs) or coupled PDE systems, is fundamental to a wide range of problems in physics,…

最优化与控制 · 数学 2026-01-19 Vladislav Bukshtynov

We propose EAGLE update rule, a novel optimization method that accelerates loss convergence during the early stages of training by leveraging both current and previous step parameter and gradient values. The update algorithm estimates…

机器学习 · 计算机科学 2025-02-04 Takumi Fujimoto , Hiroaki Nishi

We consider the problem of simultaneous control and parameter estimation when the model is available only as a differentiable physics simulator. We propose a receding-horizon control framework in which a model predictive control (MPC)…

最优化与控制 · 数学 2026-04-07 Alan Williams , Alp Sunol

This paper presents a new parameter estimation algorithm for the adaptive control of a class of time-varying plants. The main feature of this algorithm is a matrix of time-varying learning rates, which enables parameter estimation error…

最优化与控制 · 数学 2021-11-18 Joseph E. Gaudio , Anuradha M. Annaswamy , Eugene Lavretsky , Michael A. Bolender

In this paper, we introduce a policy-gradient method for model-based reinforcement learning (RL) that exploits a type of stationary distributions commonly obtained from Markov decision processes (MDPs) in stochastic networks, queueing…

机器学习 · 计算机科学 2025-10-30 Céline Comte , Matthieu Jonckheere , Jaron Sanders , Albert Senen-Cerda

We introduce optimal energy shaping as an enhancement of classical passivity-based control methods. A promising feature of passivity theory, alongside stability, has traditionally been claimed to be intuitive performance tuning along the…

系统与控制 · 电气工程与系统科学 2021-01-15 Stefano Massaroli , Michael Poli , Federico Califano , Jinkyoo Park , Atsushi Yamashita , Hajime Asama

We introduce Perturbative Gradient Training (PGT), a novel training paradigm that overcomes a critical limitation of physical reservoir computing: the inability to perform backpropagation due to the black-box nature of physical reservoirs.…

机器学习 · 计算机科学 2025-06-06 Cliff B. Abbott , Mark Elo , Dmytro A. Bozhko

Recently, adaptive control systems with relaxed persistent excitation (PE) conditions have been proposed to guarantee true parameter convergence and improve the transient response. However, in some cases, sufficient control performance and…

系统与控制 · 电气工程与系统科学 2025-03-03 Satoshi Tsuruhara , Kazuhisa Ito

Scalable and effective exploration remains a key challenge in reinforcement learning (RL). While there are methods with optimality guarantees in the setting of discrete state and action spaces, these methods cannot be applied in…

机器学习 · 计算机科学 2017-01-30 Rein Houthooft , Xi Chen , Yan Duan , John Schulman , Filip De Turck , Pieter Abbeel

Parameter identification is crucial in virtual engineering processes, yet determining appropriate system excitations for identifying specific parameters remains challenging. In practice, extensive experimental programs often fail to…

最优化与控制 · 数学 2026-05-07 Kevin Schmidt , Nicola Henkelmann , Christoph Mark , Johannes von Keler

We advocate for a practical Maximum Likelihood Estimation (MLE) approach towards designing loss functions for regression and forecasting, as an alternative to the typical approach of direct empirical risk minimization on a specific target…

机器学习 · 统计学 2021-10-12 Pranjal Awasthi , Abhimanyu Das , Rajat Sen , Ananda Theertha Suresh

This paper investigates gradient-based adaptive prediction and control for nonlinear stochastic dynamical systems under a weak convexity condition on the prediction-based loss. This condition accommodates a broad range of nonlinear models…

系统与控制 · 电气工程与系统科学 2026-02-13 Yujing Liu , Xin Zheng , Zhixin Liu , Lei Guo

Model-based reinforcement learning (MBRL) is a sample efficient technique to obtain control policies, yet unavoidable modeling errors often lead performance deterioration. The model in MBRL is often solely fitted to reconstruct dynamics,…

机器学习 · 计算机科学 2023-06-22 Claas Voelcker , Victor Liao , Animesh Garg , Amir-massoud Farahmand