中文
相关论文

相关论文: Reinforcement Solver for H-infinity Filter with Bo…

200 篇论文

The uncertainties in plant dynamics remain a challenge for nonlinear control problems. This paper develops a ternary policy iteration (TPI) algorithm for solving nonlinear robust control problems with bounded uncertainties. The controller…

系统与控制 · 电气工程与系统科学 2020-07-15 Jie Li , Shengbo Eben Li , Yang Guan , Jingliang Duan , Wenyu Li , Yuming Yin

This paper studies the mixed $H_-/H_{\infty}$ fault detection filtering of It\^o-type nonlinear stochastic systems. Mixed $H_-/H_{\infty}$ filtering combines the system robustness to the external disturbance and the sensitivity to the fault…

最优化与控制 · 数学 2018-12-21 Tianliang Zhang , Feiqi Deng , Weihai Zhang , Bor-Sen Chen

Despite its popularity in the reinforcement learning community, a provably convergent policy gradient method for continuous space-time control problems with nonlinear state dynamics has been elusive. This paper proposes proximal gradient…

最优化与控制 · 数学 2022-12-27 Christoph Reisinger , Wolfgang Stockinger , Yufei Zhang

This paper investigates the H2 and H-infinity suboptimal distributed filtering problems for continuous time linear systems. Consider a linear system monitored by a number of filters, where each of the filters receives only part of the…

最优化与控制 · 数学 2020-02-10 Junjie Jiao , Harry L. Trentelman , M. Kanat Camlibel

This paper introduces a novel framework integrating nonlinear acoustic computing and reinforcement learning to enhance advanced human-robot interaction under complex noise and reverberation. Leveraging physically informed wave equations…

机器人学 · 计算机科学 2025-05-07 Xiaoliang Chen , Xin Yu , Le Chang , Yunhe Huang , Jiashuai He , Shibo Zhang , Jin Li , Likai Lin , Ziyu Zeng , Xianling Tu , Shuyu Zhang

In this paper, we present an algorithm for identifying a parametrically described destructive unknown system based on a non-gaussianity measure. It is known that under certain conditions the output of a linear system is more gaussian than…

计算机视觉与模式识别 · 计算机科学 2013-09-20 Deborah Pereg , Doron Benzvi

In this paper, we propose two algorithms for solving linear inverse problems when the observations are corrupted by noise. A proper data fidelity term (log-likelihood) is introduced to reflect the statistics of the noise (e.g. Gaussian,…

应用统计 · 统计学 2011-03-14 François-Xavier Dupé , Jalal Fadili , Jean-Luc Starck

This paper is on learning the Kalman gain by policy optimization method. Firstly, we reformulate the finite-horizon Kalman filter as a policy optimization problem of the dual system. Secondly, we obtain the global linear convergence of…

最优化与控制 · 数学 2023-10-30 Haoran Li , Yuan-Hua Ni

Convex Q-learning is a recent approach to reinforcement learning, motivated by the possibility of a firmer theory for convergence, and the possibility of making use of greater a priori knowledge regarding policy or value function structure.…

最优化与控制 · 数学 2022-10-18 Fan Lu , Joel Mathias , Sean Meyn , Karanjit Kalsi

An algorithm based on the interior-point methodology for solving continuous nonlinearly constrained optimization problems is proposed, analyzed, and tested. The distinguishing feature of the algorithm is that it presumes that only noisy…

最优化与控制 · 数学 2025-02-18 Frank E. Curtis , Shima Dezfulian , Andreas Waechter

This paper develops an algorithm for upper- and lower-bounding the value function for a class of linear time-varying games subject to convex control sets. In particular, a two-player zero-sum differential game is considered where the…

最优化与控制 · 数学 2025-03-12 Vincent Liu , Chris Manzie , Peter M. Dower

This paper derives recursion equations for a robust smoothing problem for a class of nonlinear systems with uncertainties in modeling and exogenous noise sources. The systems considered operate in discrete-time and the uncertainties are…

最优化与控制 · 数学 2013-03-27 Abhijit G. Kallapur , Ian R. Petersen

This paper investigates the problem of controlling a linear system under possibly unbounded stochastic noise with unknown convex cost functions, known as an online control problem. In contrast to the existing work, which assumes the…

系统与控制 · 电气工程与系统科学 2025-06-03 Kaito Ito , Taira Tsuchiya

Policy iteration is a widely used technique to solve the Hamilton Jacobi Bellman (HJB) equation, which arises from nonlinear optimal feedback control theory. Its convergence analysis has attracted much attention in the unconstrained case.…

最优化与控制 · 数学 2020-05-19 Sudeep Kundu , Karl Kunisch

In this paper, we propose an interior-point method for linearly constrained optimization problems (possibly nonconvex). The method - which we call the Hessian barrier algorithm (HBA) - combines a forward Euler discretization of Hessian…

最优化与控制 · 数学 2023-09-14 Immanuel M. Bomze , Panayotis Mertikopoulos , Werner Schachinger , Mathias Staudigl

Systems equipped with modern sensing modalities such as vision and lidar gain access to increasingly high-dimensional measurements with which to enact estimation and control schemes. In this article, we examine the continuum limit of…

系统与控制 · 电气工程与系统科学 2024-09-20 Maxwell Varley , Timothy L. Molloy , Girish N. Nair

Today, the optimal performance of existing noise-suppression algorithms, both data-driven and those based on classic statistical methods, is range bound to specific levels of instantaneous input signal-to-noise ratios. In this paper, we…

机器学习 · 计算机科学 2018-07-30 Rasool Fakoor , Xiaodong He , Ivan Tashev , Shuayb Zarar

Robust environment perception is essential for decision-making on robots operating in complex domains. Principled treatment of uncertainty sources in a robot's observation model is necessary for accurate mapping and object detection. This…

计算机视觉与模式识别 · 计算机科学 2016-07-15 Shayegan Omidshafiei , Brett T. Lopez , Jonathan P. How , John Vian

Quantum metrology promises precision beyond classical limits, yet environmental noise typically degrades the quantum resources required for such enhancement. In this work, we investigate frequency estimation in noisy continuous-variable…

量子物理 · 物理学 2026-05-08 Ayan Patra , Manju , Aditi Sen De , Matteo G. A. Paris

Practical application of H[infinity] robust control relies on system identification of a valid model-set, described by a linear system in feedback with a stable norm-bounded uncertainty, which must explains all possible (or at least all…

最优化与控制 · 数学 2019-01-07 Gray C. Thomas , Luis Sentis