中文
相关论文

相关论文: Reinforcement Solver for H-infinity Filter with Bo…

200 篇论文

A fundamental question in reinforcement learning theory is: suppose the optimal value functions are linear in given features, can we learn them efficiently? This problem's counterpart in supervised learning, linear regression, can be solved…

机器学习 · 计算机科学 2023-02-28 Daniel Kane , Sihan Liu , Shachar Lovett , Gaurav Mahajan , Csaba Szepesvári , Gellért Weisz

In recent times, a variety of Reinforcement Learning (RL) algorithms have been proposed for optimal tracking problem of continuous time nonlinear systems with input constraints. Most of these algorithms are based on the notion of uniform…

系统与控制 · 电气工程与系统科学 2020-06-16 Amardeep Mishra , Satadal Ghosh

A new approach for robust Hinfty filtering for a class of Lipschitz nonlinear systems with time-varying uncertainties both in the linear and nonlinear parts of the system is proposed in an LMI framework. The admissible Lipschitz constant of…

系统与控制 · 计算机科学 2014-03-04 Masoud Abbaszadeh , Horacio J. Marquez

Solving the Hamilton-Jacobi-Bellman equation is important in many domains including control, robotics and economics. Especially for continuous control, solving this differential equation and its extension the Hamilton-Jacobi-Isaacs…

机器人学 · 计算机科学 2021-10-06 Michael Lutter , Boris Belousov , Shie Mannor , Dieter Fox , Animesh Garg , Jan Peters

We derive a method to reconstruct Gaussian signals from linear measurements with Gaussian noise. This new algorithm is intended for applications in astrophysics and other sciences. The starting point of our considerations is the principle…

天体物理仪器与方法 · 物理学 2011-10-18 Niels Oppermann , Georg Robbers , Torsten A. Ensslin

A learning-based safety filter is developed for discrete-time linear time-invariant systems with unknown models subject to Gaussian noises with unknown covariance. Safety is characterized using polytopic constraints on the states and…

机器学习 · 计算机科学 2023-05-09 Farhad Farokhi , Alex S. Leong , Mohammad Zamani , Iman Shames

This paper presents a framework to solve constrained optimization problems in an accelerated manner based on High-Order Tuners (HT). Our approach is based on reformulating the original constrained problem as the unconstrained optimization…

最优化与控制 · 数学 2022-05-27 Anjali Parashar , Priyank Srivastava , Anuradha M. Annaswamy

In this paper, we present a novel method for computing the optimal feedback gain of the infinite-horizon Linear Quadratic Regulator (LQR) problem via an ordinary differential equation. We introduce a novel continuous-time Bellman error,…

系统与控制 · 电气工程与系统科学 2026-04-17 Armin Gießler , Albertus Johannes Malan , Sören Hohmann

A new stochastic control model for the long-run environmental management of rivers is mathematically and numerically analyzed, focusing on a modern sediment replenishment problem with unique nonsmooth and nonlinear properties. Rational…

最优化与控制 · 数学 2022-03-11 Hidekazu Yoshioka , Motoh Tsujimura

We apply deep reinforcement learning techniques to design high threshold decoders for the toric code under uncorrelated noise. By rewarding the agent only if the decoding procedure preserves the logical states of the toric code, and using…

量子物理 · 物理学 2020-03-09 Laia Domingo Colomer , Michalis Skotiniotis , Ramon Muñoz-Tapia

Bayesian estimation is a vital tool in robotics as it allows systems to update the robot state belief using incomplete information from noisy sensors. To render the state estimation problem tractable, many systems assume that the motion and…

机器人学 · 计算机科学 2025-01-13 Miguel Saavedra-Ruiz , Steven A. Parkison , Ria Arora , James Richard Forbes , Liam Paull

Stochastic control with both inherent random system noise and lack of knowledge on system parameters constitutes the core and fundamental topic in reinforcement learning (RL), especially under non-episodic situations where online learning…

系统与控制 · 电气工程与系统科学 2019-06-24 Xin Huang , Duan Li , Daniel Zhuoyu Long

State filtering is a key problem in many signal processing applications. From a series of noisy measurement, one would like to estimate the state of some dynamic system. Existing techniques usually adopt a Gaussian noise assumption which…

统计方法学 · 统计学 2016-12-16 Bin Liu

In this paper, we introduce Hamilton-Jacobi-Bellman (HJB) equations for Q-functions in continuous time optimal control problems with Lipschitz continuous controls. The standard Q-function used in reinforcement learning is shown to be the…

最优化与控制 · 数学 2020-05-05 Jeongho Kim , Insoon Yang

The correct specification of reward models is a well-known challenge in reinforcement learning. Hand-crafted reward functions often lead to inefficient or suboptimal policies and may not be aligned with user values. Reinforcement learning…

The boundary control problem is a non-convex optimization and control problem in many scientific domains, including fluid mechanics, structural engineering, and heat transfer optimization. The aim is to find the optimal values for the…

机器学习 · 计算机科学 2023-10-25 Zenin Easa Panthakkalakath , Juraj Kardoš , Olaf Schenk

A generalized dynamical robust nonlinear filtering framework is established for a class of Lipschitz differential algebraic systems, in which the nonlinearities appear both in the state and measured output equations. The system is assumed…

系统与控制 · 计算机科学 2014-02-25 Masoud Abbaszadeh

In the current era of quantum computing, robust and efficient tools are essential to bridge the gap between simulations and quantum hardware execution. In this work, we introduce a machine learning approach to characterize the noise…

Classically, the optimal control problem in the presence of an adversary is formulated as a two-player zero-sum differential game or an $H_\infty$ control problem. The solution to these problems can be obtained by solving the…

最优化与控制 · 数学 2022-04-26 Alexander Krolicki , Sarang Sutavani , Umesh Vaidya

We propose a sequential quadratic programming (SQP) algorithm for inequality constrained optimization that is robust to the presence of bounded noise in function and derivative evaluations. We cover the case where constraint evaluations…

最优化与控制 · 数学 2026-04-17 Figen Oztoprak , Richard Byrd