中文
相关论文

相关论文: Learning the Kalman Filter with Fine-Grained Sampl…

200 篇论文

We explore reinforcement learning methods for finding the optimal policy in the linear quadratic regulator (LQR) problem. In particular, we consider the convergence of policy gradient methods in the setting of known and unknown parameters.…

机器学习 · 计算机科学 2021-06-25 Ben Hambly , Renyuan Xu , Huining Yang

In this paper we consider the problem of learning an $\epsilon$-optimal policy for a discounted Markov Decision Process (MDP). Given an MDP with $S$ states, $A$ actions, the discount factor $\gamma \in (0,1)$, and an approximation threshold…

机器学习 · 计算机科学 2020-12-25 Zihan Zhang , Yuan Zhou , Xiangyang Ji

Kalman Filters (KF) are fundamental to real-time state estimation applications, including radar-based tracking systems used in modern driver assistance and safety technologies. In a linear dynamical system with Gaussian noise distributions…

机器人学 · 计算机科学 2024-11-27 Arian Mehrfard , Bharanidhar Duraisamy , Stefan Haag , Florian Geiss

In this paper, we propose a probabilistic optimization method, named probabilistic incremental proximal gradient (PIPG) method, by developing a probabilistic interpretation of the incremental proximal gradient algorithm. We explicitly model…

最优化与控制 · 数学 2019-06-20 Ömer Deniz Akyildiz , Émilie Chouzenoux , Víctor Elvira , Joaquín Míguez

In this paper, we present a UKF-PF based hybrid nonlinear filter for space object tracking. Estimating the state and its associated uncertainty, also known as filtering is paramount to the tracking process. The periodicity of the Keplerian…

动力系统 · 数学 2014-09-30 Dilshad Raihan A. V. , Suman Chakravorty

Motivated by filtering tasks under a linear system with non-Gaussian heavy-tailed noise, various robust Kalman filters (RKFs) based on different heavy-tailed distributions have been proposed. Although the sub-Gaussian $\alpha$-stable…

信号处理 · 电气工程与系统科学 2023-12-29 Pengcheng Hao , Oktay Karakuş , Alin Achim

In this paper, we propose and develop a methodology for nonlinear systems health monitoring by modeling the damage and degradation mechanism dynamics as "slow" states that are augmented with the system "fast" dynamical states. This…

系统与控制 · 计算机科学 2017-10-17 Najmeh Daroogheh , Nader Meskin , Khashayar Khorasani

This article introduces a new algorithm for nonlinear state estimation based on deterministic sigma point and EKF linearized framework for priori mean and covariance respectively. This method reduces the computation cost of UKF about 50%…

系统与控制 · 电气工程与系统科学 2019-07-25 Milad Behvandi , Mohammad Azam Khosravi , Amir Abolfazl Suratgar

The Robust Markov Decision Process (RMDP) framework focuses on designing control policies that are robust against the parameter uncertainties due to the mismatches between the simulator model and real-world settings. An RMDP problem is…

机器学习 · 计算机科学 2022-05-17 Kishan Panaganti , Dileep Kalathil

We present the first finite-sample analysis of policy evaluation in robust average-reward Markov Decision Processes (MDPs). Prior work in this setting have established only asymptotic convergence guarantees, leaving open the question of…

机器学习 · 统计学 2025-12-11 Yang Xu , Washim Uddin Mondal , Vaneet Aggarwal

Model predictive control has shown potential to enhance the robustness of quantum control systems. In this work, we propose a tractable Stochastic Model Predictive Control (SMPC) framework for finite-dimensional quantum systems under…

量子物理 · 物理学 2025-12-04 Yunyan Lee , Ian R. Petersen , Daoyi Dong

In this paper, we propose a new model reduction technique for linear stochastic systems that builds upon knowledge filtering and utilizes optimal Kalman filtering techniques. This new technique will reduce the dimension of the noise…

系统与控制 · 电气工程与系统科学 2023-09-18 Maico Hendrikus Wilhelmus Engelaar , Licio Romao , Yulong Gao , Mircea Lazar , Alessandro Abate , Sofie Haesaert

State estimation in control and systems engineering traditionally requires extensive manual system identification or data-collection effort. However, transformer-based foundation models in other domains have reduced data requirements by…

系统与控制 · 电气工程与系统科学 2025-09-05 Tobin Holtmann , David Stenger , Andres Posada-Moreno , Friedrich Solowjow , Sebastian Trimpe

Direct policy gradient methods for reinforcement learning are a successful approach for a variety of reasons: they are model free, they directly optimize the performance metric of interest, and they allow for richly parameterized policies.…

机器学习 · 计算机科学 2020-08-14 Alekh Agarwal , Mikael Henaff , Sham Kakade , Wen Sun

Reinforcement learning (RL) algorithms still suffer from high sample complexity despite outstanding recent successes. The need for intensive interactions with the environment is especially observed in many widely popular policy gradient…

机器学习 · 计算机科学 2020-08-04 Samuele Tosatto , Joao Carvalho , Hany Abdulsamad , Jan Peters

Optimal decision-making under partial observability requires reasoning about the uncertainty of the environment's hidden state. However, most reinforcement learning architectures handle partial observability with sequence models that have…

机器学习 · 计算机科学 2025-02-20 Carlos E. Luis , Alessandro G. Bottero , Julia Vinogradska , Felix Berkenkamp , Jan Peters

The infinite horizon setting is widely adopted for problems of reinforcement learning (RL). These invariably result in stationary policies that are optimal. In many situations, finite horizon control problems are of interest and for such…

机器学习 · 计算机科学 2025-03-21 Soumyajit Guin , Shalabh Bhatnagar

In many physical applications, the system's state varies with spatial variables as well as time. The state of such systems is modelled by partial differential equations and evolves on an infinite-dimensional space. Systems modelled by…

最优化与控制 · 数学 2022-02-17 Sepideh Afshar , Fabian Germ , Kirsten A. Morris

Accurately reconstructing and forecasting high-resolution (HR) states from computationally cheap low-resolution (LR) observations is central to estimation-and-control of spatio-temporal PDE systems. We develop a unified superresolution…

流体动力学 · 物理学 2025-09-16 Mrigank Dhingra , Omer San

We study the sequential decision making problem of maximizing the expected total reward while satisfying a constraint on the expected total utility. We employ the natural policy gradient method to solve the discounted infinite-horizon…

最优化与控制 · 数学 2025-10-16 Dongsheng Ding , Kaiqing Zhang , Jiali Duan , Tamer Başar , Mihailo R. Jovanović