中文
相关论文

相关论文: Duality-Based Stochastic Policy Optimization for E…

200 篇论文

This paper introduces two new algorithms to accurately estimate the process noise covariance of a discrete-time Kalman filter online for robust orbit determination in the presence of dynamics model uncertainties. Common orbit determination…

信号处理 · 电气工程与系统科学 2021-05-17 Nathan Stacey , Simone D'Amico

We consider a joint sensor and controller design problem for linear Gaussian stochastic systems in which a weighted sum of quadratic control cost and the amount of information acquired by the sensor is minimized. This problem formulation is…

最优化与控制 · 数学 2015-03-09 Takashi Tanaka , Henrik Sandberg

We study the online estimation of the optimal policy of a Markov decision process (MDP). We propose a class of Stochastic Primal-Dual (SPD) methods which exploit the inherent minimax duality of Bellman equations. The SPD methods update a…

机器学习 · 统计学 2016-12-09 Yichen Chen , Mengdi Wang

Prediction error and maximum likelihood methods are powerful tools for identifying linear dynamical systems and, in particular, enable the joint estimation of model parameters and the Kalman filter used for state estimation. A key…

系统与控制 · 电气工程与系统科学 2026-04-21 Léo Simpson , Moritz Diehl

This paper focuses on the linear quadratic control (LQC) design of systems corrupted by both stochastic noise and bounded noise simultaneously. When only of these noises are considered, the LQC strategy leads to stochastic or robust…

最优化与控制 · 数学 2025-12-15 Xuehui Ma , Shiliang Zhang , Xiaohui Zhang , Jing Xin , Hector Garcia de Marina

In this work, we present a new model-free and off-policy reinforcement learning (RL) algorithm, that is capable of finding a near-optimal policy with state-action observations from arbitrary behavior policies. Our algorithm, called the…

最优化与控制 · 数学 2025-07-21 Narim Jeong , Donghwan Lee , Niao He

This article explores the estimation of parameters and states for linear stochastic systems with deterministic control inputs. It introduces a novel Kalman filtering approach called Kalman Filtering with Correlated Noises Recursive…

系统与控制 · 电气工程与系统科学 2025-07-11 Abd El Mageed Hag Elamin Khalid

In this paper, we propose a non-parametric method for state estimation of high-dimensional nonlinear stochastic dynamical systems, which evolve according to gradient flows with isotropic diffusion. We combine diffusion maps, a manifold…

信号处理 · 电气工程与系统科学 2019-02-26 Tal Shnitzer , Ronen Talmon , Jean-Jacques Slotine

In this paper, we propose an approach to address the problems with ambiguity in tuning the process and observation noises for a discrete-time linear Kalman filter. Conventional approaches to tuning (e.g. using normalized estimation error…

系统与控制 · 电气工程与系统科学 2021-08-25 Zhaozhong Chen , Christoffer Heckman , Simon Julier , Nisar Ahmed

Guided policy search algorithms have been proven to work with incredible accuracy for not only controlling a complicated dynamical system, but also learning optimal policies from various unseen instances. One assumes true nature of the…

系统与控制 · 电气工程与系统科学 2020-10-02 Prakash Mallick , Zhiyong Chen , Mohsen Zamani

This paper investigates the distributed Kalman filter (DKF) for linear systems, with specific attention on measurement fusion, which is a typical way of information sharing and is vital for enhancing stability and improving estimation…

信号处理 · 电气工程与系统科学 2025-04-14 Tuo Yang , Jiachen Qian , Zhisheng Duan , Zhiyong Sun

Dual control explicitly addresses the problem of trading off active exploration and exploitation in the optimal control of partially unknown systems. While the problem can be cast in the framework of stochastic dynamic programming, exact…

系统与控制 · 电气工程与系统科学 2019-11-12 Elena Arcari , Lukas Hewing , Melanie N. Zeilinger

We propose a method for finding approximate compilations of quantum unitary transformations, based on techniques from policy gradient reinforcement learning. The choice of a stochastic policy allows us to rephrase the optimization problem…

量子物理 · 物理学 2022-09-14 David A. Herrera-Martí

A new formulation of Stochastic Model Predictive Output Feedback Control is presented and analyzed as a translation of Stochastic Optimal Output Feedback Control into a receding horizon setting. This requires lifting the design into a…

最优化与控制 · 数学 2020-05-01 Martin A Sehr , Robert R Bitmead

Dual control denotes a class of control problems where the parameters governing the system are imperfectly known. The challenge is to find the optimal balance between probing, i.e. exciting the system to understand it more, and caution,…

最优化与控制 · 数学 2020-04-29 Martin Péron , Christopher M. Baker , Barry D. Hughes , Iadine Chadès

We study generalization properties of random features (RF) regression in high dimensions optimized by stochastic gradient descent (SGD) in under-/over-parameterized regime. In this work, we derive precise non-asymptotic error bounds of RF…

机器学习 · 统计学 2022-10-18 Fanghui Liu , Johan A. K. Suykens , Volkan Cevher

This paper considers the problem of data-driven robust control design for nonlinear systems, for instance, obtained when discretizing nonlinear partial differential equations (PDEs). A robust learning control approach is developed for…

最优化与控制 · 数学 2025-09-01 Anant A. Joshi , Saviz Mowlavi , Mouhacine Benosman

Many dynamical systems are subjected to stochastic influences, such as random excitations, noise, and unmodeled behavior. Tracking the system's state and parameters based on a physical model is a common task for which filtering algorithms,…

信号处理 · 电气工程与系统科学 2024-07-03 Jan Grashorn , Matteo Broggi , Ludovic Chamoin , Michael Beer

The optimal predictor for a linear dynamical system (with hidden state and Gaussian noise) takes the form of an autoregressive linear filter, namely the Kalman filter. However, a fundamental problem in reinforcement learning and control…

机器学习 · 计算机科学 2019-05-27 Holden Lee , Cyril Zhang

In reinforcement learning (RL), offline learning decoupled learning from data collection and is useful in dealing with exploration-exploitation tradeoff and enables data reuse in many applications. In this work, we study two offline…

机器学习 · 计算机科学 2022-02-08 Jing Dong , Xin T. Tong