English
Related papers

Related papers: Duality-Based Stochastic Policy Optimization for E…

200 papers

Motivated by the maneuvering target tracking with sensors such as radar and sonar, this paper considers the joint and recursive estimation of the dynamic state and the time-varying process noise covariance in nonlinear state space models.…

Systems and Control · Electrical Eng. & Systems 2023-05-09 Hua Lan , Jinjie Hu , Zengfu Wang , Qiang Cheng

The theory of dual control was introduced more than seven decades ago. Although it has provided rich insights to the fields of control, estimation, and system identification, dual control is generally computationally prohibitive. In recent…

Systems and Control · Electrical Eng. & Systems 2024-05-06 Mohammad S. Ramadan , Mihai Anitescu

In this paper, we study distributed estimation and control problems over graphs under partially nested information patterns. We show a duality result that is very similar to the classical duality result between state estimation and state…

Optimization and Control · Mathematics 2012-09-17 Ather Gattami , Sanjoy Mitter

Domain randomization is a simple, effective, and flexible scheme for obtaining robust feedback policies aimed at reducing the sim-to-real gap due to model mismatch. While domain randomization methods have yielded impressive demonstrations…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Alex Nguyen-Le , Nikolai Matni

This work examines the optimal covariance steering problem for systems subject to unknown parameters that enter multiplicatively with the state and control, in addition to additive disturbances. In contrast to existing works, the unknown…

Systems and Control · Electrical Eng. & Systems 2024-03-26 Jacob W. Knaup , Panagiotis Tsiotras

Optimal control of stochastic nonlinear dynamical systems is a major challenge in the domain of robot learning. Given the intractability of the global control problem, state-of-the-art algorithms focus on approximate sequential optimization…

Machine Learning · Computer Science 2020-04-23 Joe Watson , Hany Abdulsamad , Jan Peters

Reliable state estimation hinges on accurate specification of sensor noise covariances, which weigh heterogeneous measurements. In practice, these covariances are difficult to identify due to environmental variability, front-end…

Robotics · Computer Science 2025-12-18 Haoying Li , Yifan Peng , Xinghan Li , Junfeng Wu

System identification poses a significant bottleneck to characterizing and controlling complex systems. This challenge is greatest when both the system states and parameters are not directly accessible leading to a dual-estimation problem.…

Systems and Control · Electrical Eng. & Systems 2021-04-08 Matthew F. Singh , Chong Wang , Michael W. Cole , ShiNung Ching

To address the communication bottleneck challenge in distributed learning, our work introduces a novel two-stage quantization strategy designed to enhance the communication efficiency of distributed Stochastic Gradient Descent (SGD). The…

Machine Learning · Computer Science 2024-02-05 Guangfeng Yan , Tan Li , Yuanzhang Xiao , Congduan Li , Linqi Song

The extended Kalman filter is perhaps the most standard tool to estimate in real time the state of a dynamical system from noisy measurements of some function of the system, with extensive practical applications (such as position tracking…

Optimization and Control · Mathematics 2019-01-04 Yann Ollivier

A dual control problem is presented for the optimal stochastic control of a system governed by partial differential equations. Relationships between the optimal values of the original and the dual problems are investigated and two duality…

Optimization and Control · Mathematics 2017-05-03 Shinji Tanimoto

The Kalman filter is a fundamental filtering algorithm that fuses noisy sensory data, a previous state estimate, and a dynamics model to produce a principled estimate of the current state. It assumes, and is optimal for, linear models and…

Neural and Evolutionary Computing · Computer Science 2021-04-30 Beren Millidge , Alexander Tschantz , Anil Seth , Christopher Buckley

The analysis of high-dimensional dynamical systems generally requires the integration of simulation data with experimental measurements. Experimental data often has substantial amounts of measurement noise that compromises the ability to…

Numerical Analysis · Mathematics 2019-10-02 Samuel Rudy , Steven Brunton , J. Nathan Kutz

We explore reinforcement learning methods for finding the optimal policy in the linear quadratic regulator (LQR) problem. In particular, we consider the convergence of policy gradient methods in the setting of known and unknown parameters.…

Machine Learning · Computer Science 2021-06-25 Ben Hambly , Renyuan Xu , Huining Yang

We derive a decomposition for the gradient of the innovation loss with respect to the filter gain in a linear time-invariant system, decomposing as a product of an observability Gramian and a term quantifying the ``non-orthogonality"…

Optimization and Control · Mathematics 2025-07-23 M. A. Belabbas , A. Olshevsky

We systematically develop a learning-based treatment of stochastic optimal control (SOC), relying on direct optimization of parametric control policies. We propose a derivation of adjoint sensitivity results for stochastic differential…

Machine Learning · Computer Science 2021-06-08 Stefano Massaroli , Michael Poli , Stefano Peluchetti , Jinkyoo Park , Atsushi Yamashita , Hajime Asama

We present a unified framework for learning continuous control policies using backpropagation. It supports stochastic control by treating stochasticity in the Bellman equation as a deterministic function of exogenous noise. The product is a…

Machine Learning · Computer Science 2015-11-02 Nicolas Heess , Greg Wayne , David Silver , Timothy Lillicrap , Yuval Tassa , Tom Erez

This paper establishes a rigorous connection between regularized discrete-time reinforcement learning (RL) and continuous-time stochastic optimal control. Specifically, classical RL algorithms are typically solving a regularized…

Optimization and Control · Mathematics 2026-04-24 Huyên Pham , Yuming Paul Zhang , Yuhua Zhu

Estimating parameters of a diffusion process given continuous-time observations of the process via maximum likelihood approaches or, online, via stochastic gradient descent or Kalman filter formulations constitutes a well-established…

Methodology · Statistics 2025-03-17 Jan Albrecht , Sebastian Reich

This paper describes a minimax state estimation approach for linear Differential-Algebraic Equations (DAE) with uncertain parameters. The approach addresses continuous-time DAE with non-stationary rectangular matrices and uncertain bounded…

Optimization and Control · Mathematics 2011-02-28 Sergiy Zhuk