English
Related papers

Related papers: Learning Kalman Policy for Singular Unknown Covari…

200 papers

The sample covariance matrix of a random vector is a good estimate of the true covariance matrix if the sample size is much larger than the length of the vector. In high-dimensional problems, this condition is never met. As a result, in…

Data Analysis, Statistics and Probability · Physics 2024-11-12 Michael Tsyrulnikov , Arseniy Sotskiy

This paper considers the problem of data-driven robust control design for nonlinear systems, for instance, obtained when discretizing nonlinear partial differential equations (PDEs). A robust learning control approach is developed for…

Optimization and Control · Mathematics 2025-09-01 Anant A. Joshi , Saviz Mowlavi , Mouhacine Benosman

Policy evaluation is a key process in reinforcement learning. It assesses a given policy using estimation of the corresponding value function. When using a parameterized function to approximate the value, it is common to optimize the set of…

Machine Learning · Computer Science 2019-01-24 Shirli Di-Castro Shashua , Shie Mannor

A recently developed data-driven Kalman filter requires offline measurement of the process disturbance; a requirement that is often unmet for many practical applications. We propose a solution that parametrizes the Kalman filter exclusively…

Systems and Control · Electrical Eng. & Systems 2025-11-12 Mohamed Abdalmoaty , Roy S. Smith

Stochastic parameterizations are increasingly being used to represent the uncertainty associated with model errors in ensemble forecasting and data assimilation. One of the challenges associated with the use of these parameterizations is…

Computation · Statistics 2019-10-23 Guillermo Scheffler , Juan Ruiz , Manuel Pulido

In recent years, reinforcement learning (RL) systems with general goals beyond a cumulative sum of rewards have gained traction, such as in constrained problems, exploration, and acting upon prior experiences. In this paper, we consider…

Machine Learning · Computer Science 2020-07-07 Junyu Zhang , Alec Koppel , Amrit Singh Bedi , Csaba Szepesvari , Mengdi Wang

Reinforcement learning considers the problem of finding policies that maximize an expected cumulative reward in a Markov decision process with unknown transition probabilities. In this paper we consider the problem of finding optimal…

Machine Learning · Computer Science 2020-10-19 Santiago Paternain , Juan Andres Bazerque , Alejandro Ribeiro

Many robotic sensor estimation problems can characterized in terms of nonlinear measurement systems. These systems are contaminated with noise and may be underdetermined from a single observation. In order to get reliable estimation…

Systems and Control · Computer Science 2013-04-11 Greg Hager , Max Mintz

The Kalman filter is a fundamental tool for state estimation in dynamical systems. While originally developed for linear Gaussian settings, it has been extended to nonlinear problems through approaches such as the extended and unscented…

Optimization and Control · Mathematics 2025-09-10 Yuan Wu , Sicheng He

Stochastic gradient optimization is the dominant learning paradigm for a variety of scenarios, from classical supervised learning to modern self-supervised learning. We consider stochastic gradient algorithms for learning problems whose…

Machine Learning · Statistics 2025-08-29 Facheng Yu , Ronak Mehta , Alex Luedtke , Zaid Harchaoui

Second-order information -- such as curvature or data covariance -- is critical for optimisation, diagnostics, and robustness. However, in many modern settings, only the gradients are observable. We show that the gradients alone can reveal…

Machine Learning · Computer Science 2026-04-08 Arash Jamshidi , Katsiaryna Haitsiukevich , Kai Puolamäki

A new application of duality relations of stochastic processes is demonstrated. Although conventional usages of the duality relations need analytical solutions for the dual processes, we here employ numerical solutions of the dual processes…

Systems and Control · Computer Science 2015-10-14 Jun Ohkubo

This paper focus on investigating the distributed Riemannian stochastic optimization problem on the Stiefel manifold for multi-agent systems, where all the agents work collaboratively to optimize a function modeled by the average of their…

Optimization and Control · Mathematics 2025-01-17 Jishu Zhao , Xi Wang , Jinlong Lei

A computationally efficient method for online joint state inference and dynamical model learning is presented. The dynamical model combines an a priori known, physically derived, state-space model with a radial basis function expansion…

Systems and Control · Electrical Eng. & Systems 2021-07-12 Anton Kullberg , Isaac Skog , Gustaf Hendeby

It is difficult for humans to efficiently teach robots how to correctly perform a task. One intuitive solution is for the robot to iteratively learn the human's preferences from corrections, where the human improves the robot's current…

Robotics · Computer Science 2018-09-14 Dylan P. Losey , Marcia K. O'Malley

Kullback-Leibler divergence (KL) regularization is widely used in reinforcement learning, but it becomes infinite under support mismatch and can degenerate in low-noise limits. Utilizing a unified information-geometric framework, we…

Optimization and Control · Mathematics 2026-02-03 Viktor Stein , Adwait Datar , Nihat Ay

Kalman filter is a best linear unbiased state estimator. It is also comprehensible from the point view of the Bayesian estimation. However, this note gives a detailed derivation of Kalman filter from the mutual information perspective for…

Information Theory · Computer Science 2021-01-05 Yarong Luo , Jianlang Hu , Chi Guo

We consider the Ensemble Kalman Inversion which has been recently introduced as an efficient, gradient-free optimisation method to estimate unknown parameters in an inverse setting. In the case of large data sets, the Ensemble Kalman…

Numerical Analysis · Mathematics 2023-12-05 Matei Hanu , Jonas Latz , Claudia Schillings

This paper studies the optimal state estimation for a dynamic system, whose transfer function can be nonlinear and the input noise can be of arbitrary distribution. Our algorithm differs from the conventional extended Kalman filter (EKF)…

Signal Processing · Electrical Eng. & Systems 2022-04-22 Xin Liang , Yi Jiang

Reinforcement learning (RL) has become a central post-training paradigm for large language models (LLMs), but its performance is highly sensitive to the quality of training problems. This sensitivity stems from the non-stationarity of RL:…

Machine Learning · Computer Science 2026-02-26 Ningyuan Yang , Weihua Du , Weiwei Sun , Sean Welleck , Yiming Yang