English
Related papers

Related papers: Online Natural Gradient as a Kalman Filter

200 papers

Backpropagation dominates modern machine learning, yet it is not the only principled method for optimizing dynamical systems. We propose Kalman World Models (KWM), a class of learned state-space models trained via recursive Bayesian…

Machine Learning · Computer Science 2026-03-17 Andrew Kiruluta

In this paper, we revisit the Kalman filter theory. After giving the intuition on a simplified financial markets example, we revisit the maths underlying it. We then show that Kalman filter can be presented in a very different fashion using…

Statistical Finance · Quantitative Finance 2018-12-14 Eric Benhamou

A Kalman filter can be used to determine material parameters using uncertain experimental data. However, starting with inappropriate initial values for material parameters might include false local attractors or even divergence. Also,…

Materials Science · Physics 2015-02-13 Abdallah Shokry , Per Ståhle

In this work, we show that natural policy gradient, a core algorithm in reinforcement learning, admits an exact formulation as a smoothed and averaged form of policy iteration. Specifically, we introduce doubly smoothed policy iteration…

Machine Learning · Computer Science 2026-05-12 Phalguni Nanda , Zaiwei Chen

Simultaneous state and parameter estimation arises from various applicational areas but presents a major computational challenge. Most available Markov chain or sequential Monte Carlo techniques are applicable to relatively low dimensional…

Numerical Analysis · Mathematics 2017-09-28 Angwenyi David , Jana de Wiljes , Sebastian Reich

In optimization, the natural gradient method is well-known for likelihood maximization. The method uses the Kullback-Leibler divergence, corresponding infinitesimally to the Fisher-Rao metric, which is pulled back to the parameter space of…

Machine Learning · Statistics 2019-02-26 Anton Mallasto , Tom Dela Haije , Aasa Feragen

Bayesian filtering is a cornerstone of state estimation in complex systems such as aerospace systems, yet exact solutions are available only for linear Gaussian models. In practice,nonlinear systems are handled through tractable…

In this paper we propose a recursive online algorithm for estimating the parameters of a time-varying ARCH process. The estimation is done by updating the estimator at time point $t-1$ with observations about the time point $t$ to yield an…

Statistics Theory · Mathematics 2009-09-29 Rainer Dahlhaus , Suhasini Subba Rao

Reinforcement learning (RL) tackles sequential decision-making problems by creating agents that interacts with their environment. However, existing algorithms often view these problem as static, focusing on point estimates for model…

Machine Learning · Statistics 2024-03-21 Frank Shih , Faming Liang

This paper deals with estimating model parameters in graphical models. We reformulate it as an information geometric optimization problem and introduce a natural gradient descent strategy that incorporates additional meta parameters. We…

Machine Learning · Computer Science 2019-05-15 Eric Benhamou , Jamal Atif , Rida Laraki , David Saltiel

We study the natural gradient method for learning in deep Bayesian networks, including neural networks. There are two natural geometries associated with such learning systems consisting of visible and hidden units. One geometry is related…

Machine Learning · Computer Science 2020-05-22 Nihat Ay

The possible methodologies to handle the uncertain parameter are reviewed. The core idea of the desensitized Kalman filter is introduced. A new cost function consisting of a posterior covariance trace and trace of a weighted norm of the…

Information Theory · Computer Science 2015-04-21 Taishan Lou

Deep learning has the potential to dramatically impact navigation and tracking state estimation problems critical to autonomous vehicles and robotics. Measurement uncertainties in state estimation systems based on Kalman and other Bayes…

Machine Learning · Computer Science 2021-06-16 Rebecca L. Russell , Christopher Reale

This work describes a family of attitude estimators that are based on a generalization of Mahony's nonlinear complementary filter. This generalization reveals the close mathematical relationship between the nonlinear complementary filter…

Optimization and Control · Mathematics 2011-10-04 Kenneth Jensen

This paper introduces a novel proprioceptive state estimator for legged robots that combines model-based filters and deep neural networks. Recent studies have shown that neural networks such as multi-layer perceptron or recurrent neural…

Robotics · Computer Science 2024-10-28 Donghoon Youm , Hyunsik Oh , Suyoung Choi , Hyeongjun Kim , Jemin Hwangbo

Structural credit assignment for recurrent learning is challenging. An algorithm called RTRL can compute gradients for recurrent networks online but is computationally intractable for large networks. Alternatives, such as BPTT, are not…

Machine Learning · Computer Science 2021-03-11 Khurram Javed , Martha White , Rich Sutton

This report provides a brief historical evolution of the concepts in the Kalman filtering theory since ancient times to the present. A brief description of the filter equations its aesthetics, beauty, truth, fascinating perspectives and…

Methodology · Statistics 2015-03-17 Shyam Mohan M , Naren Naik , R. M. O. Gemson , M. R. Ananthasayanam

Natural language data, such as text and speech, have become readily available through social networking services and chat platforms. By leveraging human observations expressed in natural language, this paper addresses the problem of state…

Systems and Control · Electrical Eng. & Systems 2025-11-17 Yuki Miyoshi , Masaki Inoue , Yusuke Fujimoto

Inherent in virtually every iterative machine learning algorithm is the problem of hyper-parameter tuning, which includes three major design parameters: (a) the complexity of the model, e.g., the number of neurons in a neural network, (b)…

Machine Learning · Computer Science 2025-09-26 Christos Mavridis , John Baras

Motivated by the maneuvering target tracking with sensors such as radar and sonar, this paper considers the joint and recursive estimation of the dynamic state and the time-varying process noise covariance in nonlinear state space models.…

Systems and Control · Electrical Eng. & Systems 2023-05-09 Hua Lan , Jinjie Hu , Zengfu Wang , Qiang Cheng