中文
相关论文

相关论文: Learning the Kalman Filter with Fine-Grained Sampl…

200 篇论文

Using a perturbation technique, we derive a new approximate filtering and smoothing methodology generalizing along different directions several existing approaches to robust filtering based on the score and the Hessian matrix of the…

统计方法学 · 统计学 2023-06-06 Giuseppe Buccheri , Giacomo Bormetti , Fulvio Corsi , Fabrizio Lillo

In performative Reinforcement Learning (RL), an agent faces a policy-dependent environment: the reward and transition functions depend on the agent's policy. Prior work on performative RL has studied the convergence of repeated retraining…

机器学习 · 计算机科学 2025-05-12 Vasilis Pollatos , Debmalya Mandal , Goran Radanovic

Reinforcement learning algorithms such as the deep deterministic policy gradient algorithm (DDPG) has been widely used in continuous control tasks. However, the model-free DDPG algorithm suffers from high sample complexity. In this paper we…

机器学习 · 计算机科学 2019-11-14 Qingpeng Cai , Ling Pan , Pingzhong Tang

We study projection-free methods for functional constrained optimization with convex or smooth nonconvex objectives. Such problems arise in applications such as portfolio optimization and radiation therapy planning, where risk-aware…

最优化与控制 · 数学 2026-05-12 Yi Cheng , Guanghui Lan , Saeed Masiha , H. Edwin Romeijn

The ensemble Kalman filter (EnKF) is a method for combining a dynamical model with data in a sequential fashion. Despite its widespread use, there has been little analysis of its theoretical properties. Many of the algorithmic innovations…

概率论 · 数学 2015-06-17 D. T. B. Kelly , K. J. H. Law , A. M. Stuart

Policy gradient (PG) methods are successful approaches to deal with continuous reinforcement learning (RL) problems. They learn stochastic parametric (hyper)policies by either exploring in the space of actions or in the space of parameters.…

机器学习 · 计算机科学 2024-05-31 Alessandro Montenegro , Marco Mussi , Alberto Maria Metelli , Matteo Papini

We develop a general mathematical framework to analyze scaling regimes and derive explicit analytic solutions for gradient flow (GF) in large learning problems. Our key innovation is a formal power series expansion of the loss evolution,…

机器学习 · 计算机科学 2026-02-05 Dmitry Yarotsky , Eugene Golikov , Yaroslav Gusev

Several variations of the Kalman filter algorithm, such as the extended Kalman filter (EKF) and the unscented Kalman filter (UKF), are widely used in science and engineering applications. In this paper, we introduce two algorithms of…

最优化与控制 · 数学 2018-10-11 Wei Kang , Liang Xu

The real-world applications in signal processing generally involve estimating the system state or parameters in nonlinear, non-Gaussian dynamic systems. The estimation problem may get even more challenging when there are physical…

信号处理 · 电气工程与系统科学 2022-03-15 Nesrine Amor , Ghulam Rasool , Nidhal C. Bouaynaya

Policy optimization has drawn increasing attention in reinforcement learning, particularly in the context of derivative-free methods for linear quadratic regulator (LQR) problems with unknown dynamics. This paper focuses on characterizing…

最优化与控制 · 数学 2025-06-17 Weijian Li , Panagiotis Kounatidis , Zhong-Ping Jiang , Andreas A. Malikopoulos

In this paper, we consider static parameter estimation for a class of continuous-time state-space models. Our goal is to obtain an unbiased estimate of the gradient of the log-likelihood (score function), which is an estimate that is…

机器学习 · 统计学 2021-06-01 Marco Ballesio , Ajay Jasra

In this paper, a new framework, named as graphical state space model, is proposed for the real time optimal estimation of a class of nonlinear state space model. By discretizing this kind of system model as an equation which can not be…

系统与控制 · 电气工程与系统科学 2022-11-10 Shaolin Lü

This work develops a new multifidelity ensemble Kalman filter (MFEnKF) algorithm based on linear control variate framework. The approach allows for rigorous multifidelity extensions of the EnKF, where the uncertainty in coarser fidelities…

数值分析 · 数学 2020-07-03 Andrey A Popov , Changhong Mou , Traian Iliescu , Adrian Sandu

A hybrid data assimilation algorithm is developed for complex dynamical systems with partial observations. The method starts with applying a spectral decomposition to the entire spatiotemporal fields, followed by creating a machine learning…

计算物理 · 物理学 2022-12-27 Changhong Mou , Leslie M. Smith , Nan Chen

The Kalman filter computes the optimal variable-gain using prior knowledge of the initial state and random (process and measurement) noise distributions, which are assumed to be Gaussian with known variance. However, when these…

系统与控制 · 电气工程与系统科学 2022-01-31 Hugh Lachlan Kennedy

We introduce a new sequential methodology to calibrate the fixed parameters and track the stochastic dynamical variables of a state-space system. The proposed method is based on the nested hybrid filtering (NHF) framework of [1], that…

统计计算 · 统计学 2021-03-24 Sara Pérez-Vieites , Joaquín Míguez

Smoothed particle hydrodynamics (SPH) offers distinct advantages for modeling many engineering problems, yet achieving high-order consistency in its conservative formulation remains to be addressed. While zero- and higher-order…

流体动力学 · 物理学 2024-06-06 Bo Zhang , Nikolaus Adams , Xiangyu Hu

This work extends a previous study that introduced an algorithm for state estimation on manifolds within the framework of the Kalman filter. Its objective is to address the limitations of the earlier approach. The reversible Kalman filter…

系统与控制 · 电气工程与系统科学 2026-01-21 Svyatoslav Covanov , Cedric Pradalier

Direct minimization method on the complex Stiefel manifold in Kohn-Sham density functional theory is formulated to treat both finite and extended systems in a unified manner. This formulation is well-suited for scenarios where…

计算物理 · 物理学 2025-04-02 Kai Luo , Tingguang Wang , Xinguo Ren

We seek to learn an effective policy for a Markov Decision Process (MDP) with continuous states via Q-Learning. Given a set of basis functions over state action pairs we search for a corresponding set of linear weights that minimizes the…

机器学习 · 计算机科学 2013-09-27 Charles Tripp , Ross D. Shachter