中文
相关论文

相关论文: On the convergence of a Risk Sensitive like Filter

200 篇论文

For many nonlinear Bayesian state estimation problems, the posterior recursion is not analytically tractable, leading to algorithms that are influenced by numerical approximation errors. These algorithms depend on parameters that affect the…

系统与控制 · 电气工程与系统科学 2026-05-14 Ondrej Straka , Felipe Giraldo-Grueso , Renato Zanetti

In [1], it is established that a convergent observer with an infinite gain margin can be designed for a given nonlinear system when a Riemannian metric showing that the system is differentially detectable (i.e., the Lie derivative of the…

最优化与控制 · 数学 2016-06-21 Ricardo G. Sanfelice , Laurent Praly

The area of fault detection is becoming more interesting since there have been many unique designs to detect or even compensate the faults, either from sensor or actuator. This paper applies the hydraulic system with interconnected tanks by…

系统与控制 · 电气工程与系统科学 2025-07-08 Moh Kamalul Wafi

Filter convergence of vector lattice-valued measures is considered, in order to deduce theorems of convergence for their decompositions. First the $\sigma$-additive case is studied, without particular assumptions on the filter; later the…

泛函分析 · 数学 2015-08-12 Domenico Candeloro , Anna Rita Sambucini

Time-consistency is an essential requirement in risk sensitive optimal control problems to make rational decisions. An optimization problem is time consistent if its solution policy does not depend on the time sequence of solving the…

最优化与控制 · 数学 2015-03-26 Yinlam Chow , Marco Pavone

We study the convergence of stochastic fixed point iterations in the consistent case (in the sense of Butnariu and Fl{\aa}m (1995)) in several different settings, under decreasingly restrictive regularity assumptions of the fixed point…

最优化与控制 · 数学 2020-03-26 Neal Hermer , D. Russell Luke , Anja Sturm

We establish linear convergence rates for a certain class of extrapolated fixed point algorithms which are based on dynamic string-averaging methods in a real Hilbert space. This applies, in particular, to the extrapolated simultaneous and…

最优化与控制 · 数学 2018-05-11 Christian Bargetz , Victor I. Kolobov , Simeon Reich , Rafał Zalas

We study whether a risk-sensitive objective from asset-pricing theory -- recursive utility -- improves reinforcement learning for portfolio allocation. The Bellman equation under recursive utility involves a certainty equivalent (CE) of…

综合金融 · 定量金融 2026-03-25 Minkey Chang

This work theoretically studies a ubiquitous reinforcement learning policy for controlling the canonical model of continuous-time stochastic linear-quadratic systems. We show that randomized certainty equivalent policy addresses the…

机器学习 · 计算机科学 2022-08-23 Mohamad Kazem Shirani Faradonbeh

We develop a framework for quantitative convergence analysis of Picard iterations of expansive set-valued fixed point mappings. There are two key components of the analysis. The first is a natural generalization of single-valued averaged…

最优化与控制 · 数学 2018-09-24 D. Russell Luke , Nguyen H. Thao , Matthew K. Tam

This paper develops a unified framework, based on iterated random operator theory, to analyze the convergence of constant stepsize recursive stochastic algorithms (RSAs). RSAs use randomization to efficiently compute expectations, and so…

机器学习 · 计算机科学 2021-01-06 Abhishek Gupta , William B. Haskell

We address the problem of inverse reinforcement learning in Markov decision processes where the agent is risk-sensitive. In particular, we model risk-sensitivity in a reinforcement learning framework by making use of models of human…

机器学习 · 计算机科学 2017-11-23 Lillian J. Ratliff , Eric Mazumdar

The closed-loop stability and infinite-horizon performance of receding-horizon approximations are studied for non-stationary linear-quadratic regulator (LQR) problems. The approach is based on a lifted reformulation of the optimal control…

系统与控制 · 电气工程与系统科学 2023-09-06 Jintao Sun , Michael Cantoni

In this letter, we propose an iterative joint detection algorithm of Kalman filter (KF) and channel decoder for the sensor-to-controller link of wireless networked control systems, which utilizes the prior information of control system to…

信息论 · 计算机科学 2025-12-19 Jinnan Piao , Dong Li , Yiming Sun , Zhibo Li , Ming Yang , Xueting Yu

For linear time-invariant systems with uncertain parameters belonging to a finite set, we present a purely deterministic approach to multiple-model estimation and propose an algorithm based on the minimax criterion using constrained…

最优化与控制 · 数学 2022-07-18 Olle Kjellqvist , Anders Rantzer

On a polarized manifold $(X,L)$, the Bergman iteration $\phi_k^{(m)}$ is defined as a sequence of Bergman metrics on $L$ with two integer parameters $k, m$. We study the relation between the K\"ahler-Ricci flow $\phi_t$ at any time $t \geq…

微分几何 · 数学 2019-03-14 Ryosuke Takahashi

The objective is to investigate the advantages and performance of Extended Kalman Filter for the estimation of non-linear system where linearization takes place about a trajectory that was continually updated with the state estimates…

最优化与控制 · 数学 2007-07-16 Subrata Bhowmik , Chandrani Roy

In this article we consider a consistent convex feasibility problem in a real Hilbert space defined by a finite family of sets $C_i$. We are interested, in particular, in the case where for each $i$, $C_i=Fix (U_i)=\{z\in \mathcal H\mid…

最优化与控制 · 数学 2017-03-29 Victor I. Kolobov , Simeon Reich , Rafał Zalas

It has been proposed that classical filtering methods, like the Kalman filter and 3DVAR, can be used to solve linear statistical inverse problems. In the work of Iglesias, Lin, Lu, & Stuart (2017), error estimates were obtained for this…

数值分析 · 数学 2022-05-12 Felix G. Jones , Gideon Simpson

Adaptive optimal control of nonlinear dynamic systems with deterministic and known dynamics under a known undiscounted infinite-horizon cost function is investigated. Policy iteration scheme initiated using a stabilizing initial control is…

系统与控制 · 计算机科学 2015-05-21 Ali Heydari