中文
相关论文

相关论文: Approximative Policy Iteration for Exit Time Feedb…

200 篇论文

We describe an algorithm to solve Bellman optimization that replaces a sum over paths determining the optimal cost-to-go by an analytic method localized in state space. Our approach follows from the established relation between stochastic…

最优化与控制 · 数学 2022-12-02 Michael D. Schneider , Caleb Miller , George F. Chapline , Jane Pratt , Dan Merl

We propose a novel reformulation of the stochastic optimal control problem as an approximate inference problem, demonstrating, that such a interpretation leads to new practical methods for the original problem. In particular we characterise…

机器学习 · 计算机科学 2010-09-22 Konrad Rawlik , Marc Toussaint , Sethu Vijayakumar

We consider challenging dynamic programming models where the associated Bellman equation, and the value and policy iteration algorithms commonly exhibit complex and even pathological behavior. Our analysis is based on the new notion of…

最优化与控制 · 数学 2016-09-13 Dimitri P. Bertsekas

We study a constrained stochastic control problem with jumps; the jump times of the controlled process are given by a Poisson process. The cost functional comprises quadratic components for an absolutely continuous control and the…

最优化与控制 · 数学 2013-04-29 Peter Kratz

Under a Bayesian framework, we formulate the fully sequential sampling and selection decision in statistical ranking and selection as a stochastic control problem, and derive the associated Bellman equation. Using value function…

机器学习 · 计算机科学 2017-10-10 Yijie Peng , Edwin K. P. Chong , Chun-Hung Chen , Michael C. Fu

This paper studies an optimal control problem for continuous-time stochastic systems subject to reachability objectives specified in a subclass of metric interval temporal logic specifications, a temporal logic with real-time constraints.…

系统与控制 · 计算机科学 2015-04-21 Jie Fu , Ufuk Topcu

We consider the problem of learning a linear control policy for a linear dynamical system, from demonstrations of an expert regulating the system. The standard approach to this problem is policy fitting, which fits a linear policy by…

最优化与控制 · 数学 2020-01-22 Malayandi Palan , Shane Barratt , Alex McCauley , Dorsa Sadigh , Vikas Sindhwani , Stephen Boyd

A gradient-enhanced functional tensor train cross approximation method for the resolution of the Hamilton-Jacobi-Bellman (HJB) equations associated to optimal feedback control of nonlinear dynamics is presented. The procedure uses samples…

数值分析 · 数学 2023-02-23 Sergey Dolgov , Dante Kalise , Luca Saluzzi

This paper presents a method to approximately solve stochastic optimal control problems in which the cost function and the system dynamics are polynomial. For stochastic systems with polynomial dynamics, the moments of the state can be…

最优化与控制 · 数学 2017-02-24 Andrew Lamperski , Khem Raj Ghusinga , Abhyudai Singh

An infinite-dimensional bilinear optimal control problem with infinite-time horizon is considered. The associated value function can be expanded in a Taylor series around the equilibrium, the Taylor series involving multilinear forms which…

最优化与控制 · 数学 2017-09-14 Tobias Breiten , Karl Kunisch , Laurent Pfeiffer

Markov decision problems are most commonly solved via dynamic programming. Another approach is Bellman residual minimization, which directly minimizes the squared Bellman residual objective function. However, compared to dynamic…

机器学习 · 计算机科学 2026-04-28 Donghwan Lee , Hyukjun Yang

We propose a policy iteration algorithm for solving the multiplicative noise linear quadratic output feedback design problem. The algorithm solves a set of coupled Riccati equations for estimation and control arising from a partially…

系统与控制 · 电气工程与系统科学 2022-04-01 Benjamin Gravell , Matilde Gargiani , John Lygeros , Tyler H. Summers

Using results from quantum filtering theory and methods from classical control theory, we derive an optimal control strategy for an open two-level system (a qubit in interaction with the electromagnetic field) controlled by a laser. The aim…

量子物理 · 物理学 2009-11-10 Luc Bouten , Simon Edwards , V P Belavkin

Policy iteration (PI) is a recursive process of policy evaluation and improvement for solving an optimal decision-making/control problem, or in other words, a reinforcement learning (RL) problem. PI has also served as the fundamental for…

人工智能 · 计算机科学 2021-04-06 Jaeyoung Lee , Richard S. Sutton

This paper presents sufficient conditions for optimal control of systems with dynamics given by a linear operator, in order to obtain an explicit solution to the Bellman equation that can be calculated in a distributed fashion. Further, the…

最优化与控制 · 数学 2025-06-19 David Ohlin , Richard Pates , Murat Arcak

This paper studies the optimal tracking control problem for continuous-time stochastic linear systems with multiplicative noise. The solution framework involves solving a stochastic algebraic Riccati equation for the feedback gain and a…

系统与控制 · 电气工程与系统科学 2025-08-29 Jiayu Chen , Zhenhui Xu , Xinghu Wang

The solution to a stochastic optimal control problem can be determined by computing the value function from a discretization of the associated Hamilton-Jacobi-Bellman equation. Alternatively, the problem can be reformulated in terms of a…

最优化与控制 · 数学 2024-02-29 Sebastian Reich

This paper studies approximate policy iteration (API) methods which use least-squares Bellman error minimization for policy evaluation. We address several of its enhancements, namely, Bellman error minimization using instrumental variables,…

最优化与控制 · 数学 2014-01-07 Warren R. Scott , Warren B. Powell , Somayeh Moazehi

This paper addresses the numerical solution of backward stochastic differential equations (BSDEs) arising in stochastic optimal control. Specifically, we investigate two BSDEs: one derived from the Hamilton-Jacobi-Bellman equation and the…

最优化与控制 · 数学 2025-03-12 Yuhang Mei , Amirhossein Taghvaei

We consider the approximation of some optimal control problems for the Navier-Stokes equation via a Dynamic Programming approach. These control problems arise in many industrial applications and are very challenging from the numerical point…

最优化与控制 · 数学 2022-07-18 Maurizio Falcone , Gerhard Kirsten , Luca Saluzzi