中文
相关论文

相关论文: Approximative Policy Iteration for Exit Time Feedb…

200 篇论文

We propose and analyze a posteriori error estimators for an optimal control problem that involves an elliptic partial differential equation as state equation and a control variable that enters the state equation as a coefficient; pointwise…

最优化与控制 · 数学 2022-03-31 Francisco Fuica , Enrique Otarola

We investigate the adaptive robust control framework for portfolio optimization and loss-based hedging under drift and volatility uncertainty. Adaptive robust problems offer many advantages but require handling a double optimization problem…

最优化与控制 · 数学 2020-05-06 Tao Chen , Michael Ludkovski

We study the synthesis of optimal control policies for large-scale multi-agent systems. The optimal control design induces a parsimonious control intervention by means of l-1, sparsity-promoting control penalizations. We study instantaneous…

最优化与控制 · 数学 2016-11-15 Giacomo Albi , Massimo Fornasier , Dante Kalise

We study methods for solving stochastic control problems of systems of forward-backward mean-field equations with delay, in finite or infinite horizon. Necessary and sufficient maximum principles under partial information are given. The…

最优化与控制 · 数学 2016-10-31 Nacira Agram , Elin Engen Rose

We develop Bellman equation based approach for infinite time horizon optimal impulsive control problems. Both discounted and time average criteria are considered. We establish very general and at the same time natural conditions under which…

网络与互联网体系结构 · 计算机科学 2013-11-28 Konstantin Avrachenkov , Oussama Habachi , Alexei Piunovskiy , Zhang Yi

Policy iteration and value iteration are at the core of many (approximate) dynamic programming methods. For Markov Decision Processes with finite state and action spaces, we show that they are instances of semismooth Newton-type methods to…

最优化与控制 · 数学 2022-06-28 Matilde Gargiani , Andrea Zanelli , Dominic Liao-McPherson , Tyler Summers , John Lygeros

In this paper we study the problem of synthesizing optimal control policies for uncertain continuous-time nonlinear systems from syntactically co-safe linear temporal logic (scLTL) formulas. We formulate this problem as a sequence of…

系统与控制 · 电气工程与系统科学 2021-04-16 Max Cohen , Calin Belta

We describe a nonlinear generalization of dual dynamic programming theory and its application to value function estimation for deterministic control problems over continuous state and action spaces, in a discrete-time infinite horizon…

最优化与控制 · 数学 2018-10-05 Joseph Warrington , Paul N. Beuchat , John Lygeros

We present a neural network approach for approximating the value function of high-dimensional stochastic control problems. Our training process simultaneously updates our value function estimate and identifies the part of the state space…

最优化与控制 · 数学 2024-05-08 Xingjian Li , Deepanshu Verma , Lars Ruthotto

This paper is concerned with optimal control problems for control systems in continuous time, and interacting particle system methods designed to construct approximate control solutions. Particular attention is given to the linear quadratic…

系统与控制 · 电气工程与系统科学 2022-07-11 Anant Joshi , Amirhossein Taghvaei , Prashant G. Mehta , Sean P. Meyn

This paper investigates a class of optimal control problems associated with Markov processes with local state information. The decision-maker has only local access to a subset of a state vector information as often encountered in…

系统与控制 · 电气工程与系统科学 2020-05-12 Guanze Peng , Veeraruna Kavitha , Qunayan Zhu

We introduce a new numerical method to approximate the solution of a finite horizon deterministic optimal control problem. We exploit two Hamilton-Jacobi-Bellman PDE, arising by considering the dynamics in forward and backward time. This…

最优化与控制 · 数学 2023-04-21 Marianne Akian , Stéphane Gaubert , Shanqing Liu

We develop an optimization-based framework for joint real-time trajectory planning and feedback control of feedback-linearizable systems. To achieve this goal, we define a target trajectory as the optimal solution of a time-varying…

系统与控制 · 电气工程与系统科学 2020-03-17 Tianqi Zheng , John Simpson-Porco , Enrique Mallada

We develop a neural-network framework for multi-period risk--reward stochastic control problems with constrained two-step feedback policies that may be discontinuous in the state. We allow a broad class of objectives built on a…

计算金融 · 定量金融 2026-03-09 Chang Chen , Duy-Minh Dang

We provide a solution to the problem of receding horizon control for stochastic discrete-time systems with bounded control inputs and imperfect state measurements. For a suitable choice of control policies, we show that the finite-horizon…

最优化与控制 · 数学 2010-04-15 Peter Hokayem , Eugenio Cinquemani , Debasish Chatterjee , Federico Ramponi , John Lygeros

In this paper we consider discrete time stochastic optimal control problems over infinite and finite time horizons. We show that for a large class of such problems the Taylor polynomials of the solutions to the associated Dynamic…

最优化与控制 · 数学 2019-03-26 Arthur J Krener

In this paper, we aim to solve the high dimensional stochastic optimal control problem from the view of the stochastic maximum principle via deep learning. By introducing the extended Hamiltonian system which is essentially an FBSDE with a…

最优化与控制 · 数学 2021-06-23 Shaolin Ji , Shige Peng , Ying Peng , Xichuan Zhang

Time-consistency is an essential requirement in risk sensitive optimal control problems to make rational decisions. An optimization problem is time consistent if its solution policy does not depend on the time sequence of solving the…

最优化与控制 · 数学 2015-03-26 Yinlam Chow , Marco Pavone

In this paper, we present a novel method for computing the optimal feedback gain of the infinite-horizon Linear Quadratic Regulator (LQR) problem via an ordinary differential equation. We introduce a novel continuous-time Bellman error,…

系统与控制 · 电气工程与系统科学 2026-04-17 Armin Gießler , Albertus Johannes Malan , Sören Hohmann

Tensor train (TT) format is a common approach for computationally efficient work with multidimensional arrays, vectors, matrices, and discretized functions in a wide range of applications, including computational mathematics and machine…

数值分析 · 数学 2022-09-30 Andrei Chertkov , Gleb Ryzhakov , Georgii Novikov , Ivan Oseledets