English
Related papers

Related papers: Convergence of Policy Iteration for Entropy-Regula…

200 papers

A key problem in reinforcement learning for control with general function approximators (such as deep neural networks and other nonlinear functions) is that, for many algorithms employed in practice, updates to the policy or $Q$-function…

Machine Learning · Computer Science 2016-03-01 Joshua Achiam

This paper studies stochastic control problems with the action space taken to be probability measures, with the objective penalised by the relative entropy. We identify suitable metric space on which we construct a gradient flow for the…

Optimization and Control · Mathematics 2024-01-26 David Šiška , Łukasz Szpruch

We consider a problem of optimal control of an infinite horizon system governed by forward-backward stochastic differential equations with delay. Sufficient and necessary maximum principles for optimal control under partial information in…

Optimization and Control · Mathematics 2013-12-09 Nacira Agram , Bernt Øksendal

Optimal stopping is the problem of determining when to stop a stochastic system in order to maximize reward, which is of practical importance in domains such as finance, operations management and healthcare. Existing methods for…

Optimization and Control · Mathematics 2022-03-28 Xinyi Guan , Velibor V. Mišić

We investigate the existence of solutions of reversible and irreversible port-Hamiltonian systems. To this end, we utilize the associated exergy, a function that is composed of the system's Hamiltonian and entropy, to prove global existence…

Optimization and Control · Mathematics 2024-10-25 Willem Esterhuizen , Bernhard Maschke , Till Preuster , Manuel Schaller , Karl Worthmann

This paper addresses the problem of steering the distribution of the state of a discrete-time linear system to a given target distribution while minimizing an entropy-regularized cost functional. This problem is called a maximum entropy…

Optimization and Control · Mathematics 2024-12-30 Kaito Ito , Kenji Kashima

We propose a novel data-driven neural network (NN) optimization framework for solving an optimal stochastic control problem under stochastic constraints. Customized activation functions for the output layers of the NN are applied, which…

Optimization and Control · Mathematics 2023-06-21 Marc Chen , Mohammad Shirazi , Peter A. Forsyth , Yuying Li

Policy optimization (PO) is a key ingredient for reinforcement learning (RL). For control design, certain constraints are usually enforced on the policies to optimize, accounting for either the stability, robustness, or safety concerns on…

Optimization and Control · Mathematics 2021-02-16 Kaiqing Zhang , Bin Hu , Tamer Başar

Optimal feedback controllers for nonlinear systems can be derived by solving the Hamilton-Jacobi-Bellman (HJB) equation. However, because the HJB is a nonlinear partial differential equation, numerical methods typically provide only…

Optimization and Control · Mathematics 2026-03-25 Morgan Jones , Matthew Peet

We study the problem of optimal control of dissipative quantum dynamics. Although under most circumstances dissipation leads to an increase in entropy (or a decrease in purity) of the system, there is an important class of problems for…

Quantum Physics · Physics 2009-11-10 Shlomo E. Sklarz , David J. Tannor , Navin Khaneja

In this paper, we present a probability one convergence proof, under suitable conditions, of a certain class of actor-critic algorithms for finding approximate solutions to entropy-regularized MDPs using the machinery of stochastic…

Machine Learning · Computer Science 2019-10-23 Wesley Suttle , Zhuoran Yang , Kaiqing Zhang , Ji Liu

We propose a mesh-free policy iteration framework that combines classical dynamic programming with physics-informed neural networks (PINNs) to solve high-dimensional, nonconvex Hamilton--Jacobi--Isaacs (HJI) equations arising in stochastic…

Numerical Analysis · Mathematics 2025-07-24 Hee Jun Yang , Minjung Gim , Yeoneung Kim

Numerically computing global policies to optimal control problems for complex dynamical systems is mostly intractable. In consequence, a number of approximation methods have been developed. However, none of the current methods can quantify…

Robotics · Computer Science 2021-03-05 Ashwin Khadke , Hartmut Geyer

We study the error introduced by entropy regularization in infinite-horizon discrete discounted Markov decision processes. We show that this error decreases exponentially in the inverse regularization strength, both in a weighted…

Optimization and Control · Mathematics 2025-12-16 Johannes Müller , Semih Cayci

This is the first in a series of papers in which we study an efficient approximation scheme for solving the Hamilton-Jacobi-Bellman equation for multi-dimensional problems in stochastic control theory. The method is a combination of a WKB…

Computational Finance · Quantitative Finance 2014-06-26 Sakda Chaiworawitkul , Patrick S. Hagan , Andrew Lesniewski

This paper extends the optimal covariance steering problem for linear stochastic systems subject to chance constraints to account for optimal risk allocation. Previous works have assumed a uniform risk allocation to cast the optimal control…

Optimization and Control · Mathematics 2021-04-14 Joshua Pilipovsky , Panagiotis Tsiotras

This paper investigates optimal control problems for delayed systems governed by Infinitely Anticipated Backward Stochastic Differential Equations (IABSDEs). Unlike existing frameworks limited to bounded delays, we introduce a generalized…

Optimization and Control · Mathematics 2025-12-22 Guanwei Cheng

This paper considers linear-quadratic control of a non-linear dynamical system subject to arbitrary cost. I show that for this class of stochastic control problems the non-linear Hamilton-Jacobi-Bellman equation can be transformed into a…

General Physics · Physics 2009-11-11 H. J. Kappen

A new stochastic control model for the long-run environmental management of rivers is mathematically and numerically analyzed, focusing on a modern sediment replenishment problem with unique nonsmooth and nonlinear properties. Rational…

Optimization and Control · Mathematics 2022-03-11 Hidekazu Yoshioka , Motoh Tsujimura

This paper presents an inverse optimality method to solve the Hamilton-Jacobi-Bellman equation for a class of nonlinear problems for which the cost is quadratic and the dynamics are affine in the input. The method is inverse optimal because…

Optimization and Control · Mathematics 2011-10-11 Luis Rodrigues , Didier Henrion , Mehdi Abedinpour Fallah