English
Related papers

Related papers: Policy Iteration for Exploratory Hamilton--Jacobi-…

200 papers

Despite its popularity in the reinforcement learning community, a provably convergent policy gradient method for continuous space-time control problems with nonlinear state dynamics has been elusive. This paper proposes proximal gradient…

Optimization and Control · Mathematics 2022-12-27 Christoph Reisinger , Wolfgang Stockinger , Yufei Zhang

Reflected diffusions naturally arise in many problems from applications ranging from economics and mathematical biology to queueing theory. In this paper we consider a class of infinite time-horizon singular stochastic control problems for…

Optimization and Control · Mathematics 2017-11-13 Giorgio Ferrari

Decision-making problems in uncertain or stochastic domains are often formulated as Markov decision processes (MDPs). Policy iteration (PI) is a popular algorithm for searching over policy-space, the size of which is exponential in the…

Artificial Intelligence · Computer Science 2013-01-30 Yishay Mansour , Satinder Singh

Convex Q-learning is a recent approach to reinforcement learning, motivated by the possibility of a firmer theory for convergence, and the possibility of making use of greater a priori knowledge regarding policy or value function structure.…

Optimization and Control · Mathematics 2022-10-18 Fan Lu , Joel Mathias , Sean Meyn , Karanjit Kalsi

We prove convergence of the proximal policy gradient method for a class of constrained stochastic control problems with control in both the drift and diffusion of the state process. The problem requires either the running or terminal cost…

Optimization and Control · Mathematics 2025-05-27 Ashley Davey , Harry Zheng

The scaling invariance for chaotic orbits near a transition from unlimited to limited diffusion in a dissipative standard mapping is explained via the analytical solution of the diffusion equation. It gives the probability of observing a…

Chaotic Dynamics · Physics 2020-12-02 Edson D. Leonel , Celia Mayumi Kuwana , Makoto Yoshida , Juliano Antonio de Oliveira

The framework of deep operator network (DeepONet) has been widely exploited thanks to its capability of solving high dimensional partial differential equations. In this paper, we incorporate DeepONet with a recently developed policy…

Optimization and Control · Mathematics 2024-06-18 Jae Yong Lee , Yeoneung Kim

We consider discrete-time infinite horizon deterministic optimal control problems with nonnegative cost per stage, and a destination that is cost-free and absorbing. The classical linear-quadratic regulator problem is a special case. Our…

Optimization and Control · Mathematics 2017-12-20 Dimitri P. Bertsekas

In this paper, which is a continuation of the previously published discrete time paper we develop a theory for continuous time stochastic control problems which, in various ways, are time inconsistent in the sense that they do not admit a…

Optimization and Control · Mathematics 2016-12-13 Tomas Björk , Mariana Khapko , Agatha Murgoci

We study a continuous-time, finite horizon, stochastic partially reversible investment problem for a firm producing a single good in a market with frictions. The production capacity is modeled as a one-dimensional, time-homogeneous, linear…

Optimization and Control · Mathematics 2014-11-13 Tiziano De Angelis , Giorgio Ferrari

The path-integral control, which stems from the stochastic Hamilton-Jacobi-Bellman equation, is one of the methods to control stochastic nonlinear systems. This paper gives a new insight into nonlinear stochastic optimal control problems…

Optimization and Control · Mathematics 2021-09-14 Jun Ohkubo

In this paper we propose an on-line policy iteration (PI) algorithm for finite-state infinite horizon discounted dynamic programming, whereby the policy improvement operation is done on-line, only for the states that are encountered during…

Optimization and Control · Mathematics 2021-06-03 Dimitri Bertsekas

The Hamilton-Jacobi-Bellman equation (HJB) associated with the time inhomogeneous singular control problem is a parabolic partial differential equation, and the existence of a classical solution is usually difficult to prove. In this paper,…

Optimization and Control · Mathematics 2014-10-14 Yipeng Yang

This article studies a portfolio optimization problem, where the market consisting of several stocks is modeled by a multi-dimensional jump-diffusion process with age-dependent semi-Markov modulated coefficients. We study risk sensitive…

Portfolio Management · Quantitative Finance 2019-10-21 Milan Kumar Das , Anindya Goswami , Nimit Rana

We introduce some approximation schemes for linear and fully non-linear diffusion equations of Bellman-Isaacs type. Although they are not monotone one can prove their convergence to the viscosity solution of the problem. Effective…

Optimization and Control · Mathematics 2015-01-22 Xavier Warin

We examine the problem of two-point boundary optimal control of nonlinear systems over finite-horizon time periods with unknown model dynamics by employing reinforcement learning. We use techniques from singular perturbation theory to…

Optimization and Control · Mathematics 2023-06-12 Vasanth Reddy , Hoda Eldardiry , Almuatazbellah Boker

This paper deals with a class of neural SDEs and studies the limiting behavior of the associated sampled optimal control problems as the sample size grows to infinity. The neural SDEs with $N$ samples can be linked to the $N$-particle…

Optimization and Control · Mathematics 2025-06-19 Huafu Liao , Alpár R. Mészáros , Chenchen Mou , Chao Zhou

We consider a general class of stochastic optimal control problems, where the state process lives in a real separable Hilbert space and is driven by a cylindrical Brownian motion and a Poisson random measure; no special structure is imposed…

Probability · Mathematics 2018-10-04 Elena Bandini , Fulvia Confortola , Andrea Cosso

In this paper, we study distributional reinforcement learning from the perspective of statistical efficiency. We investigate distributional policy evaluation, aiming to estimate the complete return distribution (denoted $\eta^\pi$) attained…

Machine Learning · Statistics 2025-11-13 Liangyu Zhang , Yang Peng , Jiadong Liang , Wenhao Yang , Zhihua Zhang

Controlling systems of ordinary differential equations (ODEs) is ubiquitous in science and engineering. For finding an optimal feedback controller, the value function and associated fundamental equations such as the Bellman equation and the…

Optimization and Control · Mathematics 2021-04-14 Mathias Oster , Leon Sallandt , Reinhold Schneider
‹ Prev 1 4 5 6 7 8 10 Next ›