English
Related papers

Related papers: Research on Optimal Control Problem Based on Reinf…

200 papers

In this paper, we study the delayed stochastic recursive optimal control problem with a non-Lipschitz generator, in which both the dynamics of the control system and the recursive cost functional depend on the past path segment of the state…

Optimization and Control · Mathematics 2023-12-27 Jiaqiang Wen , Zhen Wu , Qi Zhang

We consider an optimal control problem arising in the context of economic theory of growth, on the lines of the works by Skiba (1978) and Askenazy - Le Van (1999). The economic framework of the model is intertemporal infinite horizon…

Optimization and Control · Mathematics 2014-09-05 Francesco Bartaloni

The convergence of policy gradient algorithms in reinforcement learning hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of…

Machine Learning · Computer Science 2023-11-01 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

This paper is concerned with a stochastic recursive optimal control problem with time delay, where the controlled system is described by a stochastic differential delayed equation (SDDE) and the cost functional is formulated as the solution…

Optimization and Control · Mathematics 2014-08-26 Jingtao Shi , Huanshui Zhang

Distributed optimal control is known to be challenging and can become intractable even for linear-quadratic regulator problems. In this work, we study a special class of such problems where distributed state feedback controllers can give…

Systems and Control · Electrical Eng. & Systems 2024-03-14 Johan Olsson , Runyu Zhang , Emma Tegling , Na Li

The framework of deep operator network (DeepONet) has been widely exploited thanks to its capability of solving high dimensional partial differential equations. In this paper, we incorporate DeepONet with a recently developed policy…

Optimization and Control · Mathematics 2024-06-18 Jae Yong Lee , Yeoneung Kim

The problem of order execution is cast as a relative entropy-regularized robust optimal control problem in this article. The order execution agent's goal is to maximize an objective functional associated with his profit-and-loss of trading…

Optimization and Control · Mathematics 2024-09-11 Meng Wang , Tai-Ho Wang

As autonomous systems become more ubiquitous in daily life, ensuring high performance with guaranteed safety is crucial. However, safety and performance could be competing objectives, which makes their co-optimization difficult.…

Robotics · Computer Science 2025-05-29 Manan Tayal , Aditya Singh , Shishir Kolathaya , Somil Bansal

Optimal control deals with optimization problems in which variables steer a dynamical system, and its outcome contributes to the objective function. Two classical approaches to solving these problems are Dynamic Programming and the…

Optimization and Control · Mathematics 2023-12-18 Alessandro Betti , Michele Casoni , Marco Gori , Simone Marullo , Stefano Melacci , Matteo Tiezzi

When randomness in demand affects the sales of a product, retailers use dynamic pricing strategies to maximize their profits. In this article, we formulate the pricing problem as a continuous-time stochastic optimal control problem and find…

Optimization and Control · Mathematics 2019-03-13 Asbjørn Nilsen Riseth

Hard constraints in reinforcement learning (RL) often degrade policy performance. Lagrangian methods offer a way to blend objectives with constraints, but require intricate reward engineering and parameter tuning. In this work, we extend…

Artificial Intelligence · Computer Science 2025-12-05 William Sharpless , Dylan Hirsch , Sander Tonkens , Nikhil Shinde , Sylvia Herbert

Without exact knowledge of the true system dynamics, optimal control of non-linear continuous-time systems requires careful treatment under epistemic uncertainty. In this work, we translate a probabilistic interpretation of the Pontryagin…

Machine Learning · Computer Science 2025-09-03 David Leeftink , Çağatay Yıldız , Steffen Ridderbusch , Max Hinne , Marcel van Gerven

We study an optimal allocation problem for a system of independent Brownian agents whose states evolve under a limited shared control. At each time, a unit of resource can be divided and allocated across components to increase their drifts,…

Optimization and Control · Mathematics 2026-03-31 Gaoyue Guo , Wenpin Tang , Nizar Touzi

Motivated by the novel paradigm developed by Van Roy and coauthors for reinforcement learning in arbitrary non-Markovian environments, we propose a related formulation and explicitly pin down the error caused by non-Markovianity of…

Systems and Control · Electrical Eng. & Systems 2024-02-15 Siddharth Chandak , Pratik Shah , Vivek S Borkar , Parth Dodhia

We study the exploration-exploitation dilemma in the linear quadratic regulator (LQR) setting. Inspired by the extended value iteration algorithm used in optimistic algorithms for finite MDPs, we propose to relax the optimistic optimization…

Machine Learning · Statistics 2020-07-14 Marc Abeille , Alessandro Lazaric

Modeling the purposeful behavior of imperfect agents from a small number of observations is a challenging task. When restricted to the single-agent decision-theoretic setting, inverse optimal control techniques assume that observed behavior…

Computer Science and Game Theory · Computer Science 2015-03-19 Kevin Waugh , Brian D. Ziebart , J. Andrew Bagnell

We examine the problem of two-point boundary optimal control of nonlinear systems over finite-horizon time periods with unknown model dynamics by employing reinforcement learning. We use techniques from singular perturbation theory to…

Optimization and Control · Mathematics 2023-06-12 Vasanth Reddy , Hoda Eldardiry , Almuatazbellah Boker

Empowerment is an information-theoretic method that can be used to intrinsically motivate learning agents. It attempts to maximize an agent's control over the environment by encouraging visiting states with a large number of reachable next…

Machine Learning · Computer Science 2020-01-09 Felix Leibfried , Sergio Pascual-Diaz , Jordi Grau-Moya

We consider a model of optimal investment and consumption with both habit formation and partial observations in incomplete It\^{o} processes market. The investor chooses his consumption under the addictive habits constraint while only…

Portfolio Management · Quantitative Finance 2014-08-12 Xiang Yu

In the context of optimal control, we consider the inverse problem of Lagrangian identification given system dynamics and optimal trajectories. Many of its theoretical and practical aspects are still open. Potential applications are very…

Optimization and Control · Mathematics 2014-03-21 Edouard Pauwels , Didier Henrion , Jean-Bernard Bernard Lasserre