English
Related papers

Related papers: On the policy improvement algorithm in continuous …

200 papers

This paper deals with time-optimal control of nonlinear continuous-time systems based on direct collocation. The underlying discretization grid is variable in time, as the time intervals are subject to optimization. This technique differs…

Systems and Control · Electrical Eng. & Systems 2020-05-26 Christoph Rösmann , Artemi Makarow , Torsten Bertram

A key problem in reinforcement learning for control with general function approximators (such as deep neural networks and other nonlinear functions) is that, for many algorithms employed in practice, updates to the policy or $Q$-function…

Machine Learning · Computer Science 2016-03-01 Joshua Achiam

In this paper we consider discrete time stochastic optimal control problems over infinite and finite time horizons. We show that for a large class of such problems the Taylor polynomials of the solutions to the associated Dynamic…

Optimization and Control · Mathematics 2019-03-26 Arthur J Krener

We consider a singular stochastic control problem, which is called the Monotone Follower Stochastic Control Problem and give sufficient conditions for the existence and uniqueness of a local-time type optimal control. To establish this…

Optimization and Control · Mathematics 2007-05-23 Erhan Bayraktar , Masahiko Egami

The paper is devoted to the study of a new class of optimal control problems for nonsmooth dynamical systems governed by nonconvex discontinuous differential inclusions of the sweeping type with involving variable time into optimization. We…

Optimization and Control · Mathematics 2025-03-05 Tan H. Cao , Boris S. Mordukhovich , Dao Nguyen , Trang Nguyen , Nguyen N. Thieu

This paper focuses on stochastic optimal control problems with constraints in law, which are rewritten as optimization (minimization) of probability measures problem on the canonical space. We introduce a penalized version of this type of…

Optimization and Control · Mathematics 2025-03-18 Thibaut Bourdais , Nadia Oudjane , Francesco Russo

We perform a detailed study of a simple mathematical model addressing the problem of optimally regulating a process subject to periodic external forcing, which is interesting both in view of its direct applications and as a prototype for…

Optimization and Control · Mathematics 2025-03-04 Nir Gavish , Guy Katriel

Reinforcement learning for control over continuous spaces typically uses high-entropy stochastic policies, such as Gaussian distributions, for local exploration and estimating policy gradient to optimize performance. Many robotic control…

Machine Learning · Computer Science 2024-04-03 Ya-Chien Chang , Sicun Gao

In this paper we develop a novel, discrete-time optimal control framework for mechanical systems with uncertain model parameters. We consider finite-horizon problems where the performance index depends on the statistical moments of the…

Optimization and Control · Mathematics 2017-05-17 George I. Boutselis , Yunpeng Pan , Gerardo De La Tore , Evangelos A. Theodorou

The Inverse Optimal Control (IOC) problem is a structured system identification problem that aims to identify the underlying objective function based on observed optimal trajectories. This provides a data-driven way to model experts'…

Optimization and Control · Mathematics 2024-02-28 Han Zhang , Axel Ringh

This paper considers the problem of designing time-dependent, real-time control policies for controllable nonlinear diffusion processes, with the goal of obtaining maximally-informative observations about parameters of interest. More…

Methodology · Statistics 2014-10-16 Giles Hooker , Kevin K. Lin , Bruce Rogers

The paper considers a stabilizing stochastic control which can be applied to a variety of unstable and even chaotic maps. Compared to previous methods introducing control by noise, we relax assumptions on the class of maps, as well as…

Dynamical Systems · Mathematics 2019-02-25 Elena Braverman , Alexandra Rodkina

This paper studies robust time-inconsistent (TIC) linear-quadratic stochastic control problems, formulated by stochastic differential games. By a spike variation approach, we derive sufficient conditions for achieving the Nash equilibrium,…

Optimization and Control · Mathematics 2025-04-29 Bingyan Han , Chi Seng Pun , Hoi Ying Wong

This research considers the ranking and selection with input uncertainty. The objective is to maximize the posterior probability of correctly selecting the best alternative under a fixed simulation budget, where each alternative is measured…

Optimization and Control · Mathematics 2023-05-15 Hui Xiao , Zhihong Wei

In this paper, we study a discrete-time stochastic optimal control problem under distribution uncertainty with convex control domain. By weak convergence method and Sion's minimax theorem, we obtain the variational inequality for cost…

Optimization and Control · Mathematics 2022-06-28 Mingshang Hu , Shaolin Ji , Xiaojuan Li

Stochastic Spatio-Temporal processes are prevalent across domains ranging from modeling of plasma to the turbulence in fluids to the wave function of quantum systems. This letter studies a measure-theoretic description of such systems by…

Optimization and Control · Mathematics 2021-05-25 George I. Boutselis , Ethan N. Evans , Marcus A. Pereira , Evangelos A. Theodorou

In this paper we study continuous-time stochastic control problems with both monotone and classical controls motivated by the so-called public good contribution problem. That is the problem of n economic agents aiming to maximize their…

Optimization and Control · Mathematics 2018-05-23 Giorgio Ferrari , Frank Riedel , Jan-Henrik Steg

This paper studies an infinite horizon optimal control problem for discrete-time linear system and quadratic criteria, both with random parameters which are independent and identically distributed with respect to time. In this general…

Optimization and Control · Mathematics 2024-03-04 Deyue Li

Policy optimization is an effective reinforcement learning approach to solve continuous control tasks. Recent achievements have shown that alternating online and offline optimization is a successful choice for efficient trajectory reuse.…

Machine Learning · Computer Science 2018-11-01 Alberto Maria Metelli , Matteo Papini , Francesco Faccio , Marcello Restelli

A variant of the optimal control problem is considered which is nonstandard in that the performance index contains "stochastic" integrals, that is, integrals against very irregular functions. The motivation for considering such performance…

Optimization and Control · Mathematics 2018-05-24 Jochen Bröcker