English
Related papers

Related papers: Convergence of Policy Iteration for Entropy-Regula…

200 papers

This paper investigates a continuous-time portfolio optimization problem with the following features: (i) a no-short selling constraint; (ii) a leverage constraint, that is, an upper limit for the sum of portfolio weights; and (iii) a…

Portfolio Management · Quantitative Finance 2022-03-08 Masashi Ieda

We consider the problem of computing optimal linear control policies for linear systems in finite-horizon. The states and the inputs are required to remain inside pre-specified safety sets at all times despite unknown disturbances. In this…

Systems and Control · Computer Science 2019-12-17 Luca Furieri , Maryam Kamgarpour

We present a midpoint policy iteration algorithm to solve linear quadratic optimal control problems in both model-based and model-free settings. The algorithm is a variation of Newton's method, and we show that in the model-based setting it…

Optimization and Control · Mathematics 2022-02-16 Benjamin Gravell , Iman Shames , Tyler Summers

We consider a mean-field control problem in which admissible controls are required to be adapted to the common noise filtration. The main objective is to show how the mean-field control problem can be approximates by time consistent…

Optimization and Control · Mathematics 2025-09-19 Bruno Bouchard , Xiaolu Tan

In ergodic singular stochastic control problems, a decision-maker can instantaneously adjust the evolution of a state variable using a control of bounded variation, with the goal of minimizing a long-term average cost functional. The cost…

Optimization and Control · Mathematics 2025-10-14 Alessandro Calvia , Federico Cannerozzi , Giorgio Ferrari

We consider the discrete-time infinite-horizon optimal control problem formalized by Markov Decision Processes. We revisit the work of Bertsekas and Ioffe, that introduced $\lambda$ Policy Iteration, a family of algorithms parameterized by…

Artificial Intelligence · Computer Science 2015-03-13 Bruno Scherrer

Inverse Optimal Control (IOC) seeks to recover an unknown cost from expert demonstrations, and it provides a systematic way of modeling experts' decision mechanisms while considering the prior information of the cost functions.…

Optimization and Control · Mathematics 2025-12-01 Ziliang Wang , Han Zhang , Axel Ringh

Inventory and queueing systems are often designed by controlling weighted combination of some time-averaged performance metrics (like cumulative holding, shortage, server-utilization or congestion costs); but real-world constraints, like…

Optimization and Control · Mathematics 2025-07-01 Madhu Dhiman , Veeraruna Kavitha , Nandyala Hemachandra

This paper investigates a class of multiscale stochastic control problems driven by $\alpha$-stable L\'evy noises, where the controlled dynamics evolve across separate slow and fast time scales. The associated value functions are governed…

Optimization and Control · Mathematics 2025-11-11 Qi Zhang , Yanjie Zhang , Ao Zhang

In the pursuit of finding an optimal policy, reinforcement learning (RL) methods generally ignore the properties of learned policies apart from their expected return. Thus, even when successful, it is difficult to characterize which…

Machine Learning · Computer Science 2025-10-10 Yash Jhaveri , Harley Wiltzer , Patrick Shafto , Marc G. Bellemare , David Meger

Analyzing and controlling system entropy is a powerful tool for regulating predictability of control systems. Applications benefiting from such approaches range from reinforcement learning and data security to human-robot collaboration. In…

Systems and Control · Electrical Eng. & Systems 2026-03-06 Menno van Zutphen , Giannis Delimpaltadakis , Duarte J. Antunes

We study the convergence problem of mean-field control theory in the presence of state constraints and non-degenerate idiosyncratic noise. Our main result is the convergence of the value functions associated to stochastic control problems…

Optimization and Control · Mathematics 2023-06-02 Samuel Daudin

This paper presents the design and analysis of a Hybrid High-Order (HHO) approximation for a distributed optimal control problem governed by the Poisson equation. We propose three distinct schemes to address unconstrained control problems…

Numerical Analysis · Mathematics 2025-01-14 Gouranga Mallik , Ramesh Chandra Sau

In this paper, we propose an interior-point method for linearly constrained optimization problems (possibly nonconvex). The method - which we call the Hessian barrier algorithm (HBA) - combines a forward Euler discretization of Hessian…

Optimization and Control · Mathematics 2023-09-14 Immanuel M. Bomze , Panayotis Mertikopoulos , Werner Schachinger , Mathias Staudigl

We consider an optimal control problem for infinite horizon systems governed by coupled forward-backward stochastic Volterra integral equations with delay. Using Hida-Malliavin calculus, we prove both sufficient and necessary maximum…

Probability · Mathematics 2026-04-02 Ibtissem Djaber , Hafiane Nawel , Samia Yakhlef

We study best-policy identification for finite-horizon risk-sensitive reinforcement learning under the entropic risk measure. Recent work established a constant gap in the exponential horizon dependence between lower and upper bounds on the…

Machine Learning · Computer Science 2026-05-14 Amer Essakine , Claire Vernade

We introduce a new and efficient numerical method for multicriterion optimal control and single criterion optimal control under integral constraints. The approach is based on extending the state space to include information on a "budget"…

Optimization and Control · Mathematics 2016-01-06 Ajeet Kumar , Alexander Vladimirsky

This work uses the entropy-regularised relaxed stochastic control perspective as a principled framework for designing reinforcement learning (RL) algorithms. Herein agent interacts with the environment by generating noisy controls…

Machine Learning · Computer Science 2023-09-18 Lukasz Szpruch , Tanut Treetanthiploet , Yufei Zhang

The paper is devoted to the optimal control of a system with two time-scales, in a regime when the limit equation is not of averaging type but, in the spirit of Wong-Zakai principle, it is a stochastic differential equation for the slow…

Optimization and Control · Mathematics 2024-11-26 Franco Flandoli , Giuseppina Guatteri , Umberto Pappalettera , Gianmario Tessitore

We consider Inverse Electrical Impedance Tomography (EIT) problem on recovering electrical conductivity and potential in the body based on the measurement of the boundary voltages on the $m$ electrodes for a given electrode current. The…

Optimization and Control · Mathematics 2024-12-17 Ugur G. Abdulla , Saleheh Seif
‹ Prev 1 4 5 6 7 8 10 Next ›