English
Related papers

Related papers: Backward Stochastic Control System with Entropy Re…

200 papers

In two-player zero-sum stochastic games, where two competing players make decisions under uncertainty, a pair of optimal strategies is traditionally described by Nash equilibrium and computed under the assumption that the players have…

Optimization and Control · Mathematics 2019-07-30 Yagiz Savas , Mohamadreza Ahmadi , Takashi Tanaka , Ufuk Topcu

We consider the problem of learning the optimal policy for Markov decision processes with safety constraints. We formulate the problem in a reach-avoid setup. Our goal is to design online reinforcement learning algorithms that ensure safety…

Machine Learning · Computer Science 2026-01-21 Abhijit Mazumdar , Rafal Wisniewski , Manuela L. Bujorianu

We examine a multi-stage stochastic optimization problem characterized by stagewise-independent, decision-dependent noises with strict constraints. The problem assumes convexity in that, following a specific relaxation, it transforms into a…

Optimization and Control · Mathematics 2023-08-28 Chen Yan , Alexandre Reiffers-Masson

Entropy regularization has been widely used in policy optimization algorithms to enhance exploration and the robustness of the optimal control; however it also introduces an additional regularization bias. This work quantifies the impact of…

Optimization and Control · Mathematics 2025-03-25 Deven Sethi , David Šiška , Yufei Zhang

In this paper, we focus on a method based on optimal control to address the optimization problem. The objective is to find the optimal solution that minimizes the objective function. We transform the optimization problem into optimal…

Optimization and Control · Mathematics 2023-09-12 Yeming Xu , Ziyuan Guo , Hongxia Wang , Huanshui Zhang

This paper investigates the so-called reward-balancing methods, a novel class of algorithms for solving discounted-return reinforcement learning (RL) problems. These methods consist of iteratively adjusting the reward function to transform…

Optimization and Control · Mathematics 2026-04-23 Simone Baroncini , Bahman Gharesifard , Giuseppe Notarstefano

In this paper, we consider a class of stochastic control problems for stochastic differential equations with random coefficients. The control domain need not to be convex but the control process is not allowed to enter in diffusion term.…

Optimization and Control · Mathematics 2020-08-06 Ishak Alia , Mohamed Sofiane Alia

This paper introduces a new formulation for stochastic optimal control and stochastic dynamic optimization that ensures safety with respect to state and control constraints. The proposed methodology brings together concepts such as…

Systems and Control · Electrical Eng. & Systems 2021-02-19 Marcus Aloysius Pereira , Ziyi Wang , Ioannis Exarchos , Evangelos A. Theodorou

In this paper, we consider optimal control of stochastic differential equations subject to an expected path constraint. The stochastic maximum principle is given for a general optimal stochastic control in terms of constrained FBSDEs. In…

Optimization and Control · Mathematics 2022-08-16 Ying Hu , Shanjian Tang , Zuo Quan Xu

While techniques have been developed for chance constrained stochastic optimal control using sample disturbance data that provide a probabilistic confidence bound for chance constraint satisfaction, far less is known about how to use sample…

Systems and Control · Electrical Eng. & Systems 2023-03-31 Shawn Priore , Meeko Oishi

We introduce a continuous policy-value iteration algorithm where the approximations of the value function of a stochastic control problem and the optimal control are simultaneously updated through Langevin-type dynamics. This framework…

Optimization and Control · Mathematics 2025-06-11 Qi Feng , Gu Wang

This paper is concerned with the maximum principle of stochastic optimal control problems, where the coefficients of the state equation and the cost functional are uncertain, and the system is generally under Markovian regime switching.…

Optimization and Control · Mathematics 2025-04-15 Tao Hao , Jiaqiang Wen , Jie Xiong

When deploying artificial agents in real-world environments where they interact with humans, it is crucial that their behavior is aligned with the values, social norms or other requirements of that environment. However, many environments…

Machine Learning · Computer Science 2023-05-05 Mattijs Baert , Pietro Mazzaglia , Sam Leroux , Pieter Simoens

In this paper, we study the maximum principle for stochastic optimal control problems of forward-backward stochastic difference systems (FBS{\Delta}Ss). Two types of FBS{\Delta}Ss are investigated. The first one is described by a partially…

Optimization and Control · Mathematics 2019-01-01 Shaolin Ji , Haodong Liu

This paper presents a novel synthesis method for designing an optimal and robust guidance law for a non-throttleable upper stage of a launch vehicle, using a convex approach. In the unperturbed scenario, a combination of lossless and…

Optimization and Control · Mathematics 2022-12-14 Boris Benedikter , Alessandro Zavoli , Guido Colasurdo , Simone Pizzurro , Enrico Cavallini

We establish an algorithm to learn feedback maps from data for a class of robust model predictive control (MPC) problems. The algorithm accounts for the approximation errors due to the learning directly at the synthesis stage, ensuring…

Optimization and Control · Mathematics 2025-10-16 Siddhartha Ganguly , Shubham Gupta , Debasish Chatterjee

This paper investigates optimal control problems for delayed systems governed by Infinitely Anticipated Backward Stochastic Differential Equations (IABSDEs). Unlike existing frameworks limited to bounded delays, we introduce a generalized…

Optimization and Control · Mathematics 2025-12-22 Guanwei Cheng

Probabilistic control design is founded on the principle that a rational agent attempts to match modelled with an arbitrary desired closed-loop system trajectory density. The framework was originally proposed as a tractable alternative to…

Machine Learning · Computer Science 2023-11-16 Tom Lefebvre

In this paper, we establish a general stochastic maximum principle for optimal control for systems described by a continuous-time Markov regime-switching stochastic recursive utilities model. The control domain is postulated not to be…

Optimization and Control · Mathematics 2019-05-02 Liangquan Zhang , Xun Li

Inverse reinforcement learning aims to infer the reward function that explains expert behavior observed through trajectories of state--action pairs. A long-standing difficulty in classical IRL is the non-uniqueness of the recovered reward:…

Machine Learning · Statistics 2025-12-09 Denis Belomestny , Alexey Naumov , Sergey Samsonov