English
Related papers

Related papers: An Entropy Regularized BSDE Approach to Bermudan O…

200 papers

In the first part of the paper, we study reflected backward stochastic differential equations (RBSDEs) with lower obstacle which is assumed to be right upper-semicontinuous but not necessarily right-continuous. We prove existence and…

Probability · Mathematics 2017-05-11 Miryana Grigorova , Peter Imkeller , Elias Offen , Youssef Ouknine , Marie-Claire Quenez

This paper studies the continuous-time reinforcement learning (RL) for optimal switching problems across multiple regimes. We consider a type of exploratory formulation under entropy regularization where the agent randomizes both the timing…

Optimization and Control · Mathematics 2025-12-23 Yijie Huang , Mengge Li , Xiang Yu , Zhou Zhou

We consider a reflected backward stochastic differential equations with default time and an optional barrier in a filtration generated by a one-dimensional Brownian motion and a defaultable process. We suppose that the barrier have…

Probability · Mathematics 2026-05-07 Badr Elmansouri , Mohamed El Otmani

We consider the non-linear optimal multiple stopping problem under general conditions on the non-linear evaluation operators, which might depend on two time indices: the time of evaluation/assessment and the horizon (when the reward or loss…

Optimization and Control · Mathematics 2025-04-21 Miryana Grigorova , Marie-Claire Quenez , Peng Yuan

A Budgeted Markov Decision Process (BMDP) is an extension of a Markov Decision Process to critical applications requiring safety constraints. It relies on a notion of risk implemented in the shape of a cost signal constrained to lie below…

Machine Learning · Computer Science 2019-05-29 Nicolas Carrara , Edouard Leurent , Romain Laroche , Tanguy Urvoy , Odalric-Ambrym Maillard , Olivier Pietquin

We introduce a new probabilistic method for solving a class of impulse control problems based on their representations as Backward Stochastic Differential Equations (BSDEs for short) with constrained jumps. As an example, our method is used…

Computational Finance · Quantitative Finance 2015-03-17 Marie Bernhart , Huyên Pham , Peter Tankov , Xavier Warin

The entropy regularization is inspired by information entropy from machine learning and the ideas of exploration and exploitation in reinforcement learning, which appears in the control problem to design an approximating algorithm for the…

Optimization and Control · Mathematics 2024-11-21 Ziyue Chen , Qi Zhang

We study an optimal control problem on infinite time horizon with semimartingale strategies, random coefficients and regime switching. The value function and the optimal strategy can be characterized in terms of three systems of backward…

Optimization and Control · Mathematics 2026-02-27 Xinman Cheng , Guanxing Fu , Xiaonyu Xia

We propose and analyze a continuous-time robust reinforcement learning framework for optimal stopping under ambiguity. In this framework, an agent chooses a robust exploratory stopping time motivated by two objectives: robust…

Optimization and Control · Mathematics 2026-04-17 Junyan Ye , Hoi Ying Wong , Kyunghyun Park

We study the optimal stopping problem for a monotonous dynamic risk measure induced by a BSDE with jumps in the Markovian case. We show that the value function is a viscosity solution of an obstacle problem for a partial…

Optimization and Control · Mathematics 2014-07-01 Roxana Dumitrescu , Marie-Claire Quenez , Agnès Sulem

This paper solves a recursive optimal stopping problem with Poisson stopping constraints using the penalized backward stochastic differential equation (PBSDE) with jumps. Stopping in this problem is only allowed at Poisson random…

Optimization and Control · Mathematics 2025-05-20 Gechun Liang , Wei Wei , Zhen Wu , Zhenda Xu

We prove existence and uniqueness of the reflected backward stochastic differential equation's (RBSDE) solution with a lower obstacle which is assumed to be right upper-semicontinuous but not necessarily right-continuous in a filtration…

Probability · Mathematics 2018-12-20 Brahim Baadi , Youssef Ouknine

Recent advances in continuous-time optimal stopping have been driven by entropy-regularized formulations of randomized stopping problems, with most existing approaches relying on partial differential equation methods. In this paper, we…

Computational Finance · Quantitative Finance 2026-02-23 Daniel Chee , Noufel Frikha , Libo Li

We address a general optimal switching problem over finite horizon for a stochastic system described by a differential equation driven by Brownian motion. The main novelty is the fact that we allow for infinitely many modes (or regimes,…

Optimization and Control · Mathematics 2019-08-07 Marco Fuhrman , Marie-Amélie Morlais

This paper studies the continuous-time reinforcement learning for stochastic singular control with the application to an infinite-horizon irreversible reinsurance problems. The singular control is equivalently characterized as a pair of…

Optimization and Control · Mathematics 2025-12-03 Zongxia Liang , Xiaodong Luo , Xiang Yu

Motivated by the trade-off between exploitation and exploration in reinforcement learning, we study a continuous-time entropy-regularized mean variance portfolio selection problem in the presence of jumps. We propose an exploratory SDE for…

Optimization and Control · Mathematics 2025-02-26 Christian Bender , Nguyen Tran Thuan

We investigate the optimal reinsurance problem under the criterion of maximizing the expected utility of terminal wealth when the insurance company has restricted information on the loss process. We propose a risk model with claim arrival…

Mathematical Finance · Quantitative Finance 2020-05-15 Matteo Brachetta , Claudia Ceci

Constrained decision-making is essential for designing safe policies in real-world control systems, yet simulated environments often fail to capture real-world adversities. We consider the problem of learning a policy that will maximize the…

Machine Learning · Computer Science 2026-02-10 Sourav Ganguly , Kishan Panaganti , Arnob Ghosh , Adam Wierman

This paper bridges reinforcement learning (RL) and risk-sensitive stochastic control by introducing a tractable exploration mechanism for policy search in risk-sensitive portfolio management, with known and unknown model parameters, that…

Portfolio Management · Quantitative Finance 2026-03-03 Sebastien Lleo , Wolfgang Runggaldier

We show a concise extension of the monotone stability approach to backward stochastic differential equations (BSDEs) that are jointly driven by a Brownian motion and a random measure for jumps, which could be of infinite activity with a…

Probability · Mathematics 2019-11-21 Dirk Becherer , Martin Büttner , Klebert Kentia