English
Related papers

Related papers: An Entropy Regularized BSDE Approach to Bermudan O…

200 papers

This paper extends our previous work to continuous-time optimal stopping, focusing on American options in an exploratory setting. Our first contribution is an entropy-regularized penalization scheme, inspired by classical penalization…

Mathematical Finance · Quantitative Finance 2026-03-04 Daniel Chee , Noufel Frikha , Libo Li

Under the assumption of no-arbitrage, the pricing of American and Bermudan options can be casted into optimal stopping problems. We propose a new adaptive simulation based algorithm for the numerical solution of optimal stopping problems in…

Probability · Mathematics 2009-09-29 Daniel Egloff , Michael Kohler , Nebojsa Todorovic

We investigate an entropy-regularized reinforcement learning (RL) approach to optimal stopping problems motivated by real option models. Classical stopping rules are strict and non-randomized, limiting natural exploration in RL settings. To…

Optimization and Control · Mathematics 2026-02-18 Jodi Dianetti , Giorgio Ferrari , Renyuan Xu

In this paper, we study the optimal stopping problem in the so-called exploratory framework, in which the agent takes actions randomly conditioning on current state and an entropy-regularized term is added to the reward functional. Such a…

Optimization and Control · Mathematics 2023-09-04 Yuchao Dong

This paper aims to establish an entropy-regularized value-based reinforcement learning method that can ensure the monotonic improvement of policies at each policy update. Unlike previously proposed lower-bounds on policy improvement in…

Machine Learning · Computer Science 2020-08-26 Lingwei Zhu , Takamitsu Matsubara

This paper explores continuous-time and state-space optimal stopping problems from a reinforcement learning perspective. We begin by formulating the stopping problem using randomized stopping times, where the decision maker's control is…

Optimization and Control · Mathematics 2026-03-12 Jodi Dianetti , Giorgio Ferrari , Renyuan Xu

We propose the Compound BSDE method, a fully forward, deep-learning-based approach for solving a broad class of problems in financial mathematics, including optimal stopping. The method is based on a reformulation of option pricing problems…

Computational Finance · Quantitative Finance 2026-02-02 Zhipeng Huang , Cornelis W. Oosterlee

Solving optimal stopping problems by backward induction in high dimensions is often very complex since the computation of conditional expectations is required. Typically, such computations are based on regression, a method that suffers from…

Probability · Mathematics 2022-05-19 Martin Redmann

We study the optimal stopping problem for dynamic risk measures represented by Backward Stochastic Differential Equations (BSDEs) with jumps and its relation with reflected BSDEs (RBSDEs). We first provide general existence, uniqueness and…

Probability · Mathematics 2013-01-01 Marie-Claire Quenez , AgnÈs Sulem

The optimal stopping problem is one of the core problems in financial markets, with broad applications such as pricing American and Bermudan options. The deep BSDE method [Han, Jentzen and E, PNAS, 115(34):8505-8510, 2018] has shown great…

Probability · Mathematics 2023-08-28 Chengfan Gao , Siping Gao , Ruimeng Hu , Zimu Zhu

We study the problem of optimal portfolio selection under stochastic volatility within a continuous time reinforcement learning framework with portfolio constraints. Exploration is modeled through entropy-regularized relaxed controls, where…

Mathematical Finance · Quantitative Finance 2026-04-27 Thai Nguyen , Pertiny Nkuize

We consider a non-Markovian optimal stopping problem on finite horizon. We prove that the value process can be represented by means of a backward stochastic differential equation (BSDE), defined on an enlarged probability space, containing…

Probability · Mathematics 2015-02-20 Marco Fuhrman , Huyên Pham , Federica Zeni

The aim of this paper is to study an optimal stopping problem for dynamic risk measures induced by backward stochastic differential equations with jumps and delayed generator. Firstly, we connect the value function of this problem to…

Probability · Mathematics 2021-10-06 Tuo Navegue , Auguste Aman

We introduce a new formulation of reflected BSDEs and doubly reflected BSDEs associated with irregular obstacles. In the first part of the paper, we consider an extension of the classical optimal stopping problem over a larger set of…

Probability · Mathematics 2023-03-31 Ihsan Arharas , Youssef Ouknine

In a Markovian framework, we consider the problem of finding the minimal initial value of a controlled process allowing to reach a stochastic target with a given level of expected loss. This question arises typically in approximate hedging…

Optimization and Control · Mathematics 2017-04-06 Géraldine Bouveret , Jean-François Chassagneux

We address an optimal stopping problem over the set of Bermudan-type strategies $\Theta$ (which we understand in a more general sense than the stopping strategies for Bermudan options in finance) and with non-linear operators (non-linear…

Optimization and Control · Mathematics 2023-01-27 Miryana Grigorova , Marie-Claire Quenez , Peng Yuan

Exploration is a crucial and distinctive aspect of reinforcement learning (RL) that remains a fundamental open problem. Several methods have been proposed to tackle this challenge. Commonly used methods inject random noise directly into the…

Machine Learning · Computer Science 2024-11-06 Sebastian Griesbach , Carlo D'Eramo

We study the optimal investment stopping problem in both continuous and discrete case, where the investor needs to choose the optimal trading strategy and optimal stopping time concurrently to maximize the expected utility of terminal…

Mathematical Finance · Quantitative Finance 2020-05-01 Dingqian Sun

We define a class of reflected backward stochastic differential equation (RBSDE) driven by a marked point process (MPP) and a Brownian motion, where the solution is constrained to stay above a given c\`adl\`ag process. The MPP is only…

Probability · Mathematics 2017-09-28 Nahuel Foresta

Policy-based reinforcement learning methods suffer from the policy collapse problem. We find valued-based reinforcement learning methods with {\epsilon}-greedy mechanism are capable of enjoying three characteristics, Closed-form Diversity,…

Machine Learning · Computer Science 2021-06-03 Changnan Xiao , Haosen Shi , Jiajun Fan , Shihong Deng
‹ Prev 1 2 3 10 Next ›