English
Related papers

Related papers: An Entropy Regularized BSDE Approach to Bermudan O…

200 papers

We study automated intrusion prevention using reinforcement learning. In a novel approach, we formulate the problem of intrusion prevention as an optimal stopping problem. This formulation allows us insight into the structure of the optimal…

Artificial Intelligence · Computer Science 2024-04-23 Kim Hammar , Rolf Stadler

In this paper we introduce and study the concept of optimal and surely optimal dual martingales in the context of dual valuation of Bermudan options, and outline the development of new algorithms in this context. We provide a…

Computational Finance · Quantitative Finance 2012-02-14 John Schoenmakers , Junbo Huang , Jianing Zhang

The aim of this study is to devise numerical methods for dealing with very high-dimensional Bermudan-style derivatives. For such problems, we quickly see that we can at best hope for price bounds, and we can only use a simulation approach.…

Computational Finance · Quantitative Finance 2016-01-06 L. C. G. Rogers

We study a speculative trading problem within the exploratory reinforcement learning (RL) framework of Wang et al. [2020]. The problem is formulated as a sequential optimal stopping problem over entry and exit times under general utility…

Mathematical Finance · Quantitative Finance 2026-04-03 Yun Zhao , Alex S. L. Tse , Harry Zheng

Sample-efficient exploration is crucial not only for discovering rewarding experiences but also for adapting to environment changes in a task-agnostic fashion. A principled treatment of the problem of optimal input synthesis for system…

Machine Learning · Computer Science 2019-10-10 Matthias Schultheis , Boris Belousov , Hany Abdulsamad , Jan Peters

We develop a continuous-time reinforcement learning framework for a class of singular stochastic control problems without entropy regularization. The optimal singular control is characterized as the optimal singular control law, which is a…

Optimization and Control · Mathematics 2026-05-14 Zongxia Liang , Xiaodong Luo , Xiang Yu

In this paper we study by probabilistic techniques the convergence of the value function for a two-scale, infinite-dimensional, stochastic controlled system as the ratio between the two evolution speeds diverges. The value function is…

Optimization and Control · Mathematics 2018-09-12 Giuseppina Guatteri , Gianmario Tessitore

We study the exploration problem with approximate linear action-value functions in episodic reinforcement learning under the notion of low inherent Bellman error, a condition normally employed to show convergence of approximate value…

Machine Learning · Computer Science 2020-06-30 Andrea Zanette , Alessandro Lazaric , Mykel Kochenderfer , Emma Brunskill

We study reflected backward stochastic differential equation (RBSDEs) on the probability space equipped with a Brownian motion. The main novelty of the paper lies in fact that we consider the following weak assumptions on the data: barriers…

Probability · Mathematics 2022-09-27 Tomasz Klimsiak , Maurycy Rzymowski

We study an optimal control problem on infinite horizon for a controlled stochastic differential equation driven by Brownian motion, with a discounted reward functional. The equation may have memory or delay effects in the coefficients,…

Optimization and Control · Mathematics 2017-10-19 F. Confortola , A. Cosso , M. Fuhrman

This paper investigates an optimal consumption-investment problem featuring recursive utility via Tsallis relative entropy. We establish a fundamental connection between this optimization problem and a quadratic backward stochastic…

Mathematical Finance · Quantitative Finance 2025-09-26 Xueying Huang , Peng Luo , Dejian Tian

In this paper we first investigate zero-sum two-player stochastic differential games with reflection with the help of theory of Reflected Backward Stochastic Differential Equations (RBSDEs). We will establish the dynamic programming…

Probability · Mathematics 2008-09-30 Rainer Buckdahn , Juan Li

This paper proposes a relaxed control regularization with general exploration rewards to design robust feedback controls for multi-dimensional continuous-time stochastic exit time problems. We establish that the regularized control problem…

Optimization and Control · Mathematics 2021-07-26 Christoph Reisinger , Yufei Zhang

We study mean-field games of optimal stopping (OS-MFGs) and introduce an entropy-regularized framework to enable learning-based solution methods. By utilizing randomized stopping times, we reformulate the OS-MFG as a mean-field game of…

Optimization and Control · Mathematics 2025-09-24 Jodi Dianetti , Roxana Dumitrescu , Giorgio Ferrari , Renyuan Xu

We consider the problem of learning the optimal policy for Markov decision processes with safety constraints. We formulate the problem in a reach-avoid setup. Our goal is to design online reinforcement learning algorithms that ensure safety…

Machine Learning · Computer Science 2026-01-21 Abhijit Mazumdar , Rafal Wisniewski , Manuela L. Bujorianu

In this paper, we introduce a non-linear Snell envelope which at each time represents the maximal value that can be achieved by stopping a BSDE with constrained jumps. We establish the existence of the Snell envelope by employing a…

Optimization and Control · Mathematics 2023-09-01 Magnus Perninge

Considering uncertainties and disturbances is an important, yet challenging, step in successful decision making. The problem becomes more challenging in safety-constrained environments. In this paper, we propose a robust and safe trajectory…

Systems and Control · Electrical Eng. & Systems 2022-03-29 Hassan Almubarak , Evangelos A. Theodorou , Nader Sadegh

We consider a classical finite horizon optimal control problem for continuous-time pure jump Markov processes described by means of a rate transition measure depending on a control parameter and controlled by a feedback law. For this class…

Probability · Mathematics 2015-01-20 Elena Bandini , Marco Fuhrman

We investigate an optimal stopping problem for the expected value of a discounted payoff on a regime-switching geometric Brownian motion under two constraints on the possible stopping times: only at exogenous random times and only during a…

Probability · Mathematics 2024-11-20 Takuji Arai , Masahiko Takenaka

We construct an aggregated version of the value processes associated with stochastic control problems, where the criterion to optimise is given by solutions to semi-martingale backward stochastic differential equations (BSDEs). The results…

Probability · Mathematics 2025-07-03 Dylan Possamaï , Marco Rodrigues , Alexandros Saplaouras
‹ Prev 1 3 4 5 6 7 10 Next ›