English
Related papers

Related papers: A Comparison of Reinforcement Learning and Deep Tr…

200 papers

The stochastic control problem of optimal market making is among the central problems in quantitative finance. In this paper, a deep reinforcement learning-based controller is trained on a weakly consistent, multivariate Hawkes…

General Finance · Quantitative Finance 2022-07-21 Bruno Gašperov , Zvonko Kostanjčar

Stock trading strategy plays a crucial role in investment companies. However, it is challenging to obtain optimal strategy in the complex and dynamic stock market. We explore the potential of deep reinforcement learning to optimize stock…

Machine Learning · Computer Science 2022-08-02 Xiao-Yang Liu , Zhuoran Xiong , Shan Zhong , Hongyang Yang , Anwar Walid

This paper studies empirical deep hedging for S&P 500 index options under a local downside-shortfall reward. It moves beyond performance comparison by asking what the learned hedge does, when it fails, and whether it can be made auditable.…

Risk Management · Quantitative Finance 2026-05-22 Kirill Zernikov

This paper studies a discrete-time mean-variance model based on reinforcement learning. Compared with its continuous-time counterpart in \cite{zhou2020mv}, the discrete-time model makes more general assumptions about the asset's return…

Mathematical Finance · Quantitative Finance 2023-12-27 Xiangyu Cui , Xun Li , Yun Shi , Si Zhao

Deep hedging trains neural networks to manage derivative risk under market frictions, but produces hedge ratios with no measure of model confidence -- a significant barrier to deployment. We introduce uncertainty quantification to the deep…

Computational Finance · Quantitative Finance 2026-03-12 Manan Poddar

We investigate the use of path signatures in a machine learning context for hedging exotic derivatives under non-Markovian stochastic volatility models. In a deep learning setting, we use signatures as features in feedforward neural…

Machine Learning · Statistics 2025-08-12 Eduardo Abi Jaber , Louis-Amand Gérard

This article leverages deep reinforcement learning (DRL) to hedge American put options, utilizing the deep deterministic policy gradient (DDPG) method. The agents are first trained and tested with Geometric Brownian Motion (GBM) asset paths…

Risk Management · Quantitative Finance 2024-05-14 Reilly Pickard , Finn Wredenhagen , Julio DeJesus , Mario Schlener , Yuri Lawryshyn

Stock trading is one of the popular ways for financial management. However, the market and the environment of economy is unstable and usually not predictable. Furthermore, engaging in stock trading requires time and effort to analyze,…

Machine Learning · Computer Science 2025-05-20 Yunfei Luo , Zhangqi Duan

The objectives of option hedging/trading extend beyond mere protection against downside risks, with a desire to seek gains also driving agent's strategies. In this study, we showcase the potential of robust risk-aware reinforcement learning…

Computational Finance · Quantitative Finance 2023-12-27 David Wu , Sebastian Jaimungal

We consider a sequential decision making problem where the agent faces the environment characterized by the stochastic discrete events and seeks an optimal intervention policy such that its long-term reward is maximized. This problem exists…

Machine Learning · Computer Science 2022-12-29 Chao Qu , Xiaoyu Tan , Siqiao Xue , Xiaoming Shi , James Zhang , Hongyuan Mei

We propose a new risk sensitive reinforcement learning approach for the dynamic hedging of options. The approach focuses on the minimization of the tail risk of the final P&L of the seller of an option. Different from most existing…

Risk Management · Quantitative Finance 2024-11-15 Xianhua Peng , Xiang Zhou , Bo Xiao , Yi Wu

Option pricing theory, such as the Black and Scholes (1973) model, provides an explicit solution to construct a strategy that perfectly hedges an option in a continuous-time setting. In practice, however, trading occurs in discrete time and…

Mathematical Finance · Quantitative Finance 2025-05-30 Pierre Brugière , Gabriel Turinici

This work provides a Deep Reinforcement Learning approach to solving a periodic review inventory control system with stochastic vendor lead times, lost sales, correlated demand, and price matching. While this dynamic program has…

Machine Learning · Computer Science 2022-11-30 Dhruv Madeka , Kari Torkkola , Carson Eisenach , Anna Luo , Dean P. Foster , Sham M. Kakade

Portfolio optimization involves determining the optimal allocation of portfolio assets in order to maximize a given investment objective. Traditionally, some form of mean-variance optimization is used with the aim of maximizing returns…

Artificial Intelligence · Computer Science 2024-03-26 Fernando Acero , Parisa Zehtabi , Nicolas Marchesotti , Michael Cashmore , Daniele Magazzeni , Manuela Veloso

This paper investigates the deep hedging framework, based on reinforcement learning (RL), for the dynamic hedging of swaptions, contrasting its performance with traditional sensitivity-based rho-hedging. We design agents under three…

Risk Management · Quantitative Finance 2025-12-09 Zaniar Ahmadi , Frédéric Godin

Dynamic portfolio optimization is the process of sequentially allocating wealth to a collection of assets in some consecutive trading periods, based on investors' return-risk profile. Automating this process with machine learning remains a…

Machine Learning · Computer Science 2019-01-28 Pengqian Yu , Joon Sern Lee , Ilya Kulyatin , Zekun Shi , Sakyasingha Dasgupta

We adopt Deep Reinforcement Learning algorithms to design trading strategies for continuous futures contracts. Both discrete and continuous action spaces are considered and volatility scaling is incorporated to create reward functions which…

Computational Finance · Quantitative Finance 2019-11-25 Zihao Zhang , Stefan Zohren , Stephen Roberts

A properly designed controller can help improve the quality of experimental measurements or force a dynamical system to follow a completely new time-evolution path. Recent developments in deep reinforcement learning have made steep advances…

Statistical Mechanics · Physics 2025-02-26 Ruslan Mukhamadiarov

In this thesis, we develop a comprehensive account of the expressive power, modelling efficiency, and performance advantages of so-called trading agents (i.e., Deep Soft Recurrent Q-Network (DSRQN) and Mixture of Score Machines (MSM)),…

Portfolio Management · Quantitative Finance 2019-09-23 Angelos Filos

We present an actor-critic-type reinforcement learning algorithm for solving the problem of hedging a portfolio of financial instruments such as securities and over-the-counter derivatives using purely historic data. The key characteristics…

Computational Finance · Quantitative Finance 2024-06-26 Hans Buehler , Phillip Murray , Ben Wood