Deep Reinforcement Learning for Optimal Stopping with Application in Financial Engineering

05/19/2021
by   Abderrahim Fathan, et al.
0

Optimal stopping is the problem of deciding the right time at which to take a particular action in a stochastic system, in order to maximize an expected reward. It has many applications in areas such as finance, healthcare, and statistics. In this paper, we employ deep Reinforcement Learning (RL) to learn optimal stopping policies in two financial engineering applications: namely option pricing, and optimal option exercise. We present for the first time a comprehensive empirical evaluation of the quality of optimal stopping policies identified by three state of the art deep RL algorithms: double deep Q-learning (DDQN), categorical distributional RL (C51), and Implicit Quantile Networks (IQN). In the case of option pricing, our findings indicate that in a theoretical Black-Schole environment, IQN successfully identifies nearly optimal prices. On the other hand, it is slightly outperformed by C51 when confronted to real stock data movements in a put option exercise problem that involves assets from the S P500 index. More importantly, the C51 algorithm is able to identify an optimal stopping policy that achieves 8 returns than the best of four natural benchmark policies. We conclude with a discussion of our findings which should pave the way for relevant future research.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/24/2021

Solving optimal stopping problems with Deep Q-Learning

We propose a reinforcement learning (RL) approach to model optimal exerc...
research
01/17/2018

The QLBS Q-Learner Goes NuQLear: Fitted Q Iteration, Inverse RL, and Option Portfolios

The QLBS model is a discrete-time option hedging and pricing model that ...
research
12/18/2018

Interpretable Optimal Stopping

Optimal stopping is the problem of deciding when to stop a stochastic sy...
research
03/25/2022

Randomized Policy Optimization for Optimal Stopping

Optimal stopping is the problem of determining when to stop a stochastic...
research
05/09/2022

A Comparative Tutorial of Bayesian Sequential Design and Reinforcement Learning

Reinforcement Learning (RL) is a computational approach to reward-driven...
research
01/11/2019

Deep Learning for Ranking Response Surfaces with Applications to Optimal Stopping Problems

In this paper, we propose deep learning algorithms for ranking response ...
research
09/22/2022

Optimal Stopping with Gaussian Processes

We propose a novel group of Gaussian Process based algorithms for fast a...

Please sign up or login with your details

Forgot password? Click here to reset