From deterministic to stochastic: an interpretable stochastic model-free reinforcement learning framework for portfolio optimization

Zitao Song; Yining Wang; Pin Qian; Sifan Song; Frans Coenen; Zhengyong Jiang; Jionglong Su

doi:10.1007/s10489-022-04217-5

From deterministic to stochastic: an interpretable stochastic model-free reinforcement learning framework for portfolio optimization

Zitao Song, Yining Wang, Pin Qian, Sifan Song, Frans Coenen, Zhengyong Jiang^*, Jionglong Su^*

^*Corresponding author for this work

School of AI and Advanced Computing

Research output: Contribution to journal › Article › peer-review

11 Citations (Scopus)

Abstract

As a fundamental problem in algorithmic trading, portfolio optimization aims to maximize the cumulative return by continuously investing in various financial derivatives within a given time period. Recent years have witnessed the transformation from traditional machine learning trading algorithms to reinforcement learning algorithms due to their superior nature of sequential decision making. However, the exponential growth of the imperfect and noisy financial data that is supposedly leveraged by the deterministic strategy in reinforcement learning, makes it increasingly challenging for one to continuously obtain a profitable portfolio. Thus, in this work, we first reconstruct several deterministic and stochastic reinforcement algorithms as benchmarks. On this basis, we introduce a risk-aware reward function to balance the risk and return. Importantly, we propose a novel interpretable stochastic reinforcement learning framework which tailors a stochastic policy parameterized by Gaussian Mixtures and a distributional critic realized by quantiles for the problem of portfolio optimization. In our experiment, the proposed algorithm demonstrates its superior performance on U.S. market stocks with a 63.1% annual rate of return while at the same time reducing the market value max drawdown by 10% when back-testing during the stock market crash around March 2020.

Original language	English
Pages (from-to)	15188-15203
Number of pages	16
Journal	Applied Intelligence
Volume	53
Issue number	12
DOIs	https://doi.org/10.1007/s10489-022-04217-5
Publication status	Published - Jun 2023

Keywords

Deep learning
Portfolio management
Quantitative finance
Reinforcement learning

Access to Document

10.1007/s10489-022-04217-5

Cite this

@article{a0289ded9abc4a189896501dac050200,

title = "From deterministic to stochastic: an interpretable stochastic model-free reinforcement learning framework for portfolio optimization",

abstract = "As a fundamental problem in algorithmic trading, portfolio optimization aims to maximize the cumulative return by continuously investing in various financial derivatives within a given time period. Recent years have witnessed the transformation from traditional machine learning trading algorithms to reinforcement learning algorithms due to their superior nature of sequential decision making. However, the exponential growth of the imperfect and noisy financial data that is supposedly leveraged by the deterministic strategy in reinforcement learning, makes it increasingly challenging for one to continuously obtain a profitable portfolio. Thus, in this work, we first reconstruct several deterministic and stochastic reinforcement algorithms as benchmarks. On this basis, we introduce a risk-aware reward function to balance the risk and return. Importantly, we propose a novel interpretable stochastic reinforcement learning framework which tailors a stochastic policy parameterized by Gaussian Mixtures and a distributional critic realized by quantiles for the problem of portfolio optimization. In our experiment, the proposed algorithm demonstrates its superior performance on U.S. market stocks with a 63.1% annual rate of return while at the same time reducing the market value max drawdown by 10% when back-testing during the stock market crash around March 2020.",

keywords = "Deep learning, Portfolio management, Quantitative finance, Reinforcement learning",

author = "Zitao Song and Yining Wang and Pin Qian and Sifan Song and Frans Coenen and Zhengyong Jiang and Jionglong Su",

note = "Publisher Copyright: {\textcopyright} 2022, The Author(s), under exclusive licence to Springer Science+Business Media, LLC, part of Springer Nature.",

year = "2023",

month = jun,

doi = "10.1007/s10489-022-04217-5",

language = "English",

volume = "53",

pages = "15188--15203",

journal = "Applied Intelligence",

issn = "0924-669X",

number = "12",

}

TY - JOUR

T1 - From deterministic to stochastic

T2 - an interpretable stochastic model-free reinforcement learning framework for portfolio optimization

AU - Song, Zitao

AU - Wang, Yining

AU - Qian, Pin

AU - Song, Sifan

AU - Coenen, Frans

AU - Jiang, Zhengyong

AU - Su, Jionglong

PY - 2023/6

Y1 - 2023/6

N2 - As a fundamental problem in algorithmic trading, portfolio optimization aims to maximize the cumulative return by continuously investing in various financial derivatives within a given time period. Recent years have witnessed the transformation from traditional machine learning trading algorithms to reinforcement learning algorithms due to their superior nature of sequential decision making. However, the exponential growth of the imperfect and noisy financial data that is supposedly leveraged by the deterministic strategy in reinforcement learning, makes it increasingly challenging for one to continuously obtain a profitable portfolio. Thus, in this work, we first reconstruct several deterministic and stochastic reinforcement algorithms as benchmarks. On this basis, we introduce a risk-aware reward function to balance the risk and return. Importantly, we propose a novel interpretable stochastic reinforcement learning framework which tailors a stochastic policy parameterized by Gaussian Mixtures and a distributional critic realized by quantiles for the problem of portfolio optimization. In our experiment, the proposed algorithm demonstrates its superior performance on U.S. market stocks with a 63.1% annual rate of return while at the same time reducing the market value max drawdown by 10% when back-testing during the stock market crash around March 2020.

AB - As a fundamental problem in algorithmic trading, portfolio optimization aims to maximize the cumulative return by continuously investing in various financial derivatives within a given time period. Recent years have witnessed the transformation from traditional machine learning trading algorithms to reinforcement learning algorithms due to their superior nature of sequential decision making. However, the exponential growth of the imperfect and noisy financial data that is supposedly leveraged by the deterministic strategy in reinforcement learning, makes it increasingly challenging for one to continuously obtain a profitable portfolio. Thus, in this work, we first reconstruct several deterministic and stochastic reinforcement algorithms as benchmarks. On this basis, we introduce a risk-aware reward function to balance the risk and return. Importantly, we propose a novel interpretable stochastic reinforcement learning framework which tailors a stochastic policy parameterized by Gaussian Mixtures and a distributional critic realized by quantiles for the problem of portfolio optimization. In our experiment, the proposed algorithm demonstrates its superior performance on U.S. market stocks with a 63.1% annual rate of return while at the same time reducing the market value max drawdown by 10% when back-testing during the stock market crash around March 2020.

KW - Deep learning

KW - Portfolio management

KW - Quantitative finance

KW - Reinforcement learning

UR - http://www.scopus.com/inward/record.url?scp=85141744916&partnerID=8YFLogxK

U2 - 10.1007/s10489-022-04217-5

DO - 10.1007/s10489-022-04217-5

M3 - Article

AN - SCOPUS:85141744916

SN - 0924-669X

VL - 53

SP - 15188

EP - 15203

JO - Applied Intelligence

JF - Applied Intelligence

IS - 12

ER -

From deterministic to stochastic: an interpretable stochastic model-free reinforcement learning framework for portfolio optimization

Abstract

Keywords

Access to Document

Other files and links

Fingerprint

Cite this