Pith. sign in

REVIEW 1 cited by

Capturing Financial markets to apply Deep Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1907.04373 v3 pith:SCQ4I5XW submitted 2019-07-09 q-fin.CP cs.LG

Capturing Financial markets to apply Deep Reinforcement Learning

classification q-fin.CP cs.LG
keywords financialmarketsdeeplearningmarketmodelreinforcementcapture
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

In this paper we explore the usage of deep reinforcement learning algorithms to automatically generate consistently profitable, robust, uncorrelated trading signals in any general financial market. In order to do this, we present a novel Markov decision process (MDP) model to capture the financial trading markets. We review and propose various modifications to existing approaches and explore different techniques like the usage of technical indicators, to succinctly capture the market dynamics to model the markets. We then go on to use deep reinforcement learning to enable the agent (the algorithm) to learn how to take profitable trades in any market on its own, while suggesting various methodology changes and leveraging the unique representation of the FMDP (financial MDP) to tackle the primary challenges faced in similar works. Through our experimentation results, we go on to show that our model could be easily extended to two very different financial markets and generates a positively robust performance in all conducted experiments.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Deep Reinforcement Learning for Reliability Based Bi-Objective Portfolio Optimization

    cs.LG 2026-07 conditional novelty 4.5

    A PPO agent with reliability-shaped rewards and GARCH–EVT–t-copula scenarios matches or approaches NSGA-II on global equity indices across pre/COVID/post-COVID regimes under variance, CVaR, and EVaR.