REVIEW 1 cited by
Deep Reinforcement Learning for Online Optimal Execution Strategies
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper tackles the challenge of learning non-Markovian optimal execution strategies in dynamic financial markets. We introduce a novel actor-critic algorithm based on Deep Deterministic Policy Gradient (DDPG) to address this issue, with a focus on transient price impact modeled by a general decay kernel. Through numerical experiments with various decay kernels, we show that our algorithm successfully approximates the optimal execution strategy. Additionally, the proposed algorithm demonstrates adaptability to evolving market conditions, where parameters fluctuate over time. Our findings also show that modern reinforcement learning algorithms can provide a solution that reduces the need for frequent and inefficient human intervention in optimal execution tasks.
Forward citations
Cited by 1 Pith paper
-
Can Reinforcement Learning Efficiently Discover Price Manipulation?
Under intermediate volatility and limited samples, model-free DDPG finds dynamic-arbitrage strategies more reliably than SLSQP run on noisily estimated Almgren-Chriss impact parameters, even though the latter knows th...
Discussion (0). Continue with ORCID to comment.