Pith. sign in

REVIEW 1 cited by

Deep Hedging of Derivatives Using Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2103.16409 v1 pith:RJBYJVC7 submitted 2021-03-29 q-fin.CP cs.CEcs.LG

classification q-fin.CPcs.CEcs.LG
keywords approachhedgingcostlearningassetpricereinforcementused
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This paper shows how reinforcement learning can be used to derive optimal hedging strategies for derivatives when there are transaction costs. The paper illustrates the approach by showing the difference between using delta hedging and optimal hedging for a short position in a call option when the objective is to minimize a function equal to the mean hedging cost plus a constant times the standard deviation of the hedging cost. Two situations are considered. In the first, the asset price follows a geometric Brownian motion. In the second, the asset price follows a stochastic volatility process. The paper extends the basic reinforcement learning approach in a number of ways. First, it uses two different Q-functions so that both the expected value of the cost and the expected value of the square of the cost are tracked for different state/action combinations. This approach increases the range of objective functions that can be used. Second, it uses a learning algorithm that allows for continuous state and action space. Third, it compares the accounting P&L approach (where the hedged position is valued at each step) and the cash flow approach (where cash inflows and outflows are used). We find that a hybrid approach involving the use of an accounting P&L approach that incorporates a relatively simple valuation model works well. The valuation model does not have to correspond to the process assumed for the underlying asset price.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Model-Free Deep Hedging with Transaction Costs and Light Data Requirements

    q-fin.MF 2025-05 conditional novelty 5.0 of 10

    A neural network hedger trained on about 256 simulated paths beats Black-Scholes and Leland hedging at high transaction costs in a synthetic GBM market, but not at low costs and not on real S&P 500 data.

Pith tools