---
title: Deep Reinforcement Learning Algorithms for Option Hedging
url: https://www.ml-quant.com/papers/arxiv/2504.05521/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2504.05521
source_url: http://arxiv.org/abs/2504.05521v1
featured: 2025-04-09
citations: 7
topic: Derivatives & Volatility
---


# Deep Reinforcement Learning Algorithms for Option Hedging

A comparison of eight Deep Reinforcement Learning algorithms for dynamic hedging found that Monte Carlo Policy Gradient and Proximal Policy Optimization performed best, with the former outperforming the Black-Scholes delta hedge baseline.

- Source: http://arxiv.org/abs/2504.05521v1
- Identifier: arXiv:2504.05521
- Released: 2025-04-07
- First featured: Quant Letter No. 92 (2025-04-09): https://www.ml-quant.com/issues/2025-04-09/
- Citations (Semantic Scholar): 7
- Published in: not yet
- Topic: Derivatives & Volatility

## Related

- [Numerical analysis of a particle system for the calibrated Heston-type local stochastic volatility model](https://www.ml-quant.com/papers/arxiv/2504.14343/): The article presents a Monte Carlo method for simulating the Heston-type local stochastic volatility model, addressing drift and diffusion coefficient challenges and proving a strong chaos propagation under certain conditions.
- [To Hedge or Not to Hedge: Optimal Strategies for Stochastic Trade Flow Management](https://www.ml-quant.com/papers/arxiv/2503.02496/): The paper proposes using reinforcement learning methods to manage stochastic trade flows, offering an alternative to traditional grid-based numerical PDE techniques.
- [A Risk Sensitive Contract-unified Reinforcement Learning Approach for Option Hedging](https://www.ml-quant.com/papers/arxiv/2411.09659/): The paper proposes a risk-sensitive reinforcement learning approach for dynamic hedging of options, reducing tail risk using historical market data.
- [ARL-Based Multi-Action Market Making with Hawkes Processes and Variable Volatility](https://www.ml-quant.com/papers/doi/10-1145-3677052-3698695/): The study combines Adversarial Reinforcement Learning, Hawkes Processes, and variable volatility to enhance market-making strategies, showing improved adaptability in high-volatility conditions and better market simulations.
- [Solving The Dynamic Volatility Fitting Problem: A Deep Reinforcement Learning Approach](https://www.ml-quant.com/papers/arxiv/2410.11789/): The article discusses the use of Deep Reinforcement Learning in solving volatility issues in equity derivatives, showing its effectiveness and adaptability in handling complex functions and online learning.
- [EX-DRL: Hedging Against Heavy Losses with EXtreme Distributional Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2408.12446/): The article introduces EXtreme DRL (EX-DRL), a new method to improve the accuracy of extreme quantile predictions in Distributional Reinforcement Learning, improving financial risk management.
