---
title: CVA Hedging by Risk-Averse Stochastic-Horizon Reinforcement Learning
url: https://www.ml-quant.com/papers/arxiv/2312.14044/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2312.14044
source_url: http://arxiv.org/abs/2312.14044
featured: 2024-01-03
citations: 2
topic: Derivatives & Volatility
---


# CVA Hedging by Risk-Averse Stochastic-Horizon Reinforcement Learning

The study explores dynamic risk management of potential credit losses on a derivatives portfolio, using recent advancements in risk-averse Reinforcement Learning for option hedging.

- Source: http://arxiv.org/abs/2312.14044
- Identifier: arXiv:2312.14044
- Released: 2023-12-21
- First featured: Quant Letter No. 31 (2024-01-03): https://www.ml-quant.com/issues/2024-01-03/
- Citations (Semantic Scholar): 2
- Published in: not yet
- Topic: Derivatives & Volatility

## Related

- [Robust Risk-Aware Option Hedging](https://www.ml-quant.com/papers/arxiv/2303.15216/): The study highlights the effectiveness of robust risk-aware reinforcement learning in managing risks related to path-dependent financial derivatives, especially in hedging barrier options, proving robust strategies are superior.
- [CVA Hedging by Risk-Averse Stochastic-Horizon Reinforcement Learning](https://www.ml-quant.com/papers/ssrn/4673150/): The study uses risk-averse Reinforcement Learning for managing potential credit losses on a derivatives portfolio, proving its effectiveness through a numerical study for a portfolio consisting of a single FX forward contract.
- [Reinforcement Learning and Deep Stochastic Optimal Control for Final Quadratic Hedging](https://www.ml-quant.com/papers/ssrn/4645455/): The study compares Reinforcement Learning and Deep Trajectory-based Stochastic Optimal Control in hedging a European call option under different market conditions.
- [A Comparison of Reinforcement Learning and Deep Trajectory Based Stochastic Control Agents for Stepwise Mean-Variance Hedging](https://www.ml-quant.com/papers/arxiv/2302.07996/): The research compares the effectiveness of Reinforcement Learning and Deep Trajectory-based Stochastic Optimal Control as data-driven hedging strategies in a simulated environment, offering guidelines for creating autonomous hedging agents.
- [Calibration of Derivative Pricing Models: a Multi-Agent Reinforcement Learning Perspective](https://www.ml-quant.com/papers/arxiv/2203.06865/): The study uses game theory and deep multi-agent reinforcement learning to create models that match market prices of specific options, aiding in understanding local volatility and path-dependence.
- [Hedging Barrier Options Using Reinforcement Learning](https://www.ml-quant.com/papers/ssrn/4566384/): The research indicates that reinforcement learning can be an effective alternative to traditional hedging methods for barrier options, potentially reducing transaction costs due to fewer trades.
