---
title: Calibration of Derivative Pricing Models: a Multi-Agent Reinforcement Learning Perspective
url: https://www.ml-quant.com/papers/arxiv/2203.06865/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2203.06865
source_url: https://arxiv.org/abs/2203.06865
featured: 2023-10-12
citations: 1
topic: Derivatives & Volatility
---


# Calibration of Derivative Pricing Models: a Multi-Agent Reinforcement Learning Perspective

The study uses game theory and deep multi-agent reinforcement learning to create models that match market prices of specific options, aiding in understanding local volatility and path-dependence.

- Source: https://arxiv.org/abs/2203.06865
- Identifier: arXiv:2203.06865
- Released: 2022-03-14
- First featured: Quant Letter No. 20 (2023-10-12): https://www.ml-quant.com/issues/2023-10-12/
- Citations (Semantic Scholar): 1
- Published in: Proceedings of the Fourth ACM International Conference on AI in Finance
- Topic: Derivatives & Volatility

## Related

- [Hedging Barrier Options Using Reinforcement Learning](https://www.ml-quant.com/papers/ssrn/4566384/): The research indicates that reinforcement learning can be an effective alternative to traditional hedging methods for barrier options, potentially reducing transaction costs due to fewer trades.
- [Reinforcement Learning and Deep Stochastic Optimal Control for Final Quadratic Hedging](https://www.ml-quant.com/papers/ssrn/4645455/): The study compares Reinforcement Learning and Deep Trajectory-based Stochastic Optimal Control in hedging a European call option under different market conditions.
- [A Comparison of Reinforcement Learning and Deep Trajectory Based Stochastic Control Agents for Stepwise Mean-Variance Hedging](https://www.ml-quant.com/papers/arxiv/2302.07996/): The research compares the effectiveness of Reinforcement Learning and Deep Trajectory-based Stochastic Optimal Control as data-driven hedging strategies in a simulated environment, offering guidelines for creating autonomous hedging agents.
- [CVA Hedging by Risk-Averse Stochastic-Horizon Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2312.14044/): The study explores dynamic risk management of potential credit losses on a derivatives portfolio, using recent advancements in risk-averse Reinforcement Learning for option hedging.
- [Robust Risk-Aware Option Hedging](https://www.ml-quant.com/papers/arxiv/2303.15216/): The study highlights the effectiveness of robust risk-aware reinforcement learning in managing risks related to path-dependent financial derivatives, especially in hedging barrier options, proving robust strategies are superior.
- [CVA Hedging by Risk-Averse Stochastic-Horizon Reinforcement Learning](https://www.ml-quant.com/papers/ssrn/4673150/): The study uses risk-averse Reinforcement Learning for managing potential credit losses on a derivatives portfolio, proving its effectiveness through a numerical study for a portfolio consisting of a single FX forward contract.
