---
title: Adversarial Training for Deep Hedging in Nonstationary Markets
url: https://www.ml-quant.com/papers/arxiv/2610.07162/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-10-09
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2610.07162
source_url: https://arxiv.org/abs/2610.07162
featured: 2026-10-09
citations: unknown
topic: Derivatives & Volatility
---


# Adversarial Training for Deep Hedging in Nonstationary Markets

WRAP combines Wasserstein reweighting and optimal-transport perturbations in a distributionally robust framework to make deep hedging policies robust to nonstationarity and distributional drift in market conditions.

- Source: https://arxiv.org/abs/2610.07162
- Identifier: arXiv:2610.07162
- Released: 2026-10-07
- First featured: Quant Letter No. 134 (2026-10-09): https://www.ml-quant.com/issues/2026-10-09/
- Citations (Semantic Scholar): not tracked
- Published in: not yet
- Topic: Derivatives & Volatility
- Authors: Philipp J. Schneider, Lukas Looser, Antoine Garin, Shuhan Liu, Daniel Kuhn

## Abstract (arXiv, CC0)

Deep hedging learns trading policies from historical or simulated market trajectories, yet under nonstationarity these training paths may not represent future market conditions. We propose WRAP (Wasserstein-Reweighting Adversarial Perturbation), a drift-aware adversarial training framework derived from a two-budget distributionally robust optimization (DRO) formulation. The formulation is anchored to a weighted empirical reference distribution whose fixed baseline weights are chosen to balance sampling uncertainty against temporal drift. Around this reference distribution, the ambiguity set addresses two complementary forms of distributional misspecification by allowing an adversary to reweight the observed trajectories subject to a $φ$-divergence constraint and perturb their paths subject to an optimal-transport (OT) constraint. We derive a joint first-order expansion in which the leading-order increase over the nominal expected loss decomposes into a reweighting contribution determined by the dispersion of hedging losses across trajectories and a transport contribution determined by the sensitivity of the loss to path perturbations. This expansion yields an explicit finite-dimensional adversarial attack that replaces the distributional inner supremum with a tractable first-order approximation. Across stationary and nonstationary Heston dynamics and a generalized affine diffusion (GAD), the experiments show complementary benefits from reweighting and transport, with joint adversarial training providing the largest gains under nonstationarity.

## Related

- [Model-Free Deep Hedging with Transaction Costs and Light Data Requirements](https://www.ml-quant.com/papers/arxiv/2505.22836/): The research shows that a neural network trained with just 256 trajectories can outperform the Black & Scholes formula and the Leland model in the Geometric Brownian Motion framework, indicating potential for real-time financial series application.
- [Deep Hedging with Options Using the Implied Volatility Surface](https://www.ml-quant.com/papers/arxiv/2504.06208/): A new deep hedging framework for index option portfolios, which includes surface-informed decisions and transaction costs, has been proposed and outperforms traditional methods in both simulated and historical data from 1996 to 2020.
- [Deep Hedging of Green PPAs in Electricity Markets](https://www.ml-quant.com/papers/arxiv/2503.13056/): The paper introduces a 'deep hedging' approach using machine learning to develop hedging strategies in power markets, specifically for Green Power Purchase Agreements, which are subject to price and weather risks.
- [Deep Hedging Bermudan Swaptions](https://www.ml-quant.com/papers/arxiv/2411.10079/): The article introduces a new method for Bermudan swaption hedging using the deep hedging framework, improving profit and loss management.
- [Enhancing Deep Hedging of Options with Implied Volatility Surface Feedback Information](https://www.ml-quant.com/papers/arxiv/2407.21138/): A new hedging strategy for S&P 500 options is introduced, using a unique reinforcement learning algorithm and hybrid neural network, which performs better than traditional benchmarks in tests and simulations.
- [Is the difference between deep hedging and delta hedging a statistical arbitrage?](https://www.ml-quant.com/papers/arxiv/2407.14736/): The research compares deep hedging and delta hedging in a GARCH-based market model, suggesting that the difference between the two can be a statistical arbitrage if the risk measure doesn't adequately consider negative outcomes.
