ML-QuantSubscribe

arXivDerivatives & Volatility

Adversarial Training for Deep Hedging in Nonstationary Markets

WRAP combines Wasserstein reweighting and optimal-transport perturbations in a distributionally robust framework to make deep hedging policies robust to nonstationarity and distributional drift in market conditions.

Featured in No. 134 on 9 Oct 2026 · 2 days after release

Loss landscapes across deep-hedging training schemes from Table 1 : (A) ERM-W , (B) \phi -W , (C) OT-W , and (D) WRAP-W
Figure 1 : Loss landscapes across deep-hedging training schemes from Table 1 : (A) ERM-W , (B) \phi -W , (C) OT-W , and (D) WRAP-W . Each surface depicts the mean hedging loss of a separately trained network on the same reference trajectories X_{n} . The axes show terminal moneyness and the root me…
Released
7 Oct 2026
First featured
No. 134 · 9 Oct 2026
Published in
Not yet, as far as Semantic Scholar knows
Fanfare
3 of 5
Identifier
arXiv:2610.07162
Authors
Philipp J. Schneider et al.

Abstract

From arXiv (CC0).

Deep hedging learns trading policies from historical or simulated market trajectories, yet under nonstationarity these training paths may not represent future market conditions. We propose WRAP (Wasserstein-Reweighting Adversarial Perturbation), a drift-aware adversarial training framework derived from a two-budget distributionally robust optimization (DRO) formulation. The formulation is anchored to a weighted empirical reference distribution whose fixed baseline weights are chosen to balance sampling uncertainty against temporal drift. Around this reference distribution, the ambiguity set addresses two complementary forms of distributional misspecification by allowing an adversary to reweight the observed trajectories subject to a $φ$-divergence constraint and perturb their paths subject to an optimal-transport (OT) constraint. We derive a joint first-order expansion in which the leading-order increase over the nominal expected loss decomposes into a reweighting contribution determined by the dispersion of hedging losses across trajectories and a transport contribution determined by the sensitivity of the loss to path perturbations. This expansion yields an explicit finite-dimensional adversarial attack that replaces the distributional inner supremum with a tractable first-order approximation. Across stationary and nonstationary Heston dynamics and a generalized affine diffusion (GAD), the experiments show complementary benefits from reweighting and transport, with joint adversarial training providing the largest gains under nonstationarity.

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page