---
title: Taming the Greeks: Option Portfolios with Inductive Biases
url: https://www.ml-quant.com/papers/arxiv/2609.33767/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-10-02
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2609.33767
source_url: https://arxiv.org/abs/2609.33767
featured: 2026-10-02
citations: unknown
topic: Portfolio & Allocation
---


# Taming the Greeks: Option Portfolios with Inductive Biases

A deep learning framework for options trading enforces portfolio-level delta neutrality and other risk constraints during training, improving risk-adjusted returns with lower directional exposure.

- Source: https://arxiv.org/abs/2609.33767
- Identifier: arXiv:2609.33767
- Released: 2026-09-29
- First featured: Quant Letter No. 133 (2026-10-02): https://www.ml-quant.com/issues/2026-10-02/
- Citations (Semantic Scholar): not tracked
- Published in: not yet
- Topic: Portfolio & Allocation
- Authors: Wee Ling Tan, Stephen Roberts, Stefan Zohren

## Abstract (arXiv, CC0)

We present an end-to-end deep learning framework for systematic options trading that directly embeds hedging behavior through explicit control of portfolio-level risk exposures. While neural networks trained to optimize risk-adjusted performance have been shown to outperform traditional rules-based strategies, such approaches remain agnostic to the sensitivities of the resulting portfolios with respect to specific underlying risk factors. We propose a general training objective that combines a performance-driven loss with a differentiable risk-sensitivity penalty, enforcing neutrality to selected risk dimensions. Unlike reinforcement learning methods that approximate optimal hedging policies via simulated market dynamics, our framework operates entirely on historical data and jointly optimizes risk-adjusted returns and targeted risk constraints in a single learning problem. We instantiate the framework on static delta-neutral straddle portfolios with the penalty directed at first-order directional exposure, and evaluate two penalty variants -- an exposure-normalized penalty and a Greek-ratio drift penalty. Empirical results on Nasdaq 100 equity options demonstrate that appropriately calibrated regularization simultaneously improves out-of-sample risk-adjusted performance relative to an unregularized baseline while reducing realized directional exposure.

## Related

- [Deep Learning for Price Trend Prediction in ETF Markets](https://www.ml-quant.com/papers/ssrn/4524134/): The channel and spatial attention convolutional neural network (CSACNN) uses deep learning to predict financial market trends, performing as well or better than models using only time series data.
- [DeePM: Regime-Robust Deep Learning for Systematic Macro Portfolio Management](https://www.ml-quant.com/papers/arxiv/2601.05975/): Deep Learning for Portfolio Management: DeePM uses deep learning to improve macro portfolio management, delivering better risk-adjusted returns than traditional methods across various economic conditions.
- [Causal Portfolio Optimization: Principles and Sensitivity-Based Solutions](https://www.ml-quant.com/papers/arxiv/2504.05743/): The article introduces a new risk management framework that uses Bayesian and neural networks for efficient portfolio optimization, based on Common Causal Manifolds.
- [The Role of Deep Learning in Financial Asset Management: A Systematic Review](https://www.ml-quant.com/papers/arxiv/2503.01591/): A review of deep learning in financial asset management identifies trends like explainable AI and deep reinforcement learning, suggesting deep learning can enhance portfolio performance and price forecasting.
- [Smart leverage? Rethinking the role of Leveraged Exchange Traded Funds in constructing portfolios to beat a benchmark](https://www.ml-quant.com/papers/arxiv/2412.05431/): The paper investigates the potential of Leveraged Exchange Traded Funds (LETFs) in long-term investment strategies, using a neural network approach to devise strategies that beat standard benchmarks.
- [DSPO: An End-to-End Framework for Direct Sorted Portfolio Construction](https://www.ml-quant.com/papers/arxiv/2405.15833/): The paper showcases Direct Sorted Portfolio Optimization (DSPO), a framework using neural networks to process stock data and construct sorted portfolios, proven effective on various markets.
