---
title: AlphaPADI: Formulaic Alpha Discovery via Pool-Aware Hierarchical Discrete Diffusion
url: https://www.ml-quant.com/papers/arxiv/2610.04959/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-10-09
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2610.04959
source_url: https://arxiv.org/abs/2610.04959
featured: 2026-10-09
citations: unknown
topic: Asset Pricing & Factors
---


# AlphaPADI: Formulaic Alpha Discovery via Pool-Aware Hierarchical Discrete Diffusion

Introduces a hierarchical discrete diffusion framework that generates pools of formulaic alphas by reconstructing complete candidate pools under current context and maximizing joint predictive performance and inner diversity.

- Source: https://arxiv.org/abs/2610.04959
- Identifier: arXiv:2610.04959
- Released: 2026-10-06
- First featured: Quant Letter No. 134 (2026-10-09): https://www.ml-quant.com/issues/2026-10-09/
- Citations (Semantic Scholar): not tracked
- Published in: not yet
- Topic: Asset Pricing & Factors
- Authors: Yanzheng Jin, Pengyang Shao, Yunshan Ma, Haowen Pan, Naixin Zhai, Chen-Hui Song, Fei Shen, Kenji Kawaguchi

## Abstract (arXiv, CC0)

Formulaic alpha discovery seeks symbolic expressions that predict cross-sectional asset returns. In deployment, multiple formulas are combined into an alpha pool, where each formula is valued through the complementary information it contributes to joint predictive performance. While Reinforcement Learning and Generative Flow Networks have emerged as promising paradigms for generating formulaic alphas, existing frameworks face three related challenges. First, generating formulas individually leaves pool context and inter-formula complementarity outside the generative state. Second, formula-wise generation lacks a unified mechanism for preserving and revising structures at different levels. Third, pool-level rewards jointly reflect predictive performance and redundancy but cannot be differentiated directly through symbolic evaluation to train the generator. To overcome these challenges, we introduce AlphaPADI (Formulaic Alpha Discovery via Pool-Aware Hierarchical Discrete Diffusion), a novel framework built around three components: (1) grammar-constrained buffer initialization that constructs syntactically valid pool candidates, (2) pool-aware hierarchical diffusion that reconstructs complete pools at multiple structural scales under the current pool context, and (3) reward-guided pool refinement that evaluates joint predictive performance and inner diversity, updates the elite buffer, and trains the reverse model through reconstruction and preference learning. Empirical results on the Chinese and U.S. stock markets demonstrate that AlphaPADI outperforms the evaluated baselines in both predictive and portfolio performance, thereby validating pool-aware generation as an effective framework for automated alpha discovery.

## Related

- [FactorBench: A Portfolio-Aware Benchmark for Automated Factor Mining](https://www.ml-quant.com/papers/arxiv/2610.06947/): A portfolio-aware benchmark comparing five thousand factors from nine automated mining methods across five equity markets finds that no discovery paradigm consistently dominates in signal quality or portfolio performance.
- [US Tariff Policy: Chaos Order](https://www.ml-quant.com/papers/ssrn/5232092/): Chaos Order: The unpredictability of U.S. tariff policy under President Trump's second administration aligns with momentum-style investing and reinforcement learning, suggesting a dynamic and reactive policy process.
- [Interpretable Machine Learning for Asset Pricing](https://www.ml-quant.com/papers/ssrn/4473746/): The paper utilizes deep neural networks to more accurately estimate equity risk premia over time, enhancing the interpretability of machine learning in economics.
- [The Cross-Section of Factor Returns](https://www.ml-quant.com/papers/ssrn/4441376/): Most of the 150 equity factors examined show positive returns but fail to deliver excess returns after accounting for risk, especially in downturns.
- [Are Penalty Shootouts Better Than a Coin Toss? Evidence From International Club Football in Europe](https://www.ml-quant.com/papers/arxiv/2510.17641/): Using UEFA penalty shootout data (2000–2025) we find outcomes are essentially random—no measurable advantage from kicking order, venue, momentum, or team strength.
- [Optimal Investment and Consumption in a Stochastic Factor Model](https://www.ml-quant.com/papers/arxiv/2509.09452/): The article discusses optimal investment and consumption in an incomplete stochastic factor model, offering a comprehensive characterization of the problem's well-posedness and an efficient numerical algorithm for computing the value function.
