---
title: Forecast Accuracy Is Not Trading Profit: Evolving Small Recurrent Networks for Stock Return Prediction
url: https://www.ml-quant.com/papers/arxiv/2610.07825/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-10-09
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2610.07825
source_url: https://arxiv.org/abs/2610.07825
featured: 2026-10-09
citations: unknown
topic: Portfolio & Allocation
---


# Forecast Accuracy Is Not Trading Profit: Evolving Small Recurrent Networks for Stock Return Prediction

Tiny neuroevolved recurrent networks rank first on both forecast accuracy and daily long-short returns across four portfolios, outperforming larger transformers and proving that forecast accuracy does not guarantee trading profit.

- Source: https://arxiv.org/abs/2610.07825
- Identifier: arXiv:2610.07825
- Released: 2026-10-07
- First featured: Quant Letter No. 134 (2026-10-09): https://www.ml-quant.com/issues/2026-10-09/
- Citations (Semantic Scholar): not tracked
- Published in: not yet
- Topic: Portfolio & Allocation
- Authors: Jonathan Chang, Zimeng Lyu

## Abstract (arXiv, CC0)

Time series forecasting models are typically compared on pointwise error, which scores a prediction in isolation from the decision it is produced for, and a lower forecast error does not imply a better decision downstream. A parallel debate asks whether modern transformer architectures forecast better than recurrent and other lightweight models. We compare linear, fixed recurrent, transformer, and mixing based architectures against recurrent networks evolved by neuroevolutionary architecture search, evaluating each on forecast accuracy and on the net return of a daily long/short strategy. All models are fit on a pooled panel, one network trained across the whole universe. Across four mid-cap portfolios and three trading years, the evolved networks rank first on both forecast accuracy and net trading performance, while the second most accurate model loses money once positions are formed and costs are charged. The advantage tracks a horizon match, since rank IC for the evolved networks rises from a one-day to a ten-day scoring horizon while every model above 300 parameters declines. They are also the cheapest end to end: a CPU-only search of 16 minutes yields 66-weight networks that predict in 10.8~$μ$s on a Raspberry Pi Zero, against transformer baselines of up to 817,153 parameters that require GPU training.

## Related

- [The Critical Line Algorithm and the Constrained LASSO: One Curve, Two Literatures](https://www.ml-quant.com/papers/arxiv/2609.25704/): Shows that mean-variance portfolio selection and the constrained LASSO trace identical piecewise-linear solution paths, mapping their parametrizations exactly.
- [Optimal Investment and Consumption in Financial Markets with Integrated Variance Clocks](https://www.ml-quant.com/papers/arxiv/2609.26349/): Characterizes optimal consumption and investment strategies in markets with stochastic volatility clocks using infinite-horizon backward SDEs, extending to rough and hyper-rough regimes.
- [DeePM: Regime-Robust Deep Learning for Systematic Macro Portfolio Management](https://www.ml-quant.com/papers/arxiv/2601.05975/): Deep Learning for Portfolio Management: DeePM uses deep learning to improve macro portfolio management, delivering better risk-adjusted returns than traditional methods across various economic conditions.
- [Regulating Cash Holdings: Assessing Lost Returns in Mutual Funds](https://www.ml-quant.com/papers/ssrn/4478272/): Israeli mutual funds hold excessive cash, indicating a need for better liquidity management to reduce redemption risks.
- [Effective and Scalable Programs to Facilitate Labor Market Transitions for Women in Technology](https://www.ml-quant.com/papers/arxiv/2211.09968/): In Poland, cheap online portfolio challenges and one‑on‑one mentoring sharply increased women’s tech employment, and data-driven targeting improved admissions.
- [A mathematical study of the excess growth rate](https://www.ml-quant.com/papers/arxiv/2510.25740/): - Excess Growth - Excess Rate - Growth Excess - Surplus Growth - Overgrowth - Growth Surplus Recommended: Excess Growth (keeps meaning but is more concise).: The paper proves that a central portfolio metric—the excess growth rate—can be exactly described using basic information‑theory ideas and a few natural axioms. In short, it shows that the extra growth a portfolio achieves is essentially an information quantity, so portfolio performance can be understood like information gain.
