---
title: MemTrial: Learning When to Trust Memory in LLM Portfolio Agents
url: https://www.ml-quant.com/papers/arxiv/2610.11732/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-10-09
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2610.11732
source_url: https://arxiv.org/abs/2610.11732
featured: 2026-10-09
citations: unknown
topic: Portfolio & Allocation
---


# MemTrial: Learning When to Trust Memory in LLM Portfolio Agents

MemTrial uses factorial design to isolate what each memory contributes to portfolio decisions by crediting experiences with their marginal effect rather than shared market moves.

- Source: https://arxiv.org/abs/2610.11732
- Identifier: arXiv:2610.11732
- Released: 2026-10-09
- First featured: Quant Letter No. 134 (2026-10-09): https://www.ml-quant.com/issues/2026-10-09/
- Citations (Semantic Scholar): not tracked
- Published in: not yet
- Topic: Portfolio & Allocation
- Authors: Guanghao Wu, Zhuo Cai, Shoujin Wang

## Abstract (arXiv, CC0)

Large language model (LLM) agents for portfolio management learn from experience: they credit each experience in their memory with the outcome of the decisions that used it. In financial markets, however, this outcome mostly reflects the market move shared by all decisions on that date, so the credit tracks the market rather than the experience, and these agents often do worse than simply holding the equal-weight (1/$N$) portfolio. We ask how an agent can credit an experience with what it changes, and answer it by putting memory on trial: drafts of the same decision with and without an experience face the same market, so the outcome they share cancels in their difference. Our agent, MemTrial, drafts each decision with eight combinations of its retrieved experiences, chosen by a fractional factorial design, and credits each experience with its Banzhaf value, the average of these differences. As each date occurs once and each draft is a noisy LLM sample, these credits are noisy and may not hold on new dates. MemTrial therefore pools them across dates and similar experiences with a hierarchical Bayesian model, acts on them only after they have predicted unseen dates, and otherwise stays anchored at a conservative reference such as 1/$N$. On four benchmarks, MemTrial not only benefits from experiences that matter (the best of 15 methods on a semi-synthetic benchmark with known experience quality) but also limits its losses when its values do not hold (at most 2.2\% below 1/$N$ on PortBench and InvestorBench, against 15--38\% for the best experience-learning agent). Averaged over five settings, it improves the utility of the best experience-learning agent by 21.2\%, and with eight LLMs it beats every LLM-based baseline on InvestorBench.

## Related

- [Generative AI in Asset Management](https://www.ml-quant.com/papers/ssrn/4786575/): Hedge funds using generative AI tools like ChatGPT since 2022 have seen increased returns compared to those not using AI.
- [The Critical Line Algorithm and the Constrained LASSO: One Curve, Two Literatures](https://www.ml-quant.com/papers/arxiv/2609.25704/): Shows that mean-variance portfolio selection and the constrained LASSO trace identical piecewise-linear solution paths, mapping their parametrizations exactly.
- [Optimal Investment and Consumption in Financial Markets with Integrated Variance Clocks](https://www.ml-quant.com/papers/arxiv/2609.26349/): Characterizes optimal consumption and investment strategies in markets with stochastic volatility clocks using infinite-horizon backward SDEs, extending to rough and hyper-rough regimes.
- [DeePM: Regime-Robust Deep Learning for Systematic Macro Portfolio Management](https://www.ml-quant.com/papers/arxiv/2601.05975/): Deep Learning for Portfolio Management: DeePM uses deep learning to improve macro portfolio management, delivering better risk-adjusted returns than traditional methods across various economic conditions.
- [Regulating Cash Holdings: Assessing Lost Returns in Mutual Funds](https://www.ml-quant.com/papers/ssrn/4478272/): Israeli mutual funds hold excessive cash, indicating a need for better liquidity management to reduce redemption risks.
- [Effective and Scalable Programs to Facilitate Labor Market Transitions for Women in Technology](https://www.ml-quant.com/papers/arxiv/2211.09968/): In Poland, cheap online portfolio challenges and one‑on‑one mentoring sharply increased women’s tech employment, and data-driven targeting improved admissions.
