ML-QuantSubscribe

arXivTrading, Microstructure & Execution

ReLOBGen: Replayable Limit Order Book Message Generation

A method generates realistic limit order book messages guaranteed to be consistent with current market state, achieving 100% replayability and 2.7–3.6× speedup over existing approaches.

Featured in No. 133 on 2 Oct 2026 · 2 days after release

Mean resting-order coverage increases with history length for GOOG and INTC stocks
Figure 10: Mean resting-order coverage by history length on GOOG and INTC. Coverage is evaluated during regular trading hours on the test split.
Released
30 Sep 2026
First featured
No. 133 · 2 Oct 2026
Published in
Not yet, as far as Semantic Scholar knows
Fanfare
3 of 5
Identifier
arXiv:2609.35867
Authors
Junoh Kang et al.

Abstract

From arXiv (CC0).

We propose ReLOBGen, a method for generating limit order book (LOB) messages that are replayable by construction. Replayability is required for closed-loop market simulation, yet existing LOB message generators may produce non-replayable raw messages, i.e., messages inconsistent with the current market state. These generators therefore rely on post-hoc correction or rejection followed by resampling, which may alter the replayed message distribution or increase inference cost. ReLOBGen instead ensures replayability during generation: it selects the referenced order from the resting orders in the current LOB and then generates the remaining message fields to be consistent with that order and the market state. For realistic reference selection, ReLOBGen samples from a learned distribution over eligible resting orders, efficiently computed from cached order representations and a context-dependent query. It then enforces the consistency of the remaining fields by masking out invalid tokens. Together, these components enable efficient generation of realistic messages without post-hoc correction or resampling. In 500-message rollouts, ReLOBGen achieves 100% replayability, improves market realism, particularly for top-of-book statistics and the relative prices of LOB messages, and provides a $2.7\text{-}3.6\times$ speedup per replayed message over the LOBS5 baseline.

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page