---
title: Packets, Transactions and Queues: Design Principles for HFT Systems from a Measurement Study of CME Market Data
url: https://www.ml-quant.com/papers/arxiv/2609.32848/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-10-02
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2609.32848
source_url: https://arxiv.org/abs/2609.32848
featured: 2026-10-02
citations: unknown
topic: Trading, Microstructure & Execution
---


# Packets, Transactions and Queues: Design Principles for HFT Systems from a Measurement Study of CME Market Data

Measurement of over a year of CME market data reveals transactions cluster within microseconds, yielding design principles: single-thread receivers never queue, two-thread splits reduce tail latency.

- Source: https://arxiv.org/abs/2609.32848
- Identifier: arXiv:2609.32848
- Released: 2026-09-29
- First featured: Quant Letter No. 133 (2026-10-02): https://www.ml-quant.com/issues/2026-10-02/
- Citations (Semantic Scholar): not tracked
- Published in: not yet
- Topic: Trading, Microstructure & Execution
- Authors: Vincent Maciejewski

## Abstract (arXiv, CC0)

HFT systems are conventionally built as a single-threaded event loop, on the rule that every thread hop adds latency. We test that rule against a measurement study of more than a year of CME market data for the NQ front-month contract, following every packet and matching-engine transaction through the feed's two exchange timestamps, and checking the results against a live production receiver. Packets arrive in near-critical self-exciting clusters that belong to the matching engine's transactions, not to how the exchange packs them. The engine often processes consecutive transactions within a fraction of a microsecond, while the market-data publisher sends at most one packet per publisher period of about 7.5 microseconds, so a burst reaches the receiver as a train of packets one period apart. This yields design principles for HFT systems. First, a receiver that handles each packet within one publisher period never queues on arrivals, however bursty the market; there one thread is best. Second, above that period a queueing tail appears, driven by the timing of transactions, not by packet rate or size, and two threads can be better than one: splitting the servicing chain into two stages on separate threads removes most of the tail at the cost of one hop on the median. Third, only the slowest stage matters, so a split pays only if it shortens it. Fourth, just under the period, where the production receiver runs, the remaining tail comes from multi-message packets and variable service times, and the levers are cost per message and spread of service, not thread count. An analytic framework, a burst-limit throughput identity and an exact reduction of the tandem to a single bottleneck server, supports these results.

## Related

- [FAST: Efficient Action Tokenization for Vision-Language-Action Models](https://www.ml-quant.com/papers/arxiv/2501.09747/): A new tokenization scheme, Frequency-space Action Sequence Tokenization (FAST), has been proposed for robot actions, facilitating the training of vision-language action policies for complex and high-frequency tasks.
- [Deep Reinforcement Learning for Active High Frequency Trading](https://www.ml-quant.com/papers/arxiv/2101.07107/): A new Deep Reinforcement Learning framework has been developed for high frequency stock trading, showing potential for profitable long-term strategies.
- [TradeFM: A Generative Foundation Model for Trade-flow and Market Microstructure](https://www.ml-quant.com/papers/arxiv/2602.23784/): A Model for Trade-Flow in Market Microstructure: TradeFM is a new AI model that improves the analysis of market structures by studying billions of trade events in stocks, leading to better simulations of financial returns.
- [Romania's Roadmap to a Greener Financial System: An analysis of Environmental, Social and Governance Reporting on the Bucharest Exchange Trading Index](https://www.ml-quant.com/papers/ssrn/4440516/): Romania struggles to attract sustainable investments because its major companies have low transparency and high greenhouse gas emissions.
- [Robust insurance pricing and liquidity management](https://www.ml-quant.com/papers/arxiv/2510.15709/): Accounting for model uncertainty makes insurers set higher, more conservative prices and liquidity buffers, widens capacity ranges, and produces much longer underwriting cycles with more time in low‑capacity states.
- [A Microstructure Analysis of Coupling in CFMMs](https://www.ml-quant.com/papers/arxiv/2510.06095/): The article investigates the impact of smart contract protocols on market dynamics, focusing on their influence on price drift, trade size, and market depth in coupled markets.
