---
title: JAX-LOB: A GPU-Accelerated limit order book simulator to unlock large scale reinforcement learning for trading
url: https://www.ml-quant.com/papers/arxiv/2308.13289/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2308.13289
source_url: https://arxiv.org/abs/2308.13289
featured: 2023-08-30
citations: 30
topic: Trading, Microstructure & Execution
---


# JAX-LOB: A GPU-Accelerated limit order book simulator to unlock large scale reinforcement learning for trading

JAX-LOB: The paper introduces JAX-LOB, the first GPU-powered limit order book simulator capable of processing multiple books simultaneously, designed for efficient large-scale simulations of LOB dynamics for research, calibration, and reinforcement learning training.

- Source: https://arxiv.org/abs/2308.13289
- Identifier: arXiv:2308.13289
- Released: 2023-08-25
- First featured: Quant Letter No. 14 (2023-08-30): https://www.ml-quant.com/issues/2023-08-30/
- Citations (Semantic Scholar): 30
- Published in: Proceedings of the Fourth ACM International Conference on AI in Finance
- Topic: Trading, Microstructure & Execution

## Related

- [Reinforcement learning for trade execution with market and limit orders](https://www.ml-quant.com/papers/arxiv/2507.06345/): The paper presents a reinforcement learning framework for optimal trade execution, using multivariate logistic-normal distributions, which outperforms traditional benchmark strategies.
- [Consistent time travel for realistic interactions with historical data: reinforcement learning for market making](https://www.ml-quant.com/papers/arxiv/2408.02322/): The article discusses the use of consistent data time travel in offline reinforcement learning for market making in limit order books.
- [Robust Market Making with Hawkes Order Flow and Price Impact via Adversarial Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2609.22785/): The research extends adversarial reinforcement learning for market making to handle self-exciting order arrivals and price impact, using an LSTM module to improve robustness in complex microstructure environments.
- [Deep Reinforcement Learning for Active High Frequency Trading](https://www.ml-quant.com/papers/arxiv/2101.07107/): A new Deep Reinforcement Learning framework has been developed for high frequency stock trading, showing potential for profitable long-term strategies.
- [Domain-adapted Learning and Imitation: DRL for Power Arbitrage](https://www.ml-quant.com/papers/arxiv/2301.08360/): Leveraging Expertise: A dual-agent reinforcement learning approach can optimize European power arbitrage trading, improving training convergence and performance, and tripling profit and loss.
- [Gray-box Adversarial Attack of Deep Reinforcement Learning-based Trading Agents*](https://www.ml-quant.com/papers/arxiv/2309.14615/): A study has shown that a gray-box method can significantly reduce the profits of a Deep Reinforcement Learning-based trading agent, highlighting the need for stronger automated trading systems.
