---
title: Reinforcement learning for trade execution with market and limit orders
url: https://www.ml-quant.com/papers/arxiv/2507.06345/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2507.06345
source_url: http://arxiv.org/abs/2507.06345v1
featured: 2025-07-10
citations: 4
topic: Trading, Microstructure & Execution
---


# Reinforcement learning for trade execution with market and limit orders

The paper presents a reinforcement learning framework for optimal trade execution, using multivariate logistic-normal distributions, which outperforms traditional benchmark strategies.

- Source: http://arxiv.org/abs/2507.06345v1
- Identifier: arXiv:2507.06345
- Released: 2025-07-08
- First featured: Quant Letter No. 105 (2025-07-10): https://www.ml-quant.com/issues/2025-07-10/
- Citations (Semantic Scholar): 4
- Published in: Quantitative Finance
- Topic: Trading, Microstructure & Execution

## Related

- [JAX-LOB: A GPU-Accelerated limit order book simulator to unlock large scale reinforcement learning for trading](https://www.ml-quant.com/papers/arxiv/2308.13289/): JAX-LOB: The paper introduces JAX-LOB, the first GPU-powered limit order book simulator capable of processing multiple books simultaneously, designed for efficient large-scale simulations of LOB dynamics for research, calibration, and reinforcement learning training.
- [Consistent time travel for realistic interactions with historical data: reinforcement learning for market making](https://www.ml-quant.com/papers/arxiv/2408.02322/): The article discusses the use of consistent data time travel in offline reinforcement learning for market making in limit order books.
- [Robust Market Making with Hawkes Order Flow and Price Impact via Adversarial Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2609.22785/): The research extends adversarial reinforcement learning for market making to handle self-exciting order arrivals and price impact, using an LSTM module to improve robustness in complex microstructure environments.
- [Deep Reinforcement Learning for Active High Frequency Trading](https://www.ml-quant.com/papers/arxiv/2101.07107/): A new Deep Reinforcement Learning framework has been developed for high frequency stock trading, showing potential for profitable long-term strategies.
- [Bimodal dynamics of the artificial limit order book stock exchange with autonomous traders](https://www.ml-quant.com/papers/arxiv/2508.17837/): The paper uncovers the inherent bistability and complex dynamics of an artificial stock market exchange, which emerge from micro-level trading rules.
- [Detecting Multilevel Manipulation from Limit Order Book via Cascaded Contrastive Representation Learning](https://www.ml-quant.com/papers/arxiv/2508.17086/): The study suggests a learning framework to enhance the detection of trade-based manipulation in financial markets, with Transformer-based architectures proving most successful.
