---
title: Reinforcement Learning for Optimal Execution When Liquidity Is Time-Varying
url: https://www.ml-quant.com/papers/arxiv/2402.12049/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2402.12049
source_url: https://arxiv.org/abs/2402.12049
featured: 2024-02-21
citations: 11
topic: Trading, Microstructure & Execution
---


# Reinforcement Learning for Optimal Execution When Liquidity Is Time-Varying

Research shows Double Deep Q-learning, a Reinforcement Learning technique, can effectively learn optimal trading strategies in fluctuating liquidity conditions.

- Source: https://arxiv.org/abs/2402.12049
- Identifier: arXiv:2402.12049
- Released: 2024-02-19
- First featured: Quant Letter No. 38 (2024-02-21): https://www.ml-quant.com/issues/2024-02-21/
- Citations (Semantic Scholar): 11
- Published in: Applied Mathematical Finance
- Topic: Trading, Microstructure & Execution

## Related

- [The QLBS Model Within the Presence of Feedback Loops Through the Impacts of a Large Trader](https://www.ml-quant.com/papers/arxiv/2311.06790/): The QLBS model is expanded to include a large trader's impact on exchange rates and contingent claim prices, using reinforcement learning to find an optimal hedging strategy, reducing transaction costs and aligning with the trader's fair price.
- [Deviations from the Nash equilibrium in a two-player optimal execution game with reinforcement learning](https://www.ml-quant.com/papers/arxiv/2408.11773/): Autonomous trading bots using advanced algorithms can disrupt markets by deviating from traditional predictions, often favoring optimal solutions over equilibrium.
- [Reinforcement Learning for Optimal Execution](https://www.ml-quant.com/papers/ssrn/4720833/): A new actor-critic reinforcement learning algorithm is introduced for optimal execution problem, featuring a recalibration step for convergence and showing linear convergence under appropriate conditions.
- [Robust Market Making with Hawkes Order Flow and Price Impact via Adversarial Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2609.22785/): The research extends adversarial reinforcement learning for market making to handle self-exciting order arrivals and price impact, using an LSTM module to improve robustness in complex microstructure environments.
- [Deep Reinforcement Learning for Active High Frequency Trading](https://www.ml-quant.com/papers/arxiv/2101.07107/): A new Deep Reinforcement Learning framework has been developed for high frequency stock trading, showing potential for profitable long-term strategies.
- [Limit Order Book Simulations: A Review](https://www.ml-quant.com/papers/ssrn/4745587/): The piece reviews models of Limit Order Books simulations, emphasizing the role of AI in improving these models and the significance of price impacts in algorithmic trading.
