---
title: Active Flow Control for Confined Square Cylinder Wake
url: https://www.ml-quant.com/papers/ssrn/5059657/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: SSRN 5059657
source_url: https://papers.ssrn.com/sol3/papers.cfm?abstract_id=5059657
featured: 2024-12-18
citations: unknown
topic: ML & AI Methods
---


# Active Flow Control for Confined Square Cylinder Wake

The article introduces a deep learning surrogate model-based reinforcement learning approach for active control of two-dimensional wake flow, which reduces computational costs while maintaining reliability.

- Source: https://papers.ssrn.com/sol3/papers.cfm?abstract_id=5059657
- Identifier: SSRN 5059657
- Released: 2024-12-16
- First featured: Quant Letter No. 79 (2024-12-18): https://www.ml-quant.com/issues/2024-12-18/
- Citations (Semantic Scholar): not tracked
- Published in: not yet
- Topic: ML & AI Methods

## Related

- [Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality](https://www.ml-quant.com/papers/arxiv/2405.21060/): The research identifies a link between state-space models and Transformers in deep learning, leading to the creation of a faster language modeling architecture, Mamba-2.
- [Mastering Diverse Domains through World Models](https://www.ml-quant.com/papers/arxiv/2301.04104/): Algorithm Mastery: DreamerV3, a universal algorithm, excels in over 150 varied tasks, including diamond collection in Minecraft without human input, expanding the scope of reinforcement learning.
- [SimPO: Simple Preference Optimization with a Reference-Free Reward](https://www.ml-quant.com/papers/arxiv/2405.14734/): Simple Preference Optimization: SimPO improves reinforcement learning from human feedback by using the average log probability of a sequence as the implicit reward, enhancing training stability and computational efficiency.
- [SOAP: Improving and Stabilizing Shampoo using Adam](https://www.ml-quant.com/papers/arxiv/2409.11321/): A new algorithm, SOAP, enhances the computational efficiency of the Shampoo preconditioning method in deep learning tasks, reducing iterations and time, with an online implementation available.
- [DPO Meets PPO: Reinforced Token Optimization for RLHF](https://www.ml-quant.com/papers/arxiv/2404.18922/): A new framework is introduced that models Reinforcement Learning from Human Feedback as a Markov decision process, using an algorithm that learns from preference data.
- [Settling the Sample Complexity of Model-Based Offline Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2204.05275/): A paper reveals that a model-based approach can achieve optimal sample complexity without burn-in cost in offline reinforcement learning for tabular Markov decision processes, providing an efficient solution for sample-starved applications.
