---
title: ACEGEN: Reinforcement Learning of Generative Chemical Agents for Drug Discovery
url: https://www.ml-quant.com/papers/arxiv/2405.04657/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2405.04657
source_url: https://arxiv.org/abs/2405.04657
featured: 2024-07-24
citations: 41
topic: ML & AI Methods
---


# ACEGEN: Reinforcement Learning of Generative Chemical Agents for Drug Discovery

RL for Drug Design: ACEGEN, a toolkit for drug design using reinforcement learning, is presented and validated, demonstrating equal or better performance than other generative models.

- Source: https://arxiv.org/abs/2405.04657
- Identifier: arXiv:2405.04657
- Released: 2024-05-07
- First featured: Quant Letter No. 58 (2024-07-24): https://www.ml-quant.com/issues/2024-07-24/
- Citations (Semantic Scholar): 41
- Published in: Journal of Chemical Information and Modeling
- Topic: ML & AI Methods

## Related

- [Mastering Diverse Domains through World Models](https://www.ml-quant.com/papers/arxiv/2301.04104/): Algorithm Mastery: DreamerV3, a universal algorithm, excels in over 150 varied tasks, including diamond collection in Minecraft without human input, expanding the scope of reinforcement learning.
- [SimPO: Simple Preference Optimization with a Reference-Free Reward](https://www.ml-quant.com/papers/arxiv/2405.14734/): Simple Preference Optimization: SimPO improves reinforcement learning from human feedback by using the average log probability of a sequence as the implicit reward, enhancing training stability and computational efficiency.
- [CAT3D: Create Anything in 3D with Multi-View Diffusion Models](https://www.ml-quant.com/papers/arxiv/2405.10314/): Multi-View Diffusion Models: CAT3D is a novel technique for generating 3D scenes from any number of images, surpassing existing methods in speed and efficiency.
- [Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models](https://www.ml-quant.com/papers/arxiv/2501.01423/): The paper proposes a new model, VA-VAE, that aligns the latent space with pre-trained vision foundation models, enabling faster convergence of Diffusion Transformers in high-dimensional latent spaces and achieving top performance on ImageNet 256x256 generation.
- [Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps](https://www.ml-quant.com/papers/arxiv/2501.09732/): The research shows that increasing computation during inference-time can enhance the quality of samples produced by diffusion models, especially in image generation.
- [ShieldGemma: Generative AI Content Moderation Based on Gemma](https://www.ml-quant.com/papers/arxiv/2407.21772/): ShieldGemma is a safety content moderation model that excels in predicting safety risks such as explicit content and hate speech, surpassing models like LlamaGuard and WildCard.
