ML-QuantSubscribe

Machine learningLLMs & Text

LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models

Enhancing Text-to-Image Models: The study suggests a two-stage process using a pretrained language model to improve image generation accuracy in diffusion models, enabling multi-round scene specification in various languages.

Featured in No. 21 on 16 Oct 2023 · · 276 citations today · published in Trans. Mach. Learn. Res.

Released
23 May 2023
First featured
No. 21 · 16 Oct 2023
Citations (Semantic Scholar)
276
Influential citations
41
Published in
Trans. Mach. Learn. Res.
Shares when featured
179
Identifier
arXiv:2305.13655

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page